GraSAME: Injecting Token-Level Structural Information to Pretrained Language Models via Graph-guided Self-Attention Mechanism
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Shuzhou, Färber, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification
by: Im, Kyuri, et al.
Published: (2026)
by: Im, Kyuri, et al.
Published: (2026)
GNNavi: Navigating the Information Flow in Large Language Models by Graph Neural Network
by: Yuan, Shuzhou, et al.
Published: (2024)
by: Yuan, Shuzhou, et al.
Published: (2024)
PoeTone: A Framework for Constrained Generation of Structured Chinese Songci with LLMs
by: Qu, Zhan, et al.
Published: (2025)
by: Qu, Zhan, et al.
Published: (2025)
From Monolingual to Bilingual: Investigating Language Conditioning in Large Language Models for Psycholinguistic Tasks
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
Why Lift so Heavy? Slimming Large Language Models by Cutting Off the Layers
by: Yuan, Shuzhou, et al.
Published: (2024)
by: Yuan, Shuzhou, et al.
Published: (2024)
ToPro: Token-Level Prompt Decomposition for Cross-Lingual Sequence Labeling Tasks
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
CoDAE: Adapting Large Language Models for Education via Chain-of-Thought Data Augmentation
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
Decomposed Prompting: Probing Multilingual Linguistic Structure Knowledge in Large Language Models
by: Nie, Ercong, et al.
Published: (2024)
by: Nie, Ercong, et al.
Published: (2024)
Hateful Person or Hateful Model? Investigating the Role of Personas in Hate Speech Detection by Large Language Models
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
by: Brannon, William, et al.
Published: (2023)
by: Brannon, William, et al.
Published: (2023)
Injecting Structured Biomedical Knowledge into Language Models: Continual Pretraining vs. GraphRAG
by: Klila, Jaafer, et al.
Published: (2026)
by: Klila, Jaafer, et al.
Published: (2026)
Graph-Guided Textual Explanation Generation Framework
by: Yuan, Shuzhou, et al.
Published: (2024)
by: Yuan, Shuzhou, et al.
Published: (2024)
Can Hallucinations Help? Boosting LLMs for Drug Discovery
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
Knowledge Graph Structure as Prompt: Improving Small Language Models Capabilities for Knowledge-based Causal Discovery
by: Susanti, Yuni, et al.
Published: (2024)
by: Susanti, Yuni, et al.
Published: (2024)
Knowledge of Pretrained Language Models on Surface Information of Tokens
by: Hiraoka, Tatsuya, et al.
Published: (2024)
by: Hiraoka, Tatsuya, et al.
Published: (2024)
Beyond Over-Refusal: Scenario-Based Diagnostics and Post-Hoc Mitigation for Exaggerated Refusals in LLMs
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models
by: Yuan, Zike, et al.
Published: (2024)
by: Yuan, Zike, et al.
Published: (2024)
GraSP: Graph-Structured Skill Compositions for LLM Agents
by: Xia, Tianle, et al.
Published: (2026)
by: Xia, Tianle, et al.
Published: (2026)
Reflection Pretraining Enables Token-Level Self-Correction in Biological Sequence Models
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
GraLMatch: Matching Groups of Entities with Graphs and Language Models
by: Pardo, Fernando De Meer, et al.
Published: (2024)
by: Pardo, Fernando De Meer, et al.
Published: (2024)
Tracing Relational Knowledge Recall in Large Language Models
by: Popovič, Nicholas, et al.
Published: (2026)
by: Popovič, Nicholas, et al.
Published: (2026)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
by: Lu, Yifan, et al.
Published: (2025)
by: Lu, Yifan, et al.
Published: (2025)
LLM in the Loop: Creating the ParaDeHate Dataset for Hate Speech Detoxification
by: Yuan, Shuzhou, et al.
Published: (2025)
by: Yuan, Shuzhou, et al.
Published: (2025)
Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
Pretraining with Token-Level Adaptive Latent Chain-of-Thought
by: Zeng, Boyi, et al.
Published: (2026)
by: Zeng, Boyi, et al.
Published: (2026)
An Encoder-Integrated PhoBERT with Graph Attention for Vietnamese Token-Level Classification
by: Nguyen, Ba-Quang
Published: (2025)
by: Nguyen, Ba-Quang
Published: (2025)
Extractive Fact Decomposition for Interpretable Natural Language Inference in one Forward Pass
by: Popovič, Nicholas, et al.
Published: (2025)
by: Popovič, Nicholas, et al.
Published: (2025)
Paths to Causality: Finding Informative Subgraphs Within Knowledge Graphs for Knowledge-Based Causal Discovery
by: Susanti, Yuni, et al.
Published: (2025)
by: Susanti, Yuni, et al.
Published: (2025)
A Knowledge-Injected Curriculum Pretraining Framework for Question Answering
by: Lin, Xin, et al.
Published: (2024)
by: Lin, Xin, et al.
Published: (2024)
Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language Models
by: Schreieder, Tobias, et al.
Published: (2025)
by: Schreieder, Tobias, et al.
Published: (2025)
Mechanics of Next Token Prediction with Self-Attention
by: Li, Yingcong, et al.
Published: (2024)
by: Li, Yingcong, et al.
Published: (2024)
Injecting New Knowledge into Large Language Models via Supervised Fine-Tuning
by: Mecklenburg, Nick, et al.
Published: (2024)
by: Mecklenburg, Nick, et al.
Published: (2024)
The Hidden Bias: A Study on Explicit and Implicit Political Stereotypes in Large Language Models
by: Löhr, Konrad, et al.
Published: (2025)
by: Löhr, Konrad, et al.
Published: (2025)
Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
by: Soldaini, Luca, et al.
Published: (2024)
by: Soldaini, Luca, et al.
Published: (2024)
Adaptive BPE Tokenization for Enhanced Vocabulary Adaptation in Finetuning Pretrained Language Models
by: Balde, Gunjan, et al.
Published: (2024)
by: Balde, Gunjan, et al.
Published: (2024)
CAT: Causal Attention Tuning For Injecting Fine-grained Causal Knowledge into Large Language Models
by: Han, Kairong, et al.
Published: (2025)
by: Han, Kairong, et al.
Published: (2025)
Adapting Pretrained Language Models for Citation Classification via Self-Supervised Contrastive Learning
by: Li, Tong, et al.
Published: (2025)
by: Li, Tong, et al.
Published: (2025)
Vision Token Reduction via Attention-Driven Self-Compression for Efficient Multimodal Large Language Models
by: Deniz, Omer Faruk, et al.
Published: (2026)
by: Deniz, Omer Faruk, et al.
Published: (2026)
Rethinking Personalization in Large Language Models at the Token Level
by: Zhang, Chenheng, et al.
Published: (2026)
by: Zhang, Chenheng, et al.
Published: (2026)
Neural Attention Search Linear: Towards Adaptive Token-Level Hybrid Attention Models
by: Deng, Difan, et al.
Published: (2026)
by: Deng, Difan, et al.
Published: (2026)
Similar Items
-
Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification
by: Im, Kyuri, et al.
Published: (2026) -
GNNavi: Navigating the Information Flow in Large Language Models by Graph Neural Network
by: Yuan, Shuzhou, et al.
Published: (2024) -
PoeTone: A Framework for Constrained Generation of Structured Chinese Songci with LLMs
by: Qu, Zhan, et al.
Published: (2025) -
From Monolingual to Bilingual: Investigating Language Conditioning in Large Language Models for Psycholinguistic Tasks
by: Yuan, Shuzhou, et al.
Published: (2025) -
Why Lift so Heavy? Slimming Large Language Models by Cutting Off the Layers
by: Yuan, Shuzhou, et al.
Published: (2024)