Improving Contextual Faithfulness of Large Language Models via Retrieval Heads-Induced Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Lei, Feng, Xiaocheng, Ma, Weitao, Fan, Yuchun, Feng, Xiachong, Ye, Yangfan, Zhong, Weihong, Gu, Yuxuan, Wang, Baoxin, Wu, Dayong, Hu, Guoping, Qin, Bing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
by: Ma, Weitao, et al.
Published: (2024)
by: Ma, Weitao, et al.
Published: (2024)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
by: Li, Qiming, et al.
Published: (2025)
by: Li, Qiming, et al.
Published: (2025)
From Hypothesis to Publication: A Comprehensive Survey of AI-Driven Research Support Systems
by: Zhou, Zekun, et al.
Published: (2025)
by: Zhou, Zekun, et al.
Published: (2025)
PERSONA: Dynamic and Compositional Inference-Time Personality Control via Activation Vector Algebra
by: Feng, Xiachong, et al.
Published: (2026)
by: Feng, Xiachong, et al.
Published: (2026)
SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution
by: Feng, Xiachong, et al.
Published: (2026)
by: Feng, Xiachong, et al.
Published: (2026)
Learning Fine-Grained Grounded Citations for Attributed Large Language Models
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
Advancing Large Language Model Attribution through Self-Improving
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
by: Ye, Yangfan, et al.
Published: (2024)
by: Ye, Yangfan, et al.
Published: (2024)
Stratagem: Learning Transferable Reasoning via Trajectory-Modulated Game Self-Play
by: Feng, Xiachong, et al.
Published: (2026)
by: Feng, Xiachong, et al.
Published: (2026)
Adaptive Backtracking for Privacy Protection in Large Language Models
by: Yao, Zhihao, et al.
Published: (2025)
by: Yao, Zhihao, et al.
Published: (2025)
LangGPS: Language Separability Guided Data Pre-Selection for Joint Multilingual Instruction Tuning
by: Ye, Yangfan, et al.
Published: (2025)
by: Ye, Yangfan, et al.
Published: (2025)
ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models
by: Qin, Chonghan, et al.
Published: (2026)
by: Qin, Chonghan, et al.
Published: (2026)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
by: Zhong, Weihong, et al.
Published: (2024)
by: Zhong, Weihong, et al.
Published: (2024)
x1: Learning to Think Adaptively Across Languages and Cultures
by: Ye, Yangfan, et al.
Published: (2026)
by: Ye, Yangfan, et al.
Published: (2026)
CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs
by: Ye, Yangfan, et al.
Published: (2026)
by: Ye, Yangfan, et al.
Published: (2026)
Can Large Language Models Simulate Human Cognition Beyond Behavioral Imitation?
by: Gu, Yuxuan, et al.
Published: (2026)
by: Gu, Yuxuan, et al.
Published: (2026)
CC-Tuning: A Cross-Lingual Connection Mechanism for Improving Joint Multilingual Supervised Fine-Tuning
by: Ye, Yangfan, et al.
Published: (2025)
by: Ye, Yangfan, et al.
Published: (2025)
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding
by: Zhao, Liang, et al.
Published: (2023)
by: Zhao, Liang, et al.
Published: (2023)
Discrete Modeling via Boundary Conditional Diffusion Processes
by: Gu, Yuxuan, et al.
Published: (2024)
by: Gu, Yuxuan, et al.
Published: (2024)
Adapter-based Selective Knowledge Distillation for Federated Multi-domain Meeting Summarization
by: Feng, Xiachong, et al.
Published: (2023)
by: Feng, Xiachong, et al.
Published: (2023)
Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges
by: Ye, Yangfan, et al.
Published: (2024)
by: Ye, Yangfan, et al.
Published: (2024)
Length Controlled Generation for Black-box LLMs
by: Gu, Yuxuan, et al.
Published: (2024)
by: Gu, Yuxuan, et al.
Published: (2024)
Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
by: Li, Qiming, et al.
Published: (2025)
by: Li, Qiming, et al.
Published: (2025)
Extending Context Window of Large Language Models from a Distributional Perspective
by: Wu, Yingsheng, et al.
Published: (2024)
by: Wu, Yingsheng, et al.
Published: (2024)
Fine-Mem: Fine-Grained Feedback Alignment for Long-Horizon Memory Management
by: Ma, Weitao, et al.
Published: (2026)
by: Ma, Weitao, et al.
Published: (2026)
CE-GOCD: Central Entity-Guided Graph Optimization for Community Detection to Augment LLM Scientific Question Answering
by: Lan, Jiayin, et al.
Published: (2026)
by: Lan, Jiayin, et al.
Published: (2026)
Context-Aware Hierarchical Taxonomy Generation for Scientific Papers via LLM-Guided Multi-Aspect Clustering
by: Zhu, Kun, et al.
Published: (2025)
by: Zhu, Kun, et al.
Published: (2025)
One for All: Update Parameterized Knowledge Across Multiple Models
by: Ma, Weitao, et al.
Published: (2025)
by: Ma, Weitao, et al.
Published: (2025)
Improving Grammatical Error Correction via Contextual Data Augmentation
by: Wang, Yixuan, et al.
Published: (2024)
by: Wang, Yixuan, et al.
Published: (2024)
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
by: Huang, Lei, et al.
Published: (2023)
by: Huang, Lei, et al.
Published: (2023)
The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models
by: Qin, Chonghan, et al.
Published: (2026)
by: Qin, Chonghan, et al.
Published: (2026)
LM-Combiner: A Contextual Rewriting Model for Chinese Grammatical Error Correction
by: Wang, Yixuan, et al.
Published: (2024)
by: Wang, Yixuan, et al.
Published: (2024)
RE$^2$: Improving Chinese Grammatical Error Correction via Retrieving Appropriate Examples with Explanation
by: Wang, Baoxin, et al.
Published: (2025)
by: Wang, Baoxin, et al.
Published: (2025)
Cross-Lingual Text-Rich Visual Comprehension: An Information Theory Perspective
by: Yu, Xinmiao, et al.
Published: (2024)
by: Yu, Xinmiao, et al.
Published: (2024)
Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation
by: Yuan, Zekun, et al.
Published: (2026)
by: Yuan, Zekun, et al.
Published: (2026)
Reasoning Does Not Necessarily Improve Role-Playing Ability
by: Feng, Xiachong, et al.
Published: (2025)
by: Feng, Xiachong, et al.
Published: (2025)
An Information Bottleneck Perspective for Effective Noise Filtering on Retrieval-Augmented Generation
by: Zhu, Kun, et al.
Published: (2024)
by: Zhu, Kun, et al.
Published: (2024)
MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents
by: Chen, Ruihan, et al.
Published: (2025)
by: Chen, Ruihan, et al.
Published: (2025)
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
by: Li, Qiming, et al.
Published: (2026)
by: Li, Qiming, et al.
Published: (2026)
Trends in Integration of Knowledge and Large Language Models: A Survey and Taxonomy of Methods, Benchmarks, and Applications
by: Feng, Zhangyin, et al.
Published: (2023)
by: Feng, Zhangyin, et al.
Published: (2023)
Similar Items
-
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
by: Ma, Weitao, et al.
Published: (2024) -
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
by: Li, Qiming, et al.
Published: (2025) -
From Hypothesis to Publication: A Comprehensive Survey of AI-Driven Research Support Systems
by: Zhou, Zekun, et al.
Published: (2025) -
PERSONA: Dynamic and Compositional Inference-Time Personality Control via Activation Vector Algebra
by: Feng, Xiachong, et al.
Published: (2026) -
SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution
by: Feng, Xiachong, et al.
Published: (2026)