In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Shiqi, Xiong, Miao, Liu, Junteng, Wu, Zhengxuan, Xiao, Teng, Gao, Siyang, He, Junxian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hallucination Detection and Hallucination Mitigation: An Investigation
di: Luo, Junliang, et al.
Pubblicazione: (2024)
di: Luo, Junliang, et al.
Pubblicazione: (2024)
DAMR: Efficient and Adaptive Context-Aware Knowledge Graph Question Answering with LLM-Guided MCTS
di: Wang, Yingxu, et al.
Pubblicazione: (2025)
di: Wang, Yingxu, et al.
Pubblicazione: (2025)
ESG-Bench: Benchmarking Long-Context ESG Reports for Hallucination Mitigation
di: Sun, Siqi, et al.
Pubblicazione: (2026)
di: Sun, Siqi, et al.
Pubblicazione: (2026)
On the Universal Truthfulness Hyperplane Inside LLMs
di: Liu, Junteng, et al.
Pubblicazione: (2024)
di: Liu, Junteng, et al.
Pubblicazione: (2024)
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective
di: Sun, Zhongxiang, et al.
Pubblicazione: (2025)
di: Sun, Zhongxiang, et al.
Pubblicazione: (2025)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
di: Sun, Zhongxiang, et al.
Pubblicazione: (2026)
di: Sun, Zhongxiang, et al.
Pubblicazione: (2026)
Towards Long Context Hallucination Detection
di: Liu, Siyi, et al.
Pubblicazione: (2025)
di: Liu, Siyi, et al.
Pubblicazione: (2025)
Learn to Reason Efficiently with Adaptive Length-based Reward Shaping
di: Liu, Wei, et al.
Pubblicazione: (2025)
di: Liu, Wei, et al.
Pubblicazione: (2025)
SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
di: Liu, Junteng, et al.
Pubblicazione: (2025)
di: Liu, Junteng, et al.
Pubblicazione: (2025)
Look Within, Why LLMs Hallucinate: A Causal Perspective
di: Li, He, et al.
Pubblicazione: (2024)
di: Li, He, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Multimodal Spatial Relations through Constraint-Aware Prompting
di: Wu, Jiarui, et al.
Pubblicazione: (2025)
di: Wu, Jiarui, et al.
Pubblicazione: (2025)
What are Models Thinking about? Understanding Large Language Model Hallucinations "Psychology" through Model Inner State Analysis
di: Wang, Peiran, et al.
Pubblicazione: (2025)
di: Wang, Peiran, et al.
Pubblicazione: (2025)
In-Context Learning State Vector with Inner and Momentum Optimization
di: Li, Dongfang, et al.
Pubblicazione: (2024)
di: Li, Dongfang, et al.
Pubblicazione: (2024)
Large Language Model Bias Mitigation from the Perspective of Knowledge Editing
di: Chen, Ruizhe, et al.
Pubblicazione: (2024)
di: Chen, Ruizhe, et al.
Pubblicazione: (2024)
Copy-Paste to Mitigate Large Language Model Hallucinations
di: Long, Yongchao, et al.
Pubblicazione: (2025)
di: Long, Yongchao, et al.
Pubblicazione: (2025)
KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning
di: Gao, Cheng, et al.
Pubblicazione: (2026)
di: Gao, Cheng, et al.
Pubblicazione: (2026)
CIP: A Plug-and-Play Causal Prompting Framework for Mitigating Hallucinations under Long-Context Noise
di: Ma, Qingsen, et al.
Pubblicazione: (2025)
di: Ma, Qingsen, et al.
Pubblicazione: (2025)
LFTF: Locating First and Then Fine-Tuning for Mitigating Gender Bias in Large Language Models
di: Qin, Zhanyue, et al.
Pubblicazione: (2025)
di: Qin, Zhanyue, et al.
Pubblicazione: (2025)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
di: Jiang, Xinyan, et al.
Pubblicazione: (2025)
di: Jiang, Xinyan, et al.
Pubblicazione: (2025)
An Evolutionary Large Language Model for Hallucination Mitigation
di: Boulesnane, Abdennour, et al.
Pubblicazione: (2024)
di: Boulesnane, Abdennour, et al.
Pubblicazione: (2024)
Delta -- Contrastive Decoding Mitigates Text Hallucinations in Large Language Models
di: Huang, Cheng Peng, et al.
Pubblicazione: (2025)
di: Huang, Cheng Peng, et al.
Pubblicazione: (2025)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
di: Yao, Siyang, et al.
Pubblicazione: (2026)
di: Yao, Siyang, et al.
Pubblicazione: (2026)
CCL-XCoT: An Efficient Cross-Lingual Knowledge Transfer Method for Mitigating Hallucination Generation
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
I-CALM: Incentivizing Confidence-Aware Abstention for LLM Hallucination Mitigation
di: Zong, Haotian, et al.
Pubblicazione: (2026)
di: Zong, Haotian, et al.
Pubblicazione: (2026)
IRCAN: Mitigating Knowledge Conflicts in LLM Generation via Identifying and Reweighting Context-Aware Neurons
di: Shi, Dan, et al.
Pubblicazione: (2024)
di: Shi, Dan, et al.
Pubblicazione: (2024)
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
Hallucination, Monofacts, and Miscalibration: An Empirical Investigation
di: Miao, Miranda Muqing, et al.
Pubblicazione: (2025)
di: Miao, Miranda Muqing, et al.
Pubblicazione: (2025)
MALM: A Multi-Information Adapter for Large Language Models to Mitigate Hallucination
di: Jia, Ao, et al.
Pubblicazione: (2025)
di: Jia, Ao, et al.
Pubblicazione: (2025)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2024)
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2024)
Theoretical Foundations and Mitigation of Hallucination in Large Language Models
di: Gumaan, Esmail
Pubblicazione: (2025)
di: Gumaan, Esmail
Pubblicazione: (2025)
DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucinations
di: Gema, Aryo Pradipta, et al.
Pubblicazione: (2024)
di: Gema, Aryo Pradipta, et al.
Pubblicazione: (2024)
Enhancing Hallucination Detection via Future Context
di: Lee, Joosung, et al.
Pubblicazione: (2025)
di: Lee, Joosung, et al.
Pubblicazione: (2025)
Improving Context Fidelity via Native Retrieval-Augmented Reasoning
di: Wang, Suyuchen, et al.
Pubblicazione: (2025)
di: Wang, Suyuchen, et al.
Pubblicazione: (2025)
Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging
di: Chen, Shiqi, et al.
Pubblicazione: (2025)
di: Chen, Shiqi, et al.
Pubblicazione: (2025)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
di: Mündler, Niels, et al.
Pubblicazione: (2023)
di: Mündler, Niels, et al.
Pubblicazione: (2023)
HDLCoRe: A Training-Free Framework for Mitigating Hallucinations in LLM-Generated HDL
di: Ping, Heng, et al.
Pubblicazione: (2025)
di: Ping, Heng, et al.
Pubblicazione: (2025)
Mitigating Hallucination in Abstractive Summarization with Domain-Conditional Mutual Information
di: Chae, Kyubyung, et al.
Pubblicazione: (2024)
di: Chae, Kyubyung, et al.
Pubblicazione: (2024)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
di: Yu, Lei, et al.
Pubblicazione: (2024)
di: Yu, Lei, et al.
Pubblicazione: (2024)
Counterfactual Probing for Hallucination Detection and Mitigation in Large Language Models
di: Feng, Yijun
Pubblicazione: (2025)
di: Feng, Yijun
Pubblicazione: (2025)
ReFT: Representation Finetuning for Language Models
di: Wu, Zhengxuan, et al.
Pubblicazione: (2024)
di: Wu, Zhengxuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hallucination Detection and Hallucination Mitigation: An Investigation
di: Luo, Junliang, et al.
Pubblicazione: (2024) -
DAMR: Efficient and Adaptive Context-Aware Knowledge Graph Question Answering with LLM-Guided MCTS
di: Wang, Yingxu, et al.
Pubblicazione: (2025) -
ESG-Bench: Benchmarking Long-Context ESG Reports for Hallucination Mitigation
di: Sun, Siqi, et al.
Pubblicazione: (2026) -
On the Universal Truthfulness Hyperplane Inside LLMs
di: Liu, Junteng, et al.
Pubblicazione: (2024) -
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective
di: Sun, Zhongxiang, et al.
Pubblicazione: (2025)