Learning to Trust Your Feelings: Leveraging Self-awareness in LLMs for Hallucination Mitigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Yuxin, Song, Zhuoyang, Wang, Hao, Zhang, Jiaxing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
Lyrics: Boosting Fine-grained Language-Vision Alignment and Comprehension via Semantic-aware Visual Objects
von: Lu, Junyu, et al.
Veröffentlicht: (2023)
von: Lu, Junyu, et al.
Veröffentlicht: (2023)
Never Lost in the Middle: Mastering Long-Context Question Answering with Position-Agnostic Decompositional Training
von: He, Junqing, et al.
Veröffentlicht: (2023)
von: He, Junqing, et al.
Veröffentlicht: (2023)
Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization
von: Compagnoni, Alberto, et al.
Veröffentlicht: (2025)
von: Compagnoni, Alberto, et al.
Veröffentlicht: (2025)
Rowen: Adaptive Retrieval-Augmented Generation for Hallucination Mitigation in LLMs
von: Ding, Hanxing, et al.
Veröffentlicht: (2024)
von: Ding, Hanxing, et al.
Veröffentlicht: (2024)
Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning
von: Liao, Yiming, et al.
Veröffentlicht: (2026)
von: Liao, Yiming, et al.
Veröffentlicht: (2026)
Can We Trust LLMs? Mitigate Overconfidence Bias in LLMs through Knowledge Transfer
von: Yang, Haoyan, et al.
Veröffentlicht: (2024)
von: Yang, Haoyan, et al.
Veröffentlicht: (2024)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
LANCET: Neural Intervention via Structural Entropy for Mitigating Faithfulness Hallucinations in LLMs
von: Wang, Chenxu, et al.
Veröffentlicht: (2026)
von: Wang, Chenxu, et al.
Veröffentlicht: (2026)
A Concise Review of Hallucinations in LLMs and their Mitigation
von: Pulkundwar, Parth, et al.
Veröffentlicht: (2025)
von: Pulkundwar, Parth, et al.
Veröffentlicht: (2025)
Do LLMs "Feel"? Emotion Circuits Discovery and Control
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs
von: Gao, Yizhao, et al.
Veröffentlicht: (2024)
von: Gao, Yizhao, et al.
Veröffentlicht: (2024)
Mitigating Multimodal Hallucination via Phase-wise Self-reward
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
von: Chang, Kai-Po, et al.
Veröffentlicht: (2025)
von: Chang, Kai-Po, et al.
Veröffentlicht: (2025)
On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
L0: Reinforcement Learning to Become General Agents
von: Zhang, Junjie, et al.
Veröffentlicht: (2025)
von: Zhang, Junjie, et al.
Veröffentlicht: (2025)
Mitigating Multimodal Hallucinations via Gradient-based Self-Reflection
von: Wang, Shan, et al.
Veröffentlicht: (2025)
von: Wang, Shan, et al.
Veröffentlicht: (2025)
Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation
von: Wu, Jingwen, et al.
Veröffentlicht: (2026)
von: Wu, Jingwen, et al.
Veröffentlicht: (2026)
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
von: Alnuhait, Deema, et al.
Veröffentlicht: (2024)
von: Alnuhait, Deema, et al.
Veröffentlicht: (2024)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
von: Chen, Junzhe, et al.
Veröffentlicht: (2024)
von: Chen, Junzhe, et al.
Veröffentlicht: (2024)
One SPACE to Rule Them All: Jointly Mitigating Factuality and Faithfulness Hallucinations in LLMs
von: Wang, Pengbo, et al.
Veröffentlicht: (2025)
von: Wang, Pengbo, et al.
Veröffentlicht: (2025)
Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation
von: Zhang, Qianchi, et al.
Veröffentlicht: (2026)
von: Zhang, Qianchi, et al.
Veröffentlicht: (2026)
The Unintended Trade-off of AI Alignment:Balancing Hallucination Mitigation and Safety in LLMs
von: Mahmoud, Omar, et al.
Veröffentlicht: (2025)
von: Mahmoud, Omar, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Healthcare LLMs with Granular Fact-Checking and Domain-Specific Adaptation
von: Zeba, Musarrat, et al.
Veröffentlicht: (2025)
von: Zeba, Musarrat, et al.
Veröffentlicht: (2025)
Less is More: Mitigating Multimodal Hallucination from an EOS Decision Perspective
von: Yue, Zihao, et al.
Veröffentlicht: (2024)
von: Yue, Zihao, et al.
Veröffentlicht: (2024)
Mitigating Hallucination on Hallucination in RAG via Ensemble Voting
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning
von: Gao, Cheng, et al.
Veröffentlicht: (2026)
von: Gao, Cheng, et al.
Veröffentlicht: (2026)
Developing a Reliable, Fast, General-Purpose Hallucination Detection and Mitigation Service
von: Wang, Song, et al.
Veröffentlicht: (2024)
von: Wang, Song, et al.
Veröffentlicht: (2024)
Hallucination Detection and Hallucination Mitigation: An Investigation
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
Self-Filtered Distillation with LLMs-generated Trust Indicators for Reliable Patent Classification
von: Yoo, Yongmin, et al.
Veröffentlicht: (2025)
von: Yoo, Yongmin, et al.
Veröffentlicht: (2025)
Bridging External and Parametric Knowledge: Mitigating Hallucination of LLMs with Shared-Private Semantic Synergy in Dual-Stream Knowledge
von: Sui, Yi, et al.
Veröffentlicht: (2025)
von: Sui, Yi, et al.
Veröffentlicht: (2025)
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination
von: Li, Zhuo, et al.
Veröffentlicht: (2026)
von: Li, Zhuo, et al.
Veröffentlicht: (2026)
SAFE: A Sparse Autoencoder-Based Framework for Robust Query Enrichment and Hallucination Mitigation in LLMs
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Volcano: Mitigating Multimodal Hallucination through Self-Feedback Guided Revision
von: Lee, Seongyun, et al.
Veröffentlicht: (2023)
von: Lee, Seongyun, et al.
Veröffentlicht: (2023)
Hallucinations Undermine Trust; Metacognition is a Way Forward
von: Yona, Gal, et al.
Veröffentlicht: (2026)
von: Yona, Gal, et al.
Veröffentlicht: (2026)
Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
Hallucination Mitigating for Medical Report Generation
von: Zhao, Ruoqing, et al.
Veröffentlicht: (2026)
von: Zhao, Ruoqing, et al.
Veröffentlicht: (2026)
Ziya2: Data-centric Learning is All LLMs Need
von: Gan, Ruyi, et al.
Veröffentlicht: (2023)
von: Gan, Ruyi, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024) -
Lyrics: Boosting Fine-grained Language-Vision Alignment and Comprehension via Semantic-aware Visual Objects
von: Lu, Junyu, et al.
Veröffentlicht: (2023) -
Never Lost in the Middle: Mastering Long-Context Question Answering with Position-Agnostic Decompositional Training
von: He, Junqing, et al.
Veröffentlicht: (2023) -
Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization
von: Compagnoni, Alberto, et al.
Veröffentlicht: (2025) -
Rowen: Adaptive Retrieval-Augmented Generation for Hallucination Mitigation in LLMs
von: Ding, Hanxing, et al.
Veröffentlicht: (2024)