Causal Decoding for Hallucination-Resistant Multimodal Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tan, Shiwei, Wang, Hengyi, Qin, Weiyi, Xu, Qi, Hua, Zhigang, Wang, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models
von: Park, Yeji, et al.
Veröffentlicht: (2024)
von: Park, Yeji, et al.
Veröffentlicht: (2024)
Woodpecker: Hallucination Correction for Multimodal Large Language Models
von: Yin, Shukang, et al.
Veröffentlicht: (2023)
von: Yin, Shukang, et al.
Veröffentlicht: (2023)
Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models
von: Huo, Fushuo, et al.
Veröffentlicht: (2024)
von: Huo, Fushuo, et al.
Veröffentlicht: (2024)
Learning to Decode Against Compositional Hallucination in Video Multimodal Large Language Models
von: Xing, Wenbin, et al.
Veröffentlicht: (2026)
von: Xing, Wenbin, et al.
Veröffentlicht: (2026)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
von: Lee, Jihoon, et al.
Veröffentlicht: (2025)
von: Lee, Jihoon, et al.
Veröffentlicht: (2025)
CoFi-Dec: Hallucination-Resistant Decoding via Coarse-to-Fine Generative Feedback in Large Vision-Language Models
von: Cao, Zongsheng, et al.
Veröffentlicht: (2025)
von: Cao, Zongsheng, et al.
Veröffentlicht: (2025)
Automatically Generating Visual Hallucination Test Cases for Multimodal Large Language Models
von: Liu, Zhongye, et al.
Veröffentlicht: (2024)
von: Liu, Zhongye, et al.
Veröffentlicht: (2024)
Mitigating Hallucinations in Video Large Language Models via Spatiotemporal-Semantic Contrastive Decoding
von: Gao, Yuansheng, et al.
Veröffentlicht: (2026)
von: Gao, Yuansheng, et al.
Veröffentlicht: (2026)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models
von: Xia, Yuxuan, et al.
Veröffentlicht: (2026)
von: Xia, Yuxuan, et al.
Veröffentlicht: (2026)
SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding
von: Wu, Chang-Hsun, et al.
Veröffentlicht: (2025)
von: Wu, Chang-Hsun, et al.
Veröffentlicht: (2025)
V-ITI: Mitigating Hallucinations in Multimodal Large Language Models via Visual Inference-Time Intervention
von: Sun, Nan, et al.
Veröffentlicht: (2025)
von: Sun, Nan, et al.
Veröffentlicht: (2025)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
von: Wu, Junfei, et al.
Veröffentlicht: (2024)
von: Wu, Junfei, et al.
Veröffentlicht: (2024)
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
NoiseBoost: Alleviating Hallucination with Noise Perturbation for Multimodal Large Language Models
von: Wu, Kai, et al.
Veröffentlicht: (2024)
von: Wu, Kai, et al.
Veröffentlicht: (2024)
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models
von: Sun, Li, et al.
Veröffentlicht: (2024)
von: Sun, Li, et al.
Veröffentlicht: (2024)
Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models
von: Miao, Yanting, et al.
Veröffentlicht: (2026)
von: Miao, Yanting, et al.
Veröffentlicht: (2026)
ClearSight: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models
von: Yin, Hao, et al.
Veröffentlicht: (2025)
von: Yin, Hao, et al.
Veröffentlicht: (2025)
Visual Hallucinations of Multi-modal Large Language Models
von: Huang, Wen, et al.
Veröffentlicht: (2024)
von: Huang, Wen, et al.
Veröffentlicht: (2024)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
von: Zhong, Weihong, et al.
Veröffentlicht: (2024)
von: Zhong, Weihong, et al.
Veröffentlicht: (2024)
Mixture of Decoding: An Attention-Inspired Adaptive Decoding Strategy to Mitigate Hallucinations in Large Vision-Language Models
von: Chen, Xinlong, et al.
Veröffentlicht: (2025)
von: Chen, Xinlong, et al.
Veröffentlicht: (2025)
Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models
von: Zhang, Gengwei, et al.
Veröffentlicht: (2026)
von: Zhang, Gengwei, et al.
Veröffentlicht: (2026)
Who Brings the Frisbee: Probing Hidden Hallucination Factors in Large Vision-Language Model via Causality Analysis
von: Huang, Po-Hsuan, et al.
Veröffentlicht: (2024)
von: Huang, Po-Hsuan, et al.
Veröffentlicht: (2024)
THRONE: An Object-based Hallucination Benchmark for the Free-form Generations of Large Vision-Language Models
von: Kaul, Prannay, et al.
Veröffentlicht: (2024)
von: Kaul, Prannay, et al.
Veröffentlicht: (2024)
Can Multimodal Large Language Models Truly Perform Multimodal In-Context Learning?
von: Chen, Shuo, et al.
Veröffentlicht: (2023)
von: Chen, Shuo, et al.
Veröffentlicht: (2023)
Speculative Decoding Reimagined for Multimodal Large Language Models
von: Lin, Luxi, et al.
Veröffentlicht: (2025)
von: Lin, Luxi, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
von: Wang, Xintong, et al.
Veröffentlicht: (2024)
von: Wang, Xintong, et al.
Veröffentlicht: (2024)
HalluRNN: Mitigating Hallucinations via Recurrent Cross-Layer Reasoning in Large Vision-Language Models
von: Yu, Le, et al.
Veröffentlicht: (2025)
von: Yu, Le, et al.
Veröffentlicht: (2025)
When Graph meets Multimodal: Benchmarking and Meditating on Multimodal Attributed Graphs Learning
von: Yan, Hao, et al.
Veröffentlicht: (2024)
von: Yan, Hao, et al.
Veröffentlicht: (2024)
Residual Decoding: Mitigating Hallucinations in Large Vision-Language Models via History-Aware Residual Guidance
von: Chen, Xinrong, et al.
Veröffentlicht: (2026)
von: Chen, Xinrong, et al.
Veröffentlicht: (2026)
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
von: Deng, Ailin, et al.
Veröffentlicht: (2024)
von: Deng, Ailin, et al.
Veröffentlicht: (2024)
Learning to Inference Adaptively for Multimodal Large Language Models
von: Xu, Zhuoyan, et al.
Veröffentlicht: (2025)
von: Xu, Zhuoyan, et al.
Veröffentlicht: (2025)
TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation
von: Gong, Han, et al.
Veröffentlicht: (2026)
von: Gong, Han, et al.
Veröffentlicht: (2026)
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images
von: You, Liangliang, et al.
Veröffentlicht: (2025)
von: You, Liangliang, et al.
Veröffentlicht: (2025)
DiG: Differential Grounding for Enhancing Fine-Grained Perception in Multimodal Large Language Model
von: Tao, Zhou, et al.
Veröffentlicht: (2025)
von: Tao, Zhou, et al.
Veröffentlicht: (2025)
Multimodal Causal Reasoning Benchmark: Challenging Vision Large Language Models to Discern Causal Links Across Modalities
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
Invariant Representation Guided Multimodal Sentiment Decoding with Sequential Variation Regularization
von: Xu, Guoyang, et al.
Veröffentlicht: (2024)
von: Xu, Guoyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models
von: Wang, Hengyi, et al.
Veröffentlicht: (2024) -
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
von: Wang, Hengyi, et al.
Veröffentlicht: (2024) -
ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models
von: Park, Yeji, et al.
Veröffentlicht: (2024) -
Woodpecker: Hallucination Correction for Multimodal Large Language Models
von: Yin, Shukang, et al.
Veröffentlicht: (2023) -
Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models
von: Huo, Fushuo, et al.
Veröffentlicht: (2024)