Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Feilong, Liu, Chengzhi, Xu, Zhongxing, Hu, Ming, Peng, Zelin, Yang, Zhiwei, Su, Jionglong, Lin, Minquan, Peng, Yifan, Cheng, Xuelian, Razzak, Imran, Ge, Zongyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
von: Zhong, Yiheng, et al.
Veröffentlicht: (2025)
von: Zhong, Yiheng, et al.
Veröffentlicht: (2025)
Thinking in Uncertainty: Mitigating Hallucinations in MLRMs with Latent Entropy-Aware Decoding
von: Xu, Zhongxing, et al.
Veröffentlicht: (2026)
von: Xu, Zhongxing, et al.
Veröffentlicht: (2026)
Neighbor Does Matter: Density-Aware Contrastive Learning for Medical Semi-supervised Segmentation
von: Tang, Feilong, et al.
Veröffentlicht: (2024)
von: Tang, Feilong, et al.
Veröffentlicht: (2024)
Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP
von: Xu, Zhongxing, et al.
Veröffentlicht: (2024)
von: Xu, Zhongxing, et al.
Veröffentlicht: (2024)
ScalingNoise: Scaling Inference-Time Search for Generating Infinite Videos
von: Yang, Haolin, et al.
Veröffentlicht: (2025)
von: Yang, Haolin, et al.
Veröffentlicht: (2025)
MMRC: A Large-Scale Benchmark for Understanding Multimodal Large Language Model in Real-World Conversation
von: Xue, Haochen, et al.
Veröffentlicht: (2025)
von: Xue, Haochen, et al.
Veröffentlicht: (2025)
Unveiling the Ignorance of MLLMs: Seeing Clearly, Answering Incorrectly
von: Liu, Yexin, et al.
Veröffentlicht: (2024)
von: Liu, Yexin, et al.
Veröffentlicht: (2024)
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
von: Kolli, Govinda, et al.
Veröffentlicht: (2026)
von: Kolli, Govinda, et al.
Veröffentlicht: (2026)
SAM-DCE: Addressing Token Uniformity and Semantic Over-Smoothing in Medical Segmentation
von: Hu, Yingzhen, et al.
Veröffentlicht: (2025)
von: Hu, Yingzhen, et al.
Veröffentlicht: (2025)
Robust Multimodal Learning for Ophthalmic Disease Grading via Disentangled Representation
von: Wang, Xinkun, et al.
Veröffentlicht: (2025)
von: Wang, Xinkun, et al.
Veröffentlicht: (2025)
Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate
von: Lu, Zhixiang, et al.
Veröffentlicht: (2026)
von: Lu, Zhixiang, et al.
Veröffentlicht: (2026)
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
Confidence-Aware Self-Distillation for Multimodal Sentiment Analysis with Incomplete Modalities
von: Luo, Yanxi, et al.
Veröffentlicht: (2025)
von: Luo, Yanxi, et al.
Veröffentlicht: (2025)
Vision-Language Introspection: Mitigating Overconfident Hallucinations in MLLMs via Interpretable Bi-Causal Steering
von: Liu, Shuliang, et al.
Veröffentlicht: (2026)
von: Liu, Shuliang, et al.
Veröffentlicht: (2026)
Optimizing LVLMs with On-Policy Data for Effective Hallucination Mitigation
von: Yu, Chengzhi, et al.
Veröffentlicht: (2025)
von: Yu, Chengzhi, et al.
Veröffentlicht: (2025)
Hunting Attributes: Context Prototype-Aware Learning for Weakly Supervised Semantic Segmentation
von: Tang, Feilong, et al.
Veröffentlicht: (2024)
von: Tang, Feilong, et al.
Veröffentlicht: (2024)
Decoding the Flow: CauseMotion for Emotional Causality Analysis in Long-form Conversations
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2025)
Discriminating retinal microvascular and neuronal differences related to migraines: Deep Learning based Crossectional Study
von: Tang, Feilong, et al.
Veröffentlicht: (2024)
von: Tang, Feilong, et al.
Veröffentlicht: (2024)
Seeing Clearly without Training: Mitigating Hallucinations in Multimodal LLMs for Remote Sensing
von: Liu, Yi, et al.
Veröffentlicht: (2026)
von: Liu, Yi, et al.
Veröffentlicht: (2026)
Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2024)
More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models
von: Liu, Chengzhi, et al.
Veröffentlicht: (2025)
von: Liu, Chengzhi, et al.
Veröffentlicht: (2025)
See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment
von: Azeez, Mohammad Anas, et al.
Veröffentlicht: (2026)
von: Azeez, Mohammad Anas, et al.
Veröffentlicht: (2026)
CCD: Mitigating Hallucinations in Radiology MLLMs via Clinical Contrastive Decoding
von: Zhang, Xi, et al.
Veröffentlicht: (2025)
von: Zhang, Xi, et al.
Veröffentlicht: (2025)
Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic Surgery
von: Hu, Ming, et al.
Veröffentlicht: (2025)
von: Hu, Ming, et al.
Veröffentlicht: (2025)
DeLo: Dual Decomposed Low-Rank Experts Collaboration for Continual Missing Modality Learning
von: Liu, Xiwei, et al.
Veröffentlicht: (2026)
von: Liu, Xiwei, et al.
Veröffentlicht: (2026)
MSWAL: 3D Multi-class Segmentation of Whole Abdominal Lesions Dataset
von: Wu, Zhaodong, et al.
Veröffentlicht: (2025)
von: Wu, Zhaodong, et al.
Veröffentlicht: (2025)
Rhythm of Opinion: A Hawkes-Graph Framework for Dynamic Propagation Analysis
von: Li, Yulong, et al.
Veröffentlicht: (2025)
von: Li, Yulong, et al.
Veröffentlicht: (2025)
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
von: Hu, Ming, et al.
Veröffentlicht: (2024)
von: Hu, Ming, et al.
Veröffentlicht: (2024)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
von: Bozorgtabar, Behzad, et al.
Veröffentlicht: (2026)
von: Bozorgtabar, Behzad, et al.
Veröffentlicht: (2026)
Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?
von: Yin, Hao, et al.
Veröffentlicht: (2025)
von: Yin, Hao, et al.
Veröffentlicht: (2025)
PRISM: A Personality-Driven Multi-Agent Framework for Social Media Simulation
von: Lu, Zhixiang, et al.
Veröffentlicht: (2025)
von: Lu, Zhixiang, et al.
Veröffentlicht: (2025)
StreamAgent: Towards Anticipatory Agents for Streaming Video Understanding
von: Yang, Haolin, et al.
Veröffentlicht: (2025)
von: Yang, Haolin, et al.
Veröffentlicht: (2025)
MRFD: Multi-Region Fusion Decoding with Self-Consistency for Mitigating Hallucinations in LVLMs
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
Mitigating Object Hallucination via Concentric Causal Attention
von: Xing, Yun, et al.
Veröffentlicht: (2024)
von: Xing, Yun, et al.
Veröffentlicht: (2024)
Delta -- Contrastive Decoding Mitigates Text Hallucinations in Large Language Models
von: Huang, Cheng Peng, et al.
Veröffentlicht: (2025)
von: Huang, Cheng Peng, et al.
Veröffentlicht: (2025)
ColonAdapter: Geometry Estimation Through Foundation Model Adaptation for Colonoscopy
von: Jiang, Zhiyi, et al.
Veröffentlicht: (2025)
von: Jiang, Zhiyi, et al.
Veröffentlicht: (2025)
Towards Efficient Medical Reasoning with Minimal Fine-Tuning Data
von: Zhuang, Xinlin, et al.
Veröffentlicht: (2025)
von: Zhuang, Xinlin, et al.
Veröffentlicht: (2025)
COPO: Causal-Oriented Policy Optimization for Hallucinations of MLLMs
von: Guo, Peizheng, et al.
Veröffentlicht: (2025)
von: Guo, Peizheng, et al.
Veröffentlicht: (2025)
Beyond Words: AuralLLM and SignMST-C for Sign Language Production and Bidirectional Accessibility
von: Li, Yulong, et al.
Veröffentlicht: (2025)
von: Li, Yulong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
von: Zhong, Yiheng, et al.
Veröffentlicht: (2025) -
Thinking in Uncertainty: Mitigating Hallucinations in MLRMs with Latent Entropy-Aware Decoding
von: Xu, Zhongxing, et al.
Veröffentlicht: (2026) -
Neighbor Does Matter: Density-Aware Contrastive Learning for Medical Semi-supervised Segmentation
von: Tang, Feilong, et al.
Veröffentlicht: (2024) -
Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP
von: Xu, Zhongxing, et al.
Veröffentlicht: (2024) -
ScalingNoise: Scaling Inference-Time Search for Generating Infinite Videos
von: Yang, Haolin, et al.
Veröffentlicht: (2025)