CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Qiming, Ye, Zekai, Feng, Xiaocheng, Zhong, Weihong, Qin, Libo, Chen, Ruihan, Huang, Lei, Li, Baohang, Jiang, Kui, Wang, Yaowei, Liu, Ting, Qin, Bing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
von: Li, Qiming, et al.
Veröffentlicht: (2025)
von: Li, Qiming, et al.
Veröffentlicht: (2025)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
von: Ye, Zekai, et al.
Veröffentlicht: (2025)
von: Ye, Zekai, et al.
Veröffentlicht: (2025)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
von: Li, Qiming, et al.
Veröffentlicht: (2025)
von: Li, Qiming, et al.
Veröffentlicht: (2025)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
von: Zhong, Weihong, et al.
Veröffentlicht: (2024)
von: Zhong, Weihong, et al.
Veröffentlicht: (2024)
Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
von: Li, Qiming, et al.
Veröffentlicht: (2025)
von: Li, Qiming, et al.
Veröffentlicht: (2025)
Not All Tokens See Equally: Perception-Grounded Policy Optimization for Large Vision-Language Models
von: Ye, Zekai, et al.
Veröffentlicht: (2026)
von: Ye, Zekai, et al.
Veröffentlicht: (2026)
Mitigating Object Hallucinations in LVLMs via Attention Imbalance Rectification
von: Sun, Han, et al.
Veröffentlicht: (2026)
von: Sun, Han, et al.
Veröffentlicht: (2026)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
von: Wang, Zihu, et al.
Veröffentlicht: (2025)
von: Wang, Zihu, et al.
Veröffentlicht: (2025)
Aligning Translation-Specific Understanding to General Understanding in Large Language Models
von: Huang, Yichong, et al.
Veröffentlicht: (2024)
von: Huang, Yichong, et al.
Veröffentlicht: (2024)
Ensemble Learning for Heterogeneous Large Language Models with Deep Parallel Collaboration
von: Huang, Yichong, et al.
Veröffentlicht: (2024)
von: Huang, Yichong, et al.
Veröffentlicht: (2024)
Mitigating Image Captioning Hallucinations in Vision-Language Models
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations
von: Chen, Boxu, et al.
Veröffentlicht: (2025)
von: Chen, Boxu, et al.
Veröffentlicht: (2025)
Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation
von: Yuan, Zekun, et al.
Veröffentlicht: (2026)
von: Yuan, Zekun, et al.
Veröffentlicht: (2026)
MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents
von: Chen, Ruihan, et al.
Veröffentlicht: (2025)
von: Chen, Ruihan, et al.
Veröffentlicht: (2025)
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
von: Huang, Lei, et al.
Veröffentlicht: (2023)
von: Huang, Lei, et al.
Veröffentlicht: (2023)
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
von: Ma, Weitao, et al.
Veröffentlicht: (2024)
von: Ma, Weitao, et al.
Veröffentlicht: (2024)
Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models
von: Wu, Jialin, et al.
Veröffentlicht: (2026)
von: Wu, Jialin, et al.
Veröffentlicht: (2026)
Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement
von: Qin, Zhenxin, et al.
Veröffentlicht: (2026)
von: Qin, Zhenxin, et al.
Veröffentlicht: (2026)
Mitigating Object Hallucination via Concentric Causal Attention
von: Xing, Yun, et al.
Veröffentlicht: (2024)
von: Xing, Yun, et al.
Veröffentlicht: (2024)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
von: Li, Bin, et al.
Veröffentlicht: (2025)
von: Li, Bin, et al.
Veröffentlicht: (2025)
Cross-Lingual Text-Rich Visual Comprehension: An Information Theory Perspective
von: Yu, Xinmiao, et al.
Veröffentlicht: (2024)
von: Yu, Xinmiao, et al.
Veröffentlicht: (2024)
Relay Decoding: Concatenating Large Language Models for Machine Translation
von: Fu, Chengpeng, et al.
Veröffentlicht: (2024)
von: Fu, Chengpeng, et al.
Veröffentlicht: (2024)
MHSA: A Lightweight Framework for Mitigating Hallucinations via Steered Attention in LVLMs
von: Ding, Wei, et al.
Veröffentlicht: (2026)
von: Ding, Wei, et al.
Veröffentlicht: (2026)
Stratagem: Learning Transferable Reasoning via Trajectory-Modulated Game Self-Play
von: Feng, Xiachong, et al.
Veröffentlicht: (2026)
von: Feng, Xiachong, et al.
Veröffentlicht: (2026)
Mitigating Audiovisual Mismatch in Visual-Guide Audio Captioning
von: Xu, Le, et al.
Veröffentlicht: (2025)
von: Xu, Le, et al.
Veröffentlicht: (2025)
Can Large Language Models Simulate Human Cognition Beyond Behavioral Imitation?
von: Gu, Yuxuan, et al.
Veröffentlicht: (2026)
von: Gu, Yuxuan, et al.
Veröffentlicht: (2026)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
von: Zhu, Younan, et al.
Veröffentlicht: (2025)
von: Zhu, Younan, et al.
Veröffentlicht: (2025)
HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding
von: Yuan, Fan, et al.
Veröffentlicht: (2024)
von: Yuan, Fan, et al.
Veröffentlicht: (2024)
COMPASS: Context-Modulated PID Attention Steering System for Hallucination Mitigation
von: Sahay, Kenji, et al.
Veröffentlicht: (2025)
von: Sahay, Kenji, et al.
Veröffentlicht: (2025)
Mitigating Open-Vocabulary Caption Hallucinations
von: Ben-Kish, Assaf, et al.
Veröffentlicht: (2023)
von: Ben-Kish, Assaf, et al.
Veröffentlicht: (2023)
Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
von: Zhang, Yuanhong, et al.
Veröffentlicht: (2026)
von: Zhang, Yuanhong, et al.
Veröffentlicht: (2026)
Look Closer! An Adversarial Parametric Editing Framework for Hallucination Mitigation in VLMs
von: Hu, Jiayu, et al.
Veröffentlicht: (2025)
von: Hu, Jiayu, et al.
Veröffentlicht: (2025)
Energy-Guided Decoding for Object Hallucination Mitigation
von: Liu, Xixi, et al.
Veröffentlicht: (2025)
von: Liu, Xixi, et al.
Veröffentlicht: (2025)
Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration
von: Xu, Yuanzhi, et al.
Veröffentlicht: (2026)
von: Xu, Yuanzhi, et al.
Veröffentlicht: (2026)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
von: An, Wenbin, et al.
Veröffentlicht: (2024)
von: An, Wenbin, et al.
Veröffentlicht: (2024)
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding
von: Zhao, Liang, et al.
Veröffentlicht: (2023)
von: Zhao, Liang, et al.
Veröffentlicht: (2023)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models
von: Zou, Zhengtao, et al.
Veröffentlicht: (2025)
von: Zou, Zhengtao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
von: Li, Qiming, et al.
Veröffentlicht: (2025) -
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
von: Ye, Zekai, et al.
Veröffentlicht: (2025) -
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
von: Li, Qiming, et al.
Veröffentlicht: (2025) -
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
von: Zhong, Weihong, et al.
Veröffentlicht: (2024) -
Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
von: Li, Qiming, et al.
Veröffentlicht: (2025)