Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Zhihe, Luo, Xufang, Han, Dongqi, Xu, Yunjian, Li, Dongsheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
di: Xie, Yuxi, et al.
Pubblicazione: (2024)
di: Xie, Yuxi, et al.
Pubblicazione: (2024)
Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning
di: Park, Dongmin, et al.
Pubblicazione: (2024)
di: Park, Dongmin, et al.
Pubblicazione: (2024)
Visual Attention Drifts,but Anchors Hold:Mitigating Hallucination in Multimodal Large Language Models via Cross-Layer Visual Anchors
di: Yang, Chengxu, et al.
Pubblicazione: (2026)
di: Yang, Chengxu, et al.
Pubblicazione: (2026)
Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens
di: Jiang, Zhangqi, et al.
Pubblicazione: (2024)
di: Jiang, Zhangqi, et al.
Pubblicazione: (2024)
VisRL: Intention-Driven Visual Perception via Reinforced Reasoning
di: Chen, Zhangquan, et al.
Pubblicazione: (2025)
di: Chen, Zhangquan, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
di: Lu, Yifan, et al.
Pubblicazione: (2025)
di: Lu, Yifan, et al.
Pubblicazione: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
di: Zhu, Younan, et al.
Pubblicazione: (2025)
di: Zhu, Younan, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Large Vision-Language Models via Causal Route Gating
di: Cheng, Zhe, et al.
Pubblicazione: (2026)
di: Cheng, Zhe, et al.
Pubblicazione: (2026)
Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation
di: Zhu, Xingyu, et al.
Pubblicazione: (2026)
di: Zhu, Xingyu, et al.
Pubblicazione: (2026)
MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction
di: Wang, Chao, et al.
Pubblicazione: (2025)
di: Wang, Chao, et al.
Pubblicazione: (2025)
Exploring Causes and Mitigation of Hallucinations in Large Vision Language Models
di: Sun, Yaqi, et al.
Pubblicazione: (2025)
di: Sun, Yaqi, et al.
Pubblicazione: (2025)
CLIP-DPO: Vision-Language Models as a Source of Preference for Fixing Hallucinations in LVLMs
di: Ouali, Yassine, et al.
Pubblicazione: (2024)
di: Ouali, Yassine, et al.
Pubblicazione: (2024)
Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection
di: Yang, Le, et al.
Pubblicazione: (2024)
di: Yang, Le, et al.
Pubblicazione: (2024)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
di: Wang, Zihu, et al.
Pubblicazione: (2025)
di: Wang, Zihu, et al.
Pubblicazione: (2025)
A Comprehensive Information-Decomposition Analysis of Large Vision-Language Models
di: Xiu, Lixin, et al.
Pubblicazione: (2026)
di: Xiu, Lixin, et al.
Pubblicazione: (2026)
Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models
di: Zhang, Jinrui, et al.
Pubblicazione: (2024)
di: Zhang, Jinrui, et al.
Pubblicazione: (2024)
HII-DPO: Eliminate Hallucination via Accurate Hallucination-Inducing Counterfactual Images
di: Yang, Yilin, et al.
Pubblicazione: (2026)
di: Yang, Yilin, et al.
Pubblicazione: (2026)
SAKED: Mitigating Hallucination in Large Vision-Language Models via Stability-Aware Knowledge Enhanced Decoding
di: Li, Zhaoxu, et al.
Pubblicazione: (2026)
di: Li, Zhaoxu, et al.
Pubblicazione: (2026)
SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
di: Li, Zhaoxu, et al.
Pubblicazione: (2025)
di: Li, Zhaoxu, et al.
Pubblicazione: (2025)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
di: Li, Bin, et al.
Pubblicazione: (2025)
di: Li, Bin, et al.
Pubblicazione: (2025)
InpaintDPO: Mitigating Spatial Relationship Hallucinations in Foreground-conditioned Inpainting via Diverse Preference Optimization
di: Li, Qirui, et al.
Pubblicazione: (2025)
di: Li, Qirui, et al.
Pubblicazione: (2025)
Residual Decoding: Mitigating Hallucinations in Large Vision-Language Models via History-Aware Residual Guidance
di: Chen, Xinrong, et al.
Pubblicazione: (2026)
di: Chen, Xinrong, et al.
Pubblicazione: (2026)
Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models
di: Zhang, Chengsheng, et al.
Pubblicazione: (2026)
di: Zhang, Chengsheng, et al.
Pubblicazione: (2026)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
di: Wang, Weihang, et al.
Pubblicazione: (2025)
di: Wang, Weihang, et al.
Pubblicazione: (2025)
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
di: Li, Qiming, et al.
Pubblicazione: (2026)
di: Li, Qiming, et al.
Pubblicazione: (2026)
Segmentation-Based Attention Entropy: Detecting and Mitigating Object Hallucinations in Large Vision-Language Models
di: Song, Jiale, et al.
Pubblicazione: (2026)
di: Song, Jiale, et al.
Pubblicazione: (2026)
A Large-scale Medical Visual Task Adaptation Benchmark
di: Mo, Shentong, et al.
Pubblicazione: (2024)
di: Mo, Shentong, et al.
Pubblicazione: (2024)
Mitigating Image Captioning Hallucinations in Vision-Language Models
di: Zhao, Fei, et al.
Pubblicazione: (2025)
di: Zhao, Fei, et al.
Pubblicazione: (2025)
KVSmooth: Mitigating Hallucination in Multi-modal Large Language Models through Key-Value Smoothing
di: Jiang, Siyu, et al.
Pubblicazione: (2026)
di: Jiang, Siyu, et al.
Pubblicazione: (2026)
pMoE: Prompting Diverse Experts Together Wins More in Visual Adaptation
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
di: Fazli, Mehrdad, et al.
Pubblicazione: (2025)
di: Fazli, Mehrdad, et al.
Pubblicazione: (2025)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
di: Zhong, Weihong, et al.
Pubblicazione: (2024)
di: Zhong, Weihong, et al.
Pubblicazione: (2024)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
YARD: Y-Architecture Register Decoding for Efficient Hallucination Mitigation in Large Vision-Language Models
di: Chen, Ting, et al.
Pubblicazione: (2026)
di: Chen, Ting, et al.
Pubblicazione: (2026)
VORD: Visual Ordinal Calibration for Mitigating Object Hallucinations in Large Vision-Language Models
di: Neo, Dexter, et al.
Pubblicazione: (2024)
di: Neo, Dexter, et al.
Pubblicazione: (2024)
HTDC: Hesitation-Triggered Differential Calibration for Mitigating Hallucination in Large Vision-Language Models
di: Liu, Xinyun
Pubblicazione: (2026)
di: Liu, Xinyun
Pubblicazione: (2026)
Mitigating Hallucinations in Video Large Language Models via Spatiotemporal-Semantic Contrastive Decoding
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
di: Zhang, Yuanhong, et al.
Pubblicazione: (2026)
di: Zhang, Yuanhong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
di: Xie, Yuxi, et al.
Pubblicazione: (2024) -
Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning
di: Park, Dongmin, et al.
Pubblicazione: (2024) -
Visual Attention Drifts,but Anchors Hold:Mitigating Hallucination in Multimodal Large Language Models via Cross-Layer Visual Anchors
di: Yang, Chengxu, et al.
Pubblicazione: (2026) -
Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens
di: Jiang, Zhangqi, et al.
Pubblicazione: (2024) -
VisRL: Intention-Driven Visual Perception via Reinforced Reasoning
di: Chen, Zhangquan, et al.
Pubblicazione: (2025)