Look, Compare, Decide: Alleviating Hallucination in Large Vision-Language Models via Multi-View Multi-Path Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qu, Xiaoye, Sun, Jiashuo, Wei, Wei, Cheng, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Alleviating Hallucination in Large Vision-Language Models with Active Retrieval Augmentation
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
Mitigating Multilingual Hallucination in Large Vision-Language Models
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
Revealing Multi-View Hallucination in Large Vision-Language Models
von: Park, Wooje, et al.
Veröffentlicht: (2026)
von: Park, Wooje, et al.
Veröffentlicht: (2026)
SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information
von: Sun, Jiashuo, et al.
Veröffentlicht: (2024)
von: Sun, Jiashuo, et al.
Veröffentlicht: (2024)
SATORI-R1: Incentivizing Multimodal Reasoning through Explicit Visual Anchoring
von: Shen, Chuming, et al.
Veröffentlicht: (2025)
von: Shen, Chuming, et al.
Veröffentlicht: (2025)
From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models
von: Huo, Fushuo, et al.
Veröffentlicht: (2024)
von: Huo, Fushuo, et al.
Veröffentlicht: (2024)
Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models
von: Zhang, Chengsheng, et al.
Veröffentlicht: (2026)
von: Zhang, Chengsheng, et al.
Veröffentlicht: (2026)
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
von: Suo, Wei, et al.
Veröffentlicht: (2025)
von: Suo, Wei, et al.
Veröffentlicht: (2025)
A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
von: Liu, Daizong, et al.
Veröffentlicht: (2024)
von: Liu, Daizong, et al.
Veröffentlicht: (2024)
Look Before You Decide: Prompting Active Deduction of MLLMs for Assumptive Reasoning
von: Li, Yian, et al.
Veröffentlicht: (2024)
von: Li, Yian, et al.
Veröffentlicht: (2024)
Multi-Object Hallucination in Vision-Language Models
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
Prescribing the Right Remedy: Mitigating Hallucinations in Large Vision-Language Models via Targeted Instruction Tuning
von: Hu, Rui, et al.
Veröffentlicht: (2024)
von: Hu, Rui, et al.
Veröffentlicht: (2024)
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
von: Zheng, Haojie, et al.
Veröffentlicht: (2024)
von: Zheng, Haojie, et al.
Veröffentlicht: (2024)
Alleviating Hallucinations in Large Vision-Language Models through Hallucination-Induced Optimization
von: Lyu, Xinyu, et al.
Veröffentlicht: (2024)
von: Lyu, Xinyu, et al.
Veröffentlicht: (2024)
NoiseBoost: Alleviating Hallucination with Noise Perturbation for Multimodal Large Language Models
von: Wu, Kai, et al.
Veröffentlicht: (2024)
von: Wu, Kai, et al.
Veröffentlicht: (2024)
From Captions to Rewards (CAREVL): Leveraging Large Language Model Experts for Enhanced Reward Modeling in Large Vision-Language Models
von: Dai, Muzhi, et al.
Veröffentlicht: (2025)
von: Dai, Muzhi, et al.
Veröffentlicht: (2025)
HalluRNN: Mitigating Hallucinations via Recurrent Cross-Layer Reasoning in Large Vision-Language Models
von: Yu, Le, et al.
Veröffentlicht: (2025)
von: Yu, Le, et al.
Veröffentlicht: (2025)
IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding
von: Zhu, Lanyun, et al.
Veröffentlicht: (2024)
von: Zhu, Lanyun, et al.
Veröffentlicht: (2024)
OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation
von: Huang, Qidong, et al.
Veröffentlicht: (2023)
von: Huang, Qidong, et al.
Veröffentlicht: (2023)
Review of Hallucination Understanding in Large Language and Vision Models
von: Ho, Zhengyi, et al.
Veröffentlicht: (2025)
von: Ho, Zhengyi, et al.
Veröffentlicht: (2025)
DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models
von: Zhang, Yudong, et al.
Veröffentlicht: (2024)
von: Zhang, Yudong, et al.
Veröffentlicht: (2024)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
von: Zhu, Younan, et al.
Veröffentlicht: (2025)
von: Zhu, Younan, et al.
Veröffentlicht: (2025)
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models
von: Lee, Yi-Lun, et al.
Veröffentlicht: (2024)
von: Lee, Yi-Lun, et al.
Veröffentlicht: (2024)
Visual Hallucinations of Multi-modal Large Language Models
von: Huang, Wen, et al.
Veröffentlicht: (2024)
von: Huang, Wen, et al.
Veröffentlicht: (2024)
Residual Decoding: Mitigating Hallucinations in Large Vision-Language Models via History-Aware Residual Guidance
von: Chen, Xinrong, et al.
Veröffentlicht: (2026)
von: Chen, Xinrong, et al.
Veröffentlicht: (2026)
INTER: Mitigating Hallucination in Large Vision-Language Models by Interaction Guidance Sampling
von: Dong, Xin, et al.
Veröffentlicht: (2025)
von: Dong, Xin, et al.
Veröffentlicht: (2025)
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors
von: Ren, Lingfeng, et al.
Veröffentlicht: (2026)
von: Ren, Lingfeng, et al.
Veröffentlicht: (2026)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
von: Zhang, Yuanhong, et al.
Veröffentlicht: (2026)
von: Zhang, Yuanhong, et al.
Veröffentlicht: (2026)
Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models
von: Seth, Ashish, et al.
Veröffentlicht: (2024)
von: Seth, Ashish, et al.
Veröffentlicht: (2024)
DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models
von: Wang, JiYang, et al.
Veröffentlicht: (2026)
von: Wang, JiYang, et al.
Veröffentlicht: (2026)
OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models
von: Yu, Chung-En Johnny, et al.
Veröffentlicht: (2025)
von: Yu, Chung-En Johnny, et al.
Veröffentlicht: (2025)
PruneHal: Reducing Hallucinations in Multi-modal Large Language Models through Adaptive KV Cache Pruning
von: Sun, Fengyuan, et al.
Veröffentlicht: (2025)
von: Sun, Fengyuan, et al.
Veröffentlicht: (2025)
Point-It-Out: Benchmarking Embodied Reasoning for Vision Language Models in Multi-Stage Visual Grounding
von: Xue, Haotian, et al.
Veröffentlicht: (2025)
von: Xue, Haotian, et al.
Veröffentlicht: (2025)
FRIEDA: Benchmarking Multi-Step Cartographic Reasoning in Vision-Language Models
von: Pyo, Jiyoon, et al.
Veröffentlicht: (2025)
von: Pyo, Jiyoon, et al.
Veröffentlicht: (2025)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
von: Moratelli, Nicholas, et al.
Veröffentlicht: (2026)
von: Moratelli, Nicholas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Alleviating Hallucination in Large Vision-Language Models with Active Retrieval Augmentation
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024) -
Mitigating Multilingual Hallucination in Large Vision-Language Models
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024) -
Revealing Multi-View Hallucination in Large Vision-Language Models
von: Park, Wooje, et al.
Veröffentlicht: (2026) -
SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information
von: Sun, Jiashuo, et al.
Veröffentlicht: (2024) -
SATORI-R1: Incentivizing Multimodal Reasoning through Explicit Visual Anchoring
von: Shen, Chuming, et al.
Veröffentlicht: (2025)