Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
Fuente:
arXiv
Salvato in:
| Autori principali: | Geigle, Gregor, Timofte, Radu, Glavaš, Goran |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
di: Geigle, Gregor, et al.
Pubblicazione: (2024)
di: Geigle, Gregor, et al.
Pubblicazione: (2024)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model
di: Geigle, Gregor, et al.
Pubblicazione: (2025)
di: Geigle, Gregor, et al.
Pubblicazione: (2025)
InstructIR: High-Quality Image Restoration Following Human Instructions
di: Conde, Marcos V., et al.
Pubblicazione: (2024)
di: Conde, Marcos V., et al.
Pubblicazione: (2024)
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
di: Shang, Yuying, et al.
Pubblicazione: (2024)
di: Shang, Yuying, et al.
Pubblicazione: (2024)
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models
di: Jing, Liqiang, et al.
Pubblicazione: (2025)
di: Jing, Liqiang, et al.
Pubblicazione: (2025)
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
di: Ha, Jiwoo, et al.
Pubblicazione: (2026)
di: Ha, Jiwoo, et al.
Pubblicazione: (2026)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
di: Ma, Ruiqi, et al.
Pubblicazione: (2025)
di: Ma, Ruiqi, et al.
Pubblicazione: (2025)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
di: Zhou, Yiyang, et al.
Pubblicazione: (2023)
di: Zhou, Yiyang, et al.
Pubblicazione: (2023)
Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded Guidance
di: Zhao, Linxi, et al.
Pubblicazione: (2024)
di: Zhao, Linxi, et al.
Pubblicazione: (2024)
Grounding Language with Vision: A Conditional Mutual Information Calibrated Decoding Strategy for Reducing Hallucinations in LVLMs
di: Fang, Hao, et al.
Pubblicazione: (2025)
di: Fang, Hao, et al.
Pubblicazione: (2025)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
di: Lu, Yifan, et al.
Pubblicazione: (2025)
di: Lu, Yifan, et al.
Pubblicazione: (2025)
Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models
di: Lovenia, Holy, et al.
Pubblicazione: (2023)
di: Lovenia, Holy, et al.
Pubblicazione: (2023)
Multi-Object Hallucination in Vision-Language Models
di: Chen, Xuweiyi, et al.
Pubblicazione: (2024)
di: Chen, Xuweiyi, et al.
Pubblicazione: (2024)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
di: Seo, Hoigi, et al.
Pubblicazione: (2025)
di: Seo, Hoigi, et al.
Pubblicazione: (2025)
Do Vision-Language Models Really Understand Visual Language?
di: Hou, Yifan, et al.
Pubblicazione: (2024)
di: Hou, Yifan, et al.
Pubblicazione: (2024)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
di: An, Wenbin, et al.
Pubblicazione: (2024)
di: An, Wenbin, et al.
Pubblicazione: (2024)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
di: Wang, Weihang, et al.
Pubblicazione: (2025)
di: Wang, Weihang, et al.
Pubblicazione: (2025)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
di: Chang, Yue, et al.
Pubblicazione: (2024)
di: Chang, Yue, et al.
Pubblicazione: (2024)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2025)
di: Woo, Sangmin, et al.
Pubblicazione: (2025)
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
di: Agrawal, Aakriti, et al.
Pubblicazione: (2025)
di: Agrawal, Aakriti, et al.
Pubblicazione: (2025)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
di: Li, Bin, et al.
Pubblicazione: (2025)
di: Li, Bin, et al.
Pubblicazione: (2025)
DAMRO: Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination
di: Gong, Xuan, et al.
Pubblicazione: (2024)
di: Gong, Xuan, et al.
Pubblicazione: (2024)
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors
di: Ren, Lingfeng, et al.
Pubblicazione: (2026)
di: Ren, Lingfeng, et al.
Pubblicazione: (2026)
Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models
di: Lu, Jiaying, et al.
Pubblicazione: (2023)
di: Lu, Jiaying, et al.
Pubblicazione: (2023)
Which Way Does Time Flow? A Psychophysics-Grounded Evaluation for Vision-Language Models
di: Matta, Shiho, et al.
Pubblicazione: (2025)
di: Matta, Shiho, et al.
Pubblicazione: (2025)
A Survey on Hallucination in Large Vision-Language Models
di: Liu, Hanchao, et al.
Pubblicazione: (2024)
di: Liu, Hanchao, et al.
Pubblicazione: (2024)
Mitigating Multilingual Hallucination in Large Vision-Language Models
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
di: Moratelli, Nicholas, et al.
Pubblicazione: (2026)
di: Moratelli, Nicholas, et al.
Pubblicazione: (2026)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
di: Ye, Zekai, et al.
Pubblicazione: (2025)
di: Ye, Zekai, et al.
Pubblicazione: (2025)
PAINT: Paying Attention to INformed Tokens to Mitigate Hallucination in Large Vision-Language Model
di: Arif, Kazi Hasan Ibn, et al.
Pubblicazione: (2025)
di: Arif, Kazi Hasan Ibn, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
di: Wang, Chao, et al.
Pubblicazione: (2025)
di: Wang, Chao, et al.
Pubblicazione: (2025)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
di: Zhang, Ce, et al.
Pubblicazione: (2025)
di: Zhang, Ce, et al.
Pubblicazione: (2025)
ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models
di: Wan, Zifu, et al.
Pubblicazione: (2025)
di: Wan, Zifu, et al.
Pubblicazione: (2025)
Mask What Matters: Mitigating Object Hallucinations in Multimodal Large Language Models with Object-Aligned Visual Contrastive Decoding
di: Chen, Boqi, et al.
Pubblicazione: (2026)
di: Chen, Boqi, et al.
Pubblicazione: (2026)
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
di: Wu, Junfei, et al.
Pubblicazione: (2024)
di: Wu, Junfei, et al.
Pubblicazione: (2024)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
di: Wang, Zihu, et al.
Pubblicazione: (2025)
di: Wang, Zihu, et al.
Pubblicazione: (2025)
Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models
di: Han, Zongbo, et al.
Pubblicazione: (2024)
di: Han, Zongbo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
di: Geigle, Gregor, et al.
Pubblicazione: (2024) -
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
di: Geigle, Gregor, et al.
Pubblicazione: (2023) -
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
di: Geigle, Gregor, et al.
Pubblicazione: (2023) -
Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model
di: Geigle, Gregor, et al.
Pubblicazione: (2025) -
InstructIR: High-Quality Image Restoration Following Human Instructions
di: Conde, Marcos V., et al.
Pubblicazione: (2024)