Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
Fuente:
arXiv
Guardado en:
| Autores principales: | Geigle, Gregor, Timofte, Radu, Glavaš, Goran |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
por: Geigle, Gregor, et al.
Publicado: (2024)
por: Geigle, Gregor, et al.
Publicado: (2024)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
por: Geigle, Gregor, et al.
Publicado: (2023)
por: Geigle, Gregor, et al.
Publicado: (2023)
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
por: Geigle, Gregor, et al.
Publicado: (2023)
por: Geigle, Gregor, et al.
Publicado: (2023)
Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model
por: Geigle, Gregor, et al.
Publicado: (2025)
por: Geigle, Gregor, et al.
Publicado: (2025)
InstructIR: High-Quality Image Restoration Following Human Instructions
por: Conde, Marcos V., et al.
Publicado: (2024)
por: Conde, Marcos V., et al.
Publicado: (2024)
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
por: Shang, Yuying, et al.
Publicado: (2024)
por: Shang, Yuying, et al.
Publicado: (2024)
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models
por: Jing, Liqiang, et al.
Publicado: (2025)
por: Jing, Liqiang, et al.
Publicado: (2025)
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
por: Ha, Jiwoo, et al.
Publicado: (2026)
por: Ha, Jiwoo, et al.
Publicado: (2026)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
por: Ma, Ruiqi, et al.
Publicado: (2025)
por: Ma, Ruiqi, et al.
Publicado: (2025)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
por: Zhou, Yiyang, et al.
Publicado: (2023)
por: Zhou, Yiyang, et al.
Publicado: (2023)
Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded Guidance
por: Zhao, Linxi, et al.
Publicado: (2024)
por: Zhao, Linxi, et al.
Publicado: (2024)
Grounding Language with Vision: A Conditional Mutual Information Calibrated Decoding Strategy for Reducing Hallucinations in LVLMs
por: Fang, Hao, et al.
Publicado: (2025)
por: Fang, Hao, et al.
Publicado: (2025)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
por: Chen, Junzhe, et al.
Publicado: (2024)
por: Chen, Junzhe, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
por: Lu, Yifan, et al.
Publicado: (2025)
por: Lu, Yifan, et al.
Publicado: (2025)
Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models
por: Lovenia, Holy, et al.
Publicado: (2023)
por: Lovenia, Holy, et al.
Publicado: (2023)
Multi-Object Hallucination in Vision-Language Models
por: Chen, Xuweiyi, et al.
Publicado: (2024)
por: Chen, Xuweiyi, et al.
Publicado: (2024)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
por: Seo, Hoigi, et al.
Publicado: (2025)
por: Seo, Hoigi, et al.
Publicado: (2025)
Do Vision-Language Models Really Understand Visual Language?
por: Hou, Yifan, et al.
Publicado: (2024)
por: Hou, Yifan, et al.
Publicado: (2024)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
por: An, Wenbin, et al.
Publicado: (2024)
por: An, Wenbin, et al.
Publicado: (2024)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
por: Wang, Weihang, et al.
Publicado: (2025)
por: Wang, Weihang, et al.
Publicado: (2025)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
por: Chang, Yue, et al.
Publicado: (2024)
por: Chang, Yue, et al.
Publicado: (2024)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
por: Woo, Sangmin, et al.
Publicado: (2025)
por: Woo, Sangmin, et al.
Publicado: (2025)
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
por: Agrawal, Aakriti, et al.
Publicado: (2025)
por: Agrawal, Aakriti, et al.
Publicado: (2025)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
por: Li, Bin, et al.
Publicado: (2025)
por: Li, Bin, et al.
Publicado: (2025)
DAMRO: Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination
por: Gong, Xuan, et al.
Publicado: (2024)
por: Gong, Xuan, et al.
Publicado: (2024)
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors
por: Ren, Lingfeng, et al.
Publicado: (2026)
por: Ren, Lingfeng, et al.
Publicado: (2026)
Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models
por: Lu, Jiaying, et al.
Publicado: (2023)
por: Lu, Jiaying, et al.
Publicado: (2023)
Which Way Does Time Flow? A Psychophysics-Grounded Evaluation for Vision-Language Models
por: Matta, Shiho, et al.
Publicado: (2025)
por: Matta, Shiho, et al.
Publicado: (2025)
A Survey on Hallucination in Large Vision-Language Models
por: Liu, Hanchao, et al.
Publicado: (2024)
por: Liu, Hanchao, et al.
Publicado: (2024)
Mitigating Multilingual Hallucination in Large Vision-Language Models
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
por: Moratelli, Nicholas, et al.
Publicado: (2026)
por: Moratelli, Nicholas, et al.
Publicado: (2026)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
por: Ye, Zekai, et al.
Publicado: (2025)
por: Ye, Zekai, et al.
Publicado: (2025)
PAINT: Paying Attention to INformed Tokens to Mitigate Hallucination in Large Vision-Language Model
por: Arif, Kazi Hasan Ibn, et al.
Publicado: (2025)
por: Arif, Kazi Hasan Ibn, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
por: Zhang, Ce, et al.
Publicado: (2025)
por: Zhang, Ce, et al.
Publicado: (2025)
ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models
por: Wan, Zifu, et al.
Publicado: (2025)
por: Wan, Zifu, et al.
Publicado: (2025)
Mask What Matters: Mitigating Object Hallucinations in Multimodal Large Language Models with Object-Aligned Visual Contrastive Decoding
por: Chen, Boqi, et al.
Publicado: (2026)
por: Chen, Boqi, et al.
Publicado: (2026)
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
por: Wu, Junfei, et al.
Publicado: (2024)
por: Wu, Junfei, et al.
Publicado: (2024)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
por: Wang, Zihu, et al.
Publicado: (2025)
por: Wang, Zihu, et al.
Publicado: (2025)
Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models
por: Han, Zongbo, et al.
Publicado: (2024)
por: Han, Zongbo, et al.
Publicado: (2024)
Ejemplares similares
-
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
por: Geigle, Gregor, et al.
Publicado: (2024) -
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
por: Geigle, Gregor, et al.
Publicado: (2023) -
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
por: Geigle, Gregor, et al.
Publicado: (2023) -
Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model
por: Geigle, Gregor, et al.
Publicado: (2025) -
InstructIR: High-Quality Image Restoration Following Human Instructions
por: Conde, Marcos V., et al.
Publicado: (2024)