DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, JiYang, Chen, Jiawei, Xiao, Mengqi, Cheng, Yu, Li, Yangfu, Yin, Zhaoxia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tex3D: Objects as Attack Surfaces via Adversarial 3D Textures for Vision-Language-Action Models
von: Chen, Jiawei, et al.
Veröffentlicht: (2026)
von: Chen, Jiawei, et al.
Veröffentlicht: (2026)
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
von: Shang, Yuying, et al.
Veröffentlicht: (2024)
von: Shang, Yuying, et al.
Veröffentlicht: (2024)
DeepScan: A Training-Free Framework for Visually Grounded Reasoning in Large Vision-Language Models
von: Li, Yangfu, et al.
Veröffentlicht: (2026)
von: Li, Yangfu, et al.
Veröffentlicht: (2026)
Who Can See Through You? Adversarial Shielding Against VLM-Based Attribute Inference Attacks
von: Fan, Yucheng, et al.
Veröffentlicht: (2025)
von: Fan, Yucheng, et al.
Veröffentlicht: (2025)
FaceCat: Enhancing Face Recognition Security with a Unified Diffusion Model
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
Detecting and Evaluating Medical Hallucinations in Large Vision Language Models
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
Object Hallucination-Free Reinforcement Unlearning for Vision-Language Models
von: Jia, Kaidi, et al.
Veröffentlicht: (2026)
von: Jia, Kaidi, et al.
Veröffentlicht: (2026)
OrdinalBench: A Benchmark Dataset for Diagnosing Generalization Limits in Ordinal Number Understanding of Vision-Language Models
von: Tozaki, Yusuke, et al.
Veröffentlicht: (2026)
von: Tozaki, Yusuke, et al.
Veröffentlicht: (2026)
Dual-Pathway Circuits of Object Hallucination in Vision-Language Models
von: Liu, Jiaxin, et al.
Veröffentlicht: (2026)
von: Liu, Jiaxin, et al.
Veröffentlicht: (2026)
Multi-Object Hallucination in Vision-Language Models
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models
von: Lovenia, Holy, et al.
Veröffentlicht: (2023)
von: Lovenia, Holy, et al.
Veröffentlicht: (2023)
Investigating and Mitigating Object Hallucinations in Pretrained Vision-Language (CLIP) Models
von: Liu, Yufang, et al.
Veröffentlicht: (2024)
von: Liu, Yufang, et al.
Veröffentlicht: (2024)
MathSight: A Benchmark Exploring Have Vision-Language Models Really Seen in University-Level Mathematical Reasoning?
von: Wang, Yuandong, et al.
Veröffentlicht: (2025)
von: Wang, Yuandong, et al.
Veröffentlicht: (2025)
Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection
von: Yang, Le, et al.
Veröffentlicht: (2024)
von: Yang, Le, et al.
Veröffentlicht: (2024)
AgroBench: Vision-Language Model Benchmark in Agriculture
von: Shinoda, Risa, et al.
Veröffentlicht: (2025)
von: Shinoda, Risa, et al.
Veröffentlicht: (2025)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
von: Ma, Ruiqi, et al.
Veröffentlicht: (2025)
von: Ma, Ruiqi, et al.
Veröffentlicht: (2025)
VORD: Visual Ordinal Calibration for Mitigating Object Hallucinations in Large Vision-Language Models
von: Neo, Dexter, et al.
Veröffentlicht: (2024)
von: Neo, Dexter, et al.
Veröffentlicht: (2024)
CDH-Bench: A Commonsense-Driven Hallucination Benchmark for Evaluating Visual Fidelity in Vision-Language Models
von: Chen, Kesheng, et al.
Veröffentlicht: (2026)
von: Chen, Kesheng, et al.
Veröffentlicht: (2026)
AutoBreach: Universal and Adaptive Jailbreaking with Efficient Wordplay-Guided Optimization
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
Evaluating and Analyzing Relationship Hallucinations in Large Vision-Language Models
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models
von: Chen, Junzhe, et al.
Veröffentlicht: (2026)
von: Chen, Junzhe, et al.
Veröffentlicht: (2026)
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)
MIHBench: Benchmarking and Mitigating Multi-Image Hallucinations in Multimodal Large Language Models
von: Li, Jiale, et al.
Veröffentlicht: (2025)
von: Li, Jiale, et al.
Veröffentlicht: (2025)
THRONE: An Object-based Hallucination Benchmark for the Free-form Generations of Large Vision-Language Models
von: Kaul, Prannay, et al.
Veröffentlicht: (2024)
von: Kaul, Prannay, et al.
Veröffentlicht: (2024)
EagleVision: Object-level Attribute Multimodal LLM for Remote Sensing
von: Jiang, Hongxiang, et al.
Veröffentlicht: (2025)
von: Jiang, Hongxiang, et al.
Veröffentlicht: (2025)
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
von: Wang, Zhaowei, et al.
Veröffentlicht: (2025)
von: Wang, Zhaowei, et al.
Veröffentlicht: (2025)
ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation
von: Wu, Yuhan, et al.
Veröffentlicht: (2025)
von: Wu, Yuhan, et al.
Veröffentlicht: (2025)
4D-Bench: Benchmarking Multi-modal Large Language Models for 4D Object Understanding
von: Zhu, Wenxuan, et al.
Veröffentlicht: (2025)
von: Zhu, Wenxuan, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models via Causal Route Gating
von: Cheng, Zhe, et al.
Veröffentlicht: (2026)
von: Cheng, Zhe, et al.
Veröffentlicht: (2026)
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
von: Jiang, Xin, et al.
Veröffentlicht: (2024)
von: Jiang, Xin, et al.
Veröffentlicht: (2024)
Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens
von: Jiang, Zhangqi, et al.
Veröffentlicht: (2024)
von: Jiang, Zhangqi, et al.
Veröffentlicht: (2024)
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
von: Li, Qiming, et al.
Veröffentlicht: (2025)
von: Li, Qiming, et al.
Veröffentlicht: (2025)
HALP: Detecting Hallucinations in Vision-Language Models without Generating a Single Token
von: Kogilathota, Sai Akhil, et al.
Veröffentlicht: (2026)
von: Kogilathota, Sai Akhil, et al.
Veröffentlicht: (2026)
Med-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models
von: Nguyen, Minh Khoi, et al.
Veröffentlicht: (2026)
von: Nguyen, Minh Khoi, et al.
Veröffentlicht: (2026)
EventHallusion: Diagnosing Event Hallucinations in Video LLMs
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models
von: Jing, Liqiang, et al.
Veröffentlicht: (2025)
von: Jing, Liqiang, et al.
Veröffentlicht: (2025)
Detecting and Preventing Hallucinations in Large Vision Language Models
von: Gunjal, Anisha, et al.
Veröffentlicht: (2023)
von: Gunjal, Anisha, et al.
Veröffentlicht: (2023)
What Makes "Good" Distractors for Object Hallucination Evaluation in Large Vision-Language Models?
von: Xie, Ming-Kun, et al.
Veröffentlicht: (2025)
von: Xie, Ming-Kun, et al.
Veröffentlicht: (2025)
Heron-Bench: A Benchmark for Evaluating Vision Language Models in Japanese
von: Inoue, Yuichi, et al.
Veröffentlicht: (2024)
von: Inoue, Yuichi, et al.
Veröffentlicht: (2024)
Alleviating Hallucinations in Large Vision-Language Models through Hallucination-Induced Optimization
von: Lyu, Xinyu, et al.
Veröffentlicht: (2024)
von: Lyu, Xinyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Tex3D: Objects as Attack Surfaces via Adversarial 3D Textures for Vision-Language-Action Models
von: Chen, Jiawei, et al.
Veröffentlicht: (2026) -
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
von: Shang, Yuying, et al.
Veröffentlicht: (2024) -
DeepScan: A Training-Free Framework for Visually Grounded Reasoning in Large Vision-Language Models
von: Li, Yangfu, et al.
Veröffentlicht: (2026) -
Who Can See Through You? Adversarial Shielding Against VLM-Based Attribute Inference Attacks
von: Fan, Yucheng, et al.
Veröffentlicht: (2025) -
FaceCat: Enhancing Face Recognition Security with a Unified Diffusion Model
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)