What Makes "Good" Distractors for Object Hallucination Evaluation in Large Vision-Language Models?
Fuente:
arXiv
Saved in:
| Main Authors: | Xie, Ming-Kun, Xiao, Jia-Hao, Niu, Gang, Feng, Lei, Kou, Zhiqiang, Zhang, Min-Ling, Sugiyama, Masashi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are Multimodal Large Language Models Good Annotators for Image Tagging?
by: Xie, Ming-Kun, et al.
Published: (2026)
by: Xie, Ming-Kun, et al.
Published: (2026)
Rethinking Toxicity Evaluation in Large Language Models: A Multi-Label Perspective
by: Kou, Zhiqiang, et al.
Published: (2025)
by: Kou, Zhiqiang, et al.
Published: (2025)
Counterfactual Reasoning for Multi-Label Image Classification via Patching-Based Training
by: Xie, Ming-Kun, et al.
Published: (2024)
by: Xie, Ming-Kun, et al.
Published: (2024)
Dual-Decoupling Learning and Metric-Adaptive Thresholding for Semi-Supervised Multi-Label Learning
by: Xiao, Jia-Hao, et al.
Published: (2024)
by: Xiao, Jia-Hao, et al.
Published: (2024)
Rethinking Consistent Multi-Label Classification Under Inexact Supervision
by: Wang, Wei, et al.
Published: (2025)
by: Wang, Wei, et al.
Published: (2025)
Multi-Label Knowledge Distillation
by: Yang, Penghui, et al.
Published: (2023)
by: Yang, Penghui, et al.
Published: (2023)
Realistic Evaluation of Deep Partial-Label Learning Algorithms
by: Wang, Wei, et al.
Published: (2025)
by: Wang, Wei, et al.
Published: (2025)
What Is Preference Optimization Doing, and Why?
by: Wang, Yue, et al.
Published: (2025)
by: Wang, Yue, et al.
Published: (2025)
Accessible, Realistic, and Fair Evaluation of Positive-Unlabeled Learning Algorithms
by: Wang, Wei, et al.
Published: (2025)
by: Wang, Wei, et al.
Published: (2025)
Decomposing the Basic Abilities of Large Language Models: Mitigating Cross-Task Interference in Multi-Task Instruct-Tuning
by: Wang, Bing, et al.
Published: (2026)
by: Wang, Bing, et al.
Published: (2026)
Label Distribution Learning with Biased Annotations by Learning Multi-Label Representation
by: Kou, Zhiqiang, et al.
Published: (2025)
by: Kou, Zhiqiang, et al.
Published: (2025)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
by: Xu, Jie, et al.
Published: (2025)
by: Xu, Jie, et al.
Published: (2025)
Towards Scalable Oversight with Collaborative Multi-Agent Debate in Error Detection
by: Chen, Yongqiang, et al.
Published: (2025)
by: Chen, Yongqiang, et al.
Published: (2025)
On the Overlooked Pitfalls of Weight Decay and How to Mitigate Them: A Gradient-Norm Perspective
by: Xie, Zeke, et al.
Published: (2020)
by: Xie, Zeke, et al.
Published: (2020)
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
by: Xie, Yuxi, et al.
Published: (2024)
by: Xie, Yuxi, et al.
Published: (2024)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
by: Lee, Jihoon, et al.
Published: (2025)
by: Lee, Jihoon, et al.
Published: (2025)
Positive-Unlabeled Reinforcement Learning Distillation for On-Premise Small Models
by: Kou, Zhiqiang, et al.
Published: (2026)
by: Kou, Zhiqiang, et al.
Published: (2026)
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
by: Cai, Xin-Qiang, et al.
Published: (2025)
by: Cai, Xin-Qiang, et al.
Published: (2025)
Generating Chain-of-Thoughts with a Pairwise-Comparison Approach to Searching for the Most Promising Intermediate Thought
by: Zhang, Zhen-Yu, et al.
Published: (2024)
by: Zhang, Zhen-Yu, et al.
Published: (2024)
Learning with Complementary Labels Revisited: The Selected-Completely-at-Random Setting Is More Practical
by: Wang, Wei, et al.
Published: (2023)
by: Wang, Wei, et al.
Published: (2023)
In-context Demonstration Matters: On Prompt Optimization for Pseudo-Supervision Refinement
by: Zhang, Zhen-Yu, et al.
Published: (2024)
by: Zhang, Zhen-Yu, et al.
Published: (2024)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
by: Chen, Junzhe, et al.
Published: (2024)
by: Chen, Junzhe, et al.
Published: (2024)
What Makes Good Few-shot Examples for Vision-Language Models?
by: Guo, Zhaojun, et al.
Published: (2024)
by: Guo, Zhaojun, et al.
Published: (2024)
Reference Service: What Makes It Good? What Makes It Ethical?
by: Palmer, Suzy Szasz
Published: (1999)
by: Palmer, Suzy Szasz
Published: (1999)
Weak-to-Strong Diffusion with Reflection
by: Bai, Lichen, et al.
Published: (2025)
by: Bai, Lichen, et al.
Published: (2025)
Evaluating Hallucination in Large Vision-Language Models based on Context-Aware Object Similarities
by: Datta, Shounak, et al.
Published: (2025)
by: Datta, Shounak, et al.
Published: (2025)
What Makes a Good Natural Language Prompt?
by: Long, Do Xuan, et al.
Published: (2025)
by: Long, Do Xuan, et al.
Published: (2025)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
by: Li, Qiming, et al.
Published: (2025)
by: Li, Qiming, et al.
Published: (2025)
Decoupling the Class Label and the Target Concept in Machine Unlearning
by: Zhu, Jianing, et al.
Published: (2024)
by: Zhu, Jianing, et al.
Published: (2024)
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
by: Huang, Zhuo, et al.
Published: (2026)
by: Huang, Zhuo, et al.
Published: (2026)
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
by: Shang, Yuying, et al.
Published: (2024)
by: Shang, Yuying, et al.
Published: (2024)
What Makes a Good AI Review? Concern-Level Diagnostics for AI Peer Review
by: Jin, Ming
Published: (2026)
by: Jin, Ming
Published: (2026)
What Makes for Good Image Captions?
by: Chen, Delong, et al.
Published: (2024)
by: Chen, Delong, et al.
Published: (2024)
What Makes a Good Book?
by: Park, Chung I.
Published: (1987)
by: Park, Chung I.
Published: (1987)
ActionCodec: What Makes for Good Action Tokenizers
by: Dong, Zibin, et al.
Published: (2026)
by: Dong, Zibin, et al.
Published: (2026)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
by: Zhou, Yiyang, et al.
Published: (2023)
by: Zhou, Yiyang, et al.
Published: (2023)
Detecting and Evaluating Medical Hallucinations in Large Vision Language Models
by: Chen, Jiawei, et al.
Published: (2024)
by: Chen, Jiawei, et al.
Published: (2024)
What Makes Good Synthetic Training Data for Zero-Shot Stereo Matching?
by: Yan, David, et al.
Published: (2025)
by: Yan, David, et al.
Published: (2025)
Hal-Eval: A Universal and Fine-grained Hallucination Evaluation Framework for Large Vision Language Models
by: Jiang, Chaoya, et al.
Published: (2024)
by: Jiang, Chaoya, et al.
Published: (2024)
VEC-SBM: Optimal Community Detection with Vectorial Edges Covariates
by: Braun, Guillaume, et al.
Published: (2024)
by: Braun, Guillaume, et al.
Published: (2024)
Similar Items
-
Are Multimodal Large Language Models Good Annotators for Image Tagging?
by: Xie, Ming-Kun, et al.
Published: (2026) -
Rethinking Toxicity Evaluation in Large Language Models: A Multi-Label Perspective
by: Kou, Zhiqiang, et al.
Published: (2025) -
Counterfactual Reasoning for Multi-Label Image Classification via Patching-Based Training
by: Xie, Ming-Kun, et al.
Published: (2024) -
Dual-Decoupling Learning and Metric-Adaptive Thresholding for Semi-Supervised Multi-Label Learning
by: Xiao, Jia-Hao, et al.
Published: (2024) -
Rethinking Consistent Multi-Label Classification Under Inexact Supervision
by: Wang, Wei, et al.
Published: (2025)