EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miyazato, Ryuhei, Kitada, Shunsuke, Harada, Kei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Prediction Performance and Model Interpretability through Attention Mechanisms from Basic and Applied Research Perspectives
von: Kitada, Shunsuke
Veröffentlicht: (2023)
von: Kitada, Shunsuke
Veröffentlicht: (2023)
HalDec-Bench: Benchmarking Hallucination Detector in Image Captioning
von: Saito, Kuniaki, et al.
Veröffentlicht: (2026)
von: Saito, Kuniaki, et al.
Veröffentlicht: (2026)
HalDec-Bench: Benchmarking Hallucination Detector in Image Captioning
von: Saito, Kuniaki, et al.
Veröffentlicht: (2025)
von: Saito, Kuniaki, et al.
Veröffentlicht: (2025)
SciGA: A Comprehensive Dataset for Designing Graphical Abstracts in Academic Papers
von: Kawada, Takuro, et al.
Veröffentlicht: (2025)
von: Kawada, Takuro, et al.
Veröffentlicht: (2025)
Systematic Reward Gap Optimization for Mitigating VLM Hallucinations
von: He, Lehan, et al.
Veröffentlicht: (2024)
von: He, Lehan, et al.
Veröffentlicht: (2024)
OViP: Online Vision-Language Preference Learning for VLM Hallucination
von: Liu, Shujun, et al.
Veröffentlicht: (2025)
von: Liu, Shujun, et al.
Veröffentlicht: (2025)
ChartHal: A Fine-grained Framework Evaluating Hallucination of Large Vision Language Models in Chart Understanding
von: Wang, Xingqi, et al.
Veröffentlicht: (2025)
von: Wang, Xingqi, et al.
Veröffentlicht: (2025)
Pre-Training Multimodal Hallucination Detectors with Corrupted Grounding Data
von: Whitehead, Spencer, et al.
Veröffentlicht: (2024)
von: Whitehead, Spencer, et al.
Veröffentlicht: (2024)
From Behavioral Performance to Internal Competence: Interpreting Vision-Language Models with VLM-Lens
von: Sheta, Hala, et al.
Veröffentlicht: (2025)
von: Sheta, Hala, et al.
Veröffentlicht: (2025)
GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning
von: Ebouky, Brown, et al.
Veröffentlicht: (2026)
von: Ebouky, Brown, et al.
Veröffentlicht: (2026)
VidHal: Benchmarking Temporal Hallucinations in Vision LLMs
von: Choong, Wey Yeh, et al.
Veröffentlicht: (2024)
von: Choong, Wey Yeh, et al.
Veröffentlicht: (2024)
Mitigating Object Hallucination via Robust Local Perception Search
von: Gao, Zixian, et al.
Veröffentlicht: (2025)
von: Gao, Zixian, et al.
Veröffentlicht: (2025)
InfoDet: A Dataset for Infographic Element Detection
von: Zhu, Jiangning, et al.
Veröffentlicht: (2025)
von: Zhu, Jiangning, et al.
Veröffentlicht: (2025)
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base
von: Nguyen, Cong-Duy, et al.
Veröffentlicht: (2025)
von: Nguyen, Cong-Duy, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs
von: Nath, Sujoy, et al.
Veröffentlicht: (2025)
von: Nath, Sujoy, et al.
Veröffentlicht: (2025)
HalLoc: Token-level Localization of Hallucinations for Vision Language Models
von: Park, Eunkyu, et al.
Veröffentlicht: (2025)
von: Park, Eunkyu, et al.
Veröffentlicht: (2025)
Seeing is Believing: Rich-Context Hallucination Detection for MLLMs via Backward Visual Grounding
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
Noisy Deep Ensemble: Accelerating Deep Ensemble Learning via Noise Injection
von: Sakai, Shunsuke, et al.
Veröffentlicht: (2025)
von: Sakai, Shunsuke, et al.
Veröffentlicht: (2025)
VASCAR: Content-Aware Layout Generation via Visual-Aware Self-Correction
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
BookAsSumQA: An Evaluation Framework for Aspect-Based Book Summarization via Question Answering
von: Miyazato, Ryuhei, et al.
Veröffentlicht: (2025)
von: Miyazato, Ryuhei, et al.
Veröffentlicht: (2025)
OmDet: Large-scale vision-language multi-dataset pre-training with multimodal detection network
von: Zhao, Tiancheng, et al.
Veröffentlicht: (2022)
von: Zhao, Tiancheng, et al.
Veröffentlicht: (2022)
Mitigating Object Hallucination via Concentric Causal Attention
von: Xing, Yun, et al.
Veröffentlicht: (2024)
von: Xing, Yun, et al.
Veröffentlicht: (2024)
Praxis-VLM: Vision-Grounded Decision Making via Text-Driven Reinforcement Learning
von: Hu, Zhe, et al.
Veröffentlicht: (2025)
von: Hu, Zhe, et al.
Veröffentlicht: (2025)
Do Vision Encoders Truly Explain Object Hallucination?: Mitigating Object Hallucination via Simple Fine-Grained CLIPScore
von: Oh, Hongseok, et al.
Veröffentlicht: (2025)
von: Oh, Hongseok, et al.
Veröffentlicht: (2025)
RA-Det: Towards Universal Detection of AI-Generated Images via Robustness Asymmetry
von: Wang, Xinchang, et al.
Veröffentlicht: (2026)
von: Wang, Xinchang, et al.
Veröffentlicht: (2026)
FMG-Det: Foundation Model Guided Robust Object Detection
von: Hannan, Darryl, et al.
Veröffentlicht: (2025)
von: Hannan, Darryl, et al.
Veröffentlicht: (2025)
Investigating VLM Hallucination from a Cognitive Psychology Perspective: A First Step Toward Interpretation with Intriguing Observations
von: Liu, Xiangrui, et al.
Veröffentlicht: (2025)
von: Liu, Xiangrui, et al.
Veröffentlicht: (2025)
Plain-Det: A Plain Multi-Dataset Object Detector
von: Shi, Cheng, et al.
Veröffentlicht: (2024)
von: Shi, Cheng, et al.
Veröffentlicht: (2024)
HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025)
Mitigating Multimodal Hallucination via Phase-wise Self-reward
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
Mitigating Multimodal Hallucinations via Gradient-based Self-Reflection
von: Wang, Shan, et al.
Veröffentlicht: (2025)
von: Wang, Shan, et al.
Veröffentlicht: (2025)
LongHalQA: Long-Context Hallucination Evaluation for MultiModal Large Language Models
von: Qiu, Han, et al.
Veröffentlicht: (2024)
von: Qiu, Han, et al.
Veröffentlicht: (2024)
GLSim: Detecting Object Hallucinations in LVLMs via Global-Local Similarity
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
PersonaVLM: Long-Term Personalized Multimodal LLMs
von: Nie, Chang, et al.
Veröffentlicht: (2026)
von: Nie, Chang, et al.
Veröffentlicht: (2026)
Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2023)
Adaptive Detector-Verifier Framework for Zero-Shot Polyp Detection in Open-World Settings
von: Xu, Shengkai, et al.
Veröffentlicht: (2025)
von: Xu, Shengkai, et al.
Veröffentlicht: (2025)
Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens
von: Zheng, Haohan, et al.
Veröffentlicht: (2025)
von: Zheng, Haohan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Improving Prediction Performance and Model Interpretability through Attention Mechanisms from Basic and Applied Research Perspectives
von: Kitada, Shunsuke
Veröffentlicht: (2023) -
HalDec-Bench: Benchmarking Hallucination Detector in Image Captioning
von: Saito, Kuniaki, et al.
Veröffentlicht: (2026) -
HalDec-Bench: Benchmarking Hallucination Detector in Image Captioning
von: Saito, Kuniaki, et al.
Veröffentlicht: (2025) -
SciGA: A Comprehensive Dataset for Designing Graphical Abstracts in Academic Papers
von: Kawada, Takuro, et al.
Veröffentlicht: (2025) -
Systematic Reward Gap Optimization for Mitigating VLM Hallucinations
von: He, Lehan, et al.
Veröffentlicht: (2024)