When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise
Fuente:
arXiv
Guardado en:
| Autores principales: | Shin, Philip Wootaek, Sridhar, Ajay Narayanan, Devarapalli, Sivani, Zhang, Rui, Sampson, Jack, Narayanan, Vijaykrishnan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can Prompt Modifiers Control Bias? A Comparative Analysis of Text-to-Image Generative Models
por: Shin, Philip Wootaek, et al.
Publicado: (2024)
por: Shin, Philip Wootaek, et al.
Publicado: (2024)
Losing the Plot: How VLM responses degrade on imperfect charts
por: Shin, Philip Wootaek, et al.
Publicado: (2025)
por: Shin, Philip Wootaek, et al.
Publicado: (2025)
Disharmony: Forensics using Reverse Lighting Harmonization
por: Shin, Philip Wootaek, et al.
Publicado: (2025)
por: Shin, Philip Wootaek, et al.
Publicado: (2025)
Towards High-Resolution Alignment and Super-Resolution of Multi-Sensor Satellite Imagery
por: Shin, Philip Wootaek, et al.
Publicado: (2025)
por: Shin, Philip Wootaek, et al.
Publicado: (2025)
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
por: Zhao, Shu, et al.
Publicado: (2025)
por: Zhao, Shu, et al.
Publicado: (2025)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
por: Zhou, Yiyang, et al.
Publicado: (2023)
por: Zhou, Yiyang, et al.
Publicado: (2023)
Evaluating Large Language Models on Rare Disease Diagnosis: A Case Study using House M.D
por: Gupta, Arsh, et al.
Publicado: (2025)
por: Gupta, Arsh, et al.
Publicado: (2025)
Parts-Mamba: Augmenting Joint Context with Part-Level Scanning for Occluded Human Skeleton
por: Shen, Tianyi, et al.
Publicado: (2025)
por: Shen, Tianyi, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
por: Lu, Yifan, et al.
Publicado: (2025)
por: Lu, Yifan, et al.
Publicado: (2025)
TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning
por: Pan, Muyu, et al.
Publicado: (2026)
por: Pan, Muyu, et al.
Publicado: (2026)
MoRA: Missing Modality Low-Rank Adaptation for Visual Recognition
por: Zhao, Shu, et al.
Publicado: (2025)
por: Zhao, Shu, et al.
Publicado: (2025)
KALAHash: Knowledge-Anchored Low-Resource Adaptation for Deep Hashing
por: Zhao, Shu, et al.
Publicado: (2024)
por: Zhao, Shu, et al.
Publicado: (2024)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
por: Wang, Weihang, et al.
Publicado: (2025)
por: Wang, Weihang, et al.
Publicado: (2025)
Optimal Multi-bit Generative Watermarking Schemes Under Worst-Case False-Alarm Constraints
por: Huang, Yu-Shin, et al.
Publicado: (2026)
por: Huang, Yu-Shin, et al.
Publicado: (2026)
Analyzing and Mitigating Object Hallucination: A Training Bias Perspective
por: Li, Yifan, et al.
Publicado: (2025)
por: Li, Yifan, et al.
Publicado: (2025)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
por: Khayatan, Pegah, et al.
Publicado: (2026)
por: Khayatan, Pegah, et al.
Publicado: (2026)
Multi-Object Hallucination in Vision-Language Models
por: Chen, Xuweiyi, et al.
Publicado: (2024)
por: Chen, Xuweiyi, et al.
Publicado: (2024)
OViP: Online Vision-Language Preference Learning for VLM Hallucination
por: Liu, Shujun, et al.
Publicado: (2025)
por: Liu, Shujun, et al.
Publicado: (2025)
AutoHallusion: Automatic Generation of Hallucination Benchmarks for Vision-Language Models
por: Wu, Xiyang, et al.
Publicado: (2024)
por: Wu, Xiyang, et al.
Publicado: (2024)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
por: Chang, Yue, et al.
Publicado: (2024)
por: Chang, Yue, et al.
Publicado: (2024)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
por: Kim, Minchan, et al.
Publicado: (2024)
por: Kim, Minchan, et al.
Publicado: (2024)
Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
por: Zheng, Kening, et al.
Publicado: (2024)
por: Zheng, Kening, et al.
Publicado: (2024)
ODE: Open-Set Evaluation of Hallucinations in Multimodal Large Language Models
por: Tu, Yahan, et al.
Publicado: (2024)
por: Tu, Yahan, et al.
Publicado: (2024)
Mitigating Hallucinations in Multimodal Spatial Relations through Constraint-Aware Prompting
por: Wu, Jiarui, et al.
Publicado: (2025)
por: Wu, Jiarui, et al.
Publicado: (2025)
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
por: Agrawal, Aakriti, et al.
Publicado: (2025)
por: Agrawal, Aakriti, et al.
Publicado: (2025)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
por: Geigle, Gregor, et al.
Publicado: (2024)
por: Geigle, Gregor, et al.
Publicado: (2024)
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models
por: Jing, Liqiang, et al.
Publicado: (2025)
por: Jing, Liqiang, et al.
Publicado: (2025)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
por: Li, Bin, et al.
Publicado: (2025)
por: Li, Bin, et al.
Publicado: (2025)
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
por: Shang, Yuying, et al.
Publicado: (2024)
por: Shang, Yuying, et al.
Publicado: (2024)
A Survey on Hallucination in Large Vision-Language Models
por: Liu, Hanchao, et al.
Publicado: (2024)
por: Liu, Hanchao, et al.
Publicado: (2024)
Mitigating Multilingual Hallucination in Large Vision-Language Models
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
por: Moratelli, Nicholas, et al.
Publicado: (2026)
por: Moratelli, Nicholas, et al.
Publicado: (2026)
Prioritizing Image-Related Tokens Enhances Vision-Language Pre-Training
por: Chen, Yangyi, et al.
Publicado: (2025)
por: Chen, Yangyi, et al.
Publicado: (2025)
NaviSense: A Multimodal Assistive Mobile application for Object Retrieval by Persons with Visual Impairment
por: Sridhar, Ajay Narayanan, et al.
Publicado: (2025)
por: Sridhar, Ajay Narayanan, et al.
Publicado: (2025)
Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens
por: Zheng, Haohan, et al.
Publicado: (2025)
por: Zheng, Haohan, et al.
Publicado: (2025)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
por: Wang, Zihu, et al.
Publicado: (2025)
por: Wang, Zihu, et al.
Publicado: (2025)
PAINT: Paying Attention to INformed Tokens to Mitigate Hallucination in Large Vision-Language Model
por: Arif, Kazi Hasan Ibn, et al.
Publicado: (2025)
por: Arif, Kazi Hasan Ibn, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
por: Zhang, Ce, et al.
Publicado: (2025)
por: Zhang, Ce, et al.
Publicado: (2025)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
por: Ma, Ruiqi, et al.
Publicado: (2025)
por: Ma, Ruiqi, et al.
Publicado: (2025)
Ejemplares similares
-
Can Prompt Modifiers Control Bias? A Comparative Analysis of Text-to-Image Generative Models
por: Shin, Philip Wootaek, et al.
Publicado: (2024) -
Losing the Plot: How VLM responses degrade on imperfect charts
por: Shin, Philip Wootaek, et al.
Publicado: (2025) -
Disharmony: Forensics using Reverse Lighting Harmonization
por: Shin, Philip Wootaek, et al.
Publicado: (2025) -
Towards High-Resolution Alignment and Super-Resolution of Multi-Sensor Satellite Imagery
por: Shin, Philip Wootaek, et al.
Publicado: (2025) -
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
por: Zhao, Shu, et al.
Publicado: (2025)