Mitigating Hallucinations in Multimodal Spatial Relations through Constraint-Aware Prompting
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Jiarui, Liu, Zhuo, He, Hangfeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images
por: You, Liangliang, et al.
Publicado: (2025)
por: You, Liangliang, et al.
Publicado: (2025)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024)
por: Zhong, Weihong, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models via Entity-Centric Multimodal Preference Optimization
por: Wu, Jiulong, et al.
Publicado: (2025)
por: Wu, Jiulong, et al.
Publicado: (2025)
Magnifier Prompt: Tackling Multimodal Hallucination via Extremely Simple Instructions
por: Fu, Yuhan, et al.
Publicado: (2024)
por: Fu, Yuhan, et al.
Publicado: (2024)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
por: Woo, Sangmin, et al.
Publicado: (2025)
por: Woo, Sangmin, et al.
Publicado: (2025)
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
por: Fu, Yuhan, et al.
Publicado: (2024)
por: Fu, Yuhan, et al.
Publicado: (2024)
Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization
por: Compagnoni, Alberto, et al.
Publicado: (2025)
por: Compagnoni, Alberto, et al.
Publicado: (2025)
More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models
por: Liu, Chengzhi, et al.
Publicado: (2025)
por: Liu, Chengzhi, et al.
Publicado: (2025)
Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation
por: Jia, Sihang, et al.
Publicado: (2026)
por: Jia, Sihang, et al.
Publicado: (2026)
Mixture of Decoding: An Attention-Inspired Adaptive Decoding Strategy to Mitigate Hallucinations in Large Vision-Language Models
por: Chen, Xinlong, et al.
Publicado: (2025)
por: Chen, Xinlong, et al.
Publicado: (2025)
ChartCap: Mitigating Hallucination of Dense Chart Captioning
por: Lim, Junyoung, et al.
Publicado: (2025)
por: Lim, Junyoung, et al.
Publicado: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
SpatialPrompting: Keyframe-driven Zero-Shot Spatial Reasoning with Off-the-Shelf Multimodal Large Language Models
por: Taguchi, Shun, et al.
Publicado: (2025)
por: Taguchi, Shun, et al.
Publicado: (2025)
Mitigating Object Hallucinations in MLLMs via Multi-Frequency Perturbations
por: Li, Shuo, et al.
Publicado: (2025)
por: Li, Shuo, et al.
Publicado: (2025)
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
por: Wada, Yuiga, et al.
Publicado: (2025)
por: Wada, Yuiga, et al.
Publicado: (2025)
PAR: Prompt-Aware Token Reduction Method for Efficient Large Multimodal Models
por: Liu, Yingen, et al.
Publicado: (2024)
por: Liu, Yingen, et al.
Publicado: (2024)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
por: Chang, Kai-Po, et al.
Publicado: (2025)
por: Chang, Kai-Po, et al.
Publicado: (2025)
Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation
por: Hua, Zhenglin, et al.
Publicado: (2025)
por: Hua, Zhenglin, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
por: Min, Kyungmin, et al.
Publicado: (2024)
por: Min, Kyungmin, et al.
Publicado: (2024)
Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models
por: Lee, Yi-Lun, et al.
Publicado: (2024)
por: Lee, Yi-Lun, et al.
Publicado: (2024)
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models
por: Van, Minh-Hao, et al.
Publicado: (2025)
por: Van, Minh-Hao, et al.
Publicado: (2025)
Euclid: Supercharging Multimodal LLMs with Synthetic High-Fidelity Visual Descriptions
por: Zhang, Jiarui, et al.
Publicado: (2024)
por: Zhang, Jiarui, et al.
Publicado: (2024)
The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?
por: Yin, Hao, et al.
Publicado: (2025)
por: Yin, Hao, et al.
Publicado: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
por: An, Wenbin, et al.
Publicado: (2024)
por: An, Wenbin, et al.
Publicado: (2024)
What if...?: Thinking Counterfactual Keywords Helps to Mitigate Hallucination in Large Multi-modal Models
por: Kim, Junho, et al.
Publicado: (2024)
por: Kim, Junho, et al.
Publicado: (2024)
SAGE: Spuriousness-Aware Guided Prompt Exploration for Mitigating Multimodal Bias
por: Ye, Wenqian, et al.
Publicado: (2025)
por: Ye, Wenqian, et al.
Publicado: (2025)
Steering the Verifiability of Multimodal AI Hallucinations
por: Pang, Jianhong, et al.
Publicado: (2026)
por: Pang, Jianhong, et al.
Publicado: (2026)
Beyond Superficial Unlearning: Sharpness-Aware Robust Erasure of Hallucinations in Multimodal LLMs
por: Fang, Xianya, et al.
Publicado: (2026)
por: Fang, Xianya, et al.
Publicado: (2026)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
por: Yin, Jianghao, et al.
Publicado: (2026)
por: Yin, Jianghao, et al.
Publicado: (2026)
MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models
por: Yan, Qiao, et al.
Publicado: (2025)
por: Yan, Qiao, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)
por: Manevich, Avshalom, et al.
Publicado: (2024)
por: Manevich, Avshalom, et al.
Publicado: (2024)
Hierarchy-Aware Multimodal Unlearning for Medical AI
por: Wu, Fengli, et al.
Publicado: (2025)
por: Wu, Fengli, et al.
Publicado: (2025)
Focus Matters: Phase-Aware Suppression for Hallucination in Vision-Language Models
por: Kim, Sohyeon, et al.
Publicado: (2026)
por: Kim, Sohyeon, et al.
Publicado: (2026)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
por: Khayatan, Pegah, et al.
Publicado: (2026)
por: Khayatan, Pegah, et al.
Publicado: (2026)
Hallucination Benchmark in Medical Visual Question Answering
por: Wu, Jinge, et al.
Publicado: (2024)
por: Wu, Jinge, et al.
Publicado: (2024)
MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs
por: Zhang, Jiarui, et al.
Publicado: (2025)
por: Zhang, Jiarui, et al.
Publicado: (2025)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
por: Ye, Zekai, et al.
Publicado: (2025)
por: Ye, Zekai, et al.
Publicado: (2025)
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
por: Ha, Jiwoo, et al.
Publicado: (2026)
por: Ha, Jiwoo, et al.
Publicado: (2026)
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
por: Wang, Xintong, et al.
Publicado: (2024)
por: Wang, Xintong, et al.
Publicado: (2024)
Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination
por: Li, Xinzhuo, et al.
Publicado: (2025)
por: Li, Xinzhuo, et al.
Publicado: (2025)
Ejemplares similares
-
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images
por: You, Liangliang, et al.
Publicado: (2025) -
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024) -
Mitigating Hallucinations in Large Vision-Language Models via Entity-Centric Multimodal Preference Optimization
por: Wu, Jiulong, et al.
Publicado: (2025) -
Magnifier Prompt: Tackling Multimodal Hallucination via Extremely Simple Instructions
por: Fu, Yuhan, et al.
Publicado: (2024) -
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
por: Woo, Sangmin, et al.
Publicado: (2025)