Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
Fuente:
arXiv
Guardado en:
| Autores principales: | Agrawal, Aakriti, KV, Gouthaman, Aralikatti, Rohith, Jagatap, Gauri, Yuan, Jiaxin, Kamarshi, Vijay, Fanelli, Andrea, Huang, Furong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Moment Sampling in Video LLMs for Long-Form Video QA
por: Chasmai, Mustafa, et al.
Publicado: (2025)
por: Chasmai, Mustafa, et al.
Publicado: (2025)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
por: Agrawal, Aakriti, et al.
Publicado: (2025)
por: Agrawal, Aakriti, et al.
Publicado: (2025)
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval
por: Stewart, Shanti, et al.
Publicado: (2024)
por: Stewart, Shanti, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
por: Lu, Yifan, et al.
Publicado: (2025)
por: Lu, Yifan, et al.
Publicado: (2025)
Vision-Enhanced Large Language Models for High-Resolution Image Synthesis and Multimodal Data Interpretation
por: KV, Karthikeya
Publicado: (2025)
por: KV, Karthikeya
Publicado: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024)
por: Zhong, Weihong, et al.
Publicado: (2024)
Quantity Matters: Towards Assessing and Mitigating Number Hallucination in Large Vision-Language Models
por: Zhang, Huixuan, et al.
Publicado: (2024)
por: Zhang, Huixuan, et al.
Publicado: (2024)
Segmentation-Based Attention Entropy: Detecting and Mitigating Object Hallucinations in Large Vision-Language Models
por: Song, Jiale, et al.
Publicado: (2026)
por: Song, Jiale, et al.
Publicado: (2026)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
por: Zhou, Yiyang, et al.
Publicado: (2023)
por: Zhou, Yiyang, et al.
Publicado: (2023)
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
por: Agrawal, Aakriti, et al.
Publicado: (2025)
por: Agrawal, Aakriti, et al.
Publicado: (2025)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
por: Chang, Yue, et al.
Publicado: (2024)
por: Chang, Yue, et al.
Publicado: (2024)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
por: Wang, Weihang, et al.
Publicado: (2025)
por: Wang, Weihang, et al.
Publicado: (2025)
Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation
por: Lokegaonkar, Vaibhavi, et al.
Publicado: (2026)
por: Lokegaonkar, Vaibhavi, et al.
Publicado: (2026)
CPR: Mitigating Large Language Model Hallucinations with Curative Prompt Refinement
por: Shim, Jung-Woo, et al.
Publicado: (2025)
por: Shim, Jung-Woo, et al.
Publicado: (2025)
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
por: Shim, Jung-Woo, et al.
Publicado: (2025)
por: Shim, Jung-Woo, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
por: Wang, Xintong, et al.
Publicado: (2024)
por: Wang, Xintong, et al.
Publicado: (2024)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
por: Li, Bin, et al.
Publicado: (2025)
por: Li, Bin, et al.
Publicado: (2025)
MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models
por: Wu, Xiyang, et al.
Publicado: (2025)
por: Wu, Xiyang, et al.
Publicado: (2025)
Exploring Causes and Mitigation of Hallucinations in Large Vision Language Models
por: Sun, Yaqi, et al.
Publicado: (2025)
por: Sun, Yaqi, et al.
Publicado: (2025)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
por: Fieback, Laura, et al.
Publicado: (2025)
por: Fieback, Laura, et al.
Publicado: (2025)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
por: Fazli, Mehrdad, et al.
Publicado: (2025)
por: Fazli, Mehrdad, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
por: Min, Kyungmin, et al.
Publicado: (2024)
por: Min, Kyungmin, et al.
Publicado: (2024)
Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models
por: Lee, Yi-Lun, et al.
Publicado: (2024)
por: Lee, Yi-Lun, et al.
Publicado: (2024)
PAINT: Paying Attention to INformed Tokens to Mitigate Hallucination in Large Vision-Language Model
por: Arif, Kazi Hasan Ibn, et al.
Publicado: (2025)
por: Arif, Kazi Hasan Ibn, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
por: Zhang, Ce, et al.
Publicado: (2025)
por: Zhang, Ce, et al.
Publicado: (2025)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
por: Ma, Ruiqi, et al.
Publicado: (2025)
por: Ma, Ruiqi, et al.
Publicado: (2025)
ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models
por: Wan, Zifu, et al.
Publicado: (2025)
por: Wan, Zifu, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Adaptively Constraining Information Flow
por: Bai, Jiaqi, et al.
Publicado: (2025)
por: Bai, Jiaqi, et al.
Publicado: (2025)
Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models
por: Zhang, Jinrui, et al.
Publicado: (2024)
por: Zhang, Jinrui, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation
por: Zhu, Xingyu, et al.
Publicado: (2026)
por: Zhu, Xingyu, et al.
Publicado: (2026)
Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)
por: Manevich, Avshalom, et al.
Publicado: (2024)
por: Manevich, Avshalom, et al.
Publicado: (2024)
Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation
por: Jia, Sihang, et al.
Publicado: (2026)
por: Jia, Sihang, et al.
Publicado: (2026)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
por: An, Wenbin, et al.
Publicado: (2024)
por: An, Wenbin, et al.
Publicado: (2024)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
por: Wang, Zihu, et al.
Publicado: (2025)
por: Wang, Zihu, et al.
Publicado: (2025)
Fine-Refine: Iterative Fine-grained Refinement for Mitigating Dialogue Hallucination
por: Chen, Xiangyan, et al.
Publicado: (2026)
por: Chen, Xiangyan, et al.
Publicado: (2026)
HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models
por: Guan, Tianrui, et al.
Publicado: (2023)
por: Guan, Tianrui, et al.
Publicado: (2023)
Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models
por: Zhang, Chengsheng, et al.
Publicado: (2026)
por: Zhang, Chengsheng, et al.
Publicado: (2026)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
por: Zhang, Yuanhong, et al.
Publicado: (2026)
por: Zhang, Yuanhong, et al.
Publicado: (2026)
Ejemplares similares
-
Moment Sampling in Video LLMs for Long-Form Video QA
por: Chasmai, Mustafa, et al.
Publicado: (2025) -
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
por: Agrawal, Aakriti, et al.
Publicado: (2025) -
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval
por: Stewart, Shanti, et al.
Publicado: (2024) -
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
por: Lu, Yifan, et al.
Publicado: (2025) -
Vision-Enhanced Large Language Models for High-Resolution Image Synthesis and Multimodal Data Interpretation
por: KV, Karthikeya
Publicado: (2025)