Toward More Reliable Artificial Intelligence: Reducing Hallucinations in Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sanogo, Kassoum, Ardiccioni, Renzo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models
von: Han, Zongbo, et al.
Veröffentlicht: (2024)
von: Han, Zongbo, et al.
Veröffentlicht: (2024)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift
von: Xu, Qinwu
Veröffentlicht: (2026)
von: Xu, Qinwu
Veröffentlicht: (2026)
Unified Triplet-Level Hallucination Evaluation for Large Vision-Language Models
von: Wu, Junjie, et al.
Veröffentlicht: (2024)
von: Wu, Junjie, et al.
Veröffentlicht: (2024)
SegSub: Evaluating Robustness to Knowledge Conflicts and Hallucinations in Vision-Language Models
von: Carragher, Peter, et al.
Veröffentlicht: (2025)
von: Carragher, Peter, et al.
Veröffentlicht: (2025)
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
von: Wu, Junfei, et al.
Veröffentlicht: (2024)
von: Wu, Junfei, et al.
Veröffentlicht: (2024)
Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded Guidance
von: Zhao, Linxi, et al.
Veröffentlicht: (2024)
von: Zhao, Linxi, et al.
Veröffentlicht: (2024)
Detecting and Mitigating Hallucination in Large Vision Language Models via Fine-Grained AI Feedback
von: Xiao, Wenyi, et al.
Veröffentlicht: (2024)
von: Xiao, Wenyi, et al.
Veröffentlicht: (2024)
Cyclic Vision-Language Manipulator: Towards Reliable and Fine-Grained Image Interpretation for Automated Report Generation
von: Fang, Yingying, et al.
Veröffentlicht: (2024)
von: Fang, Yingying, et al.
Veröffentlicht: (2024)
Woodpecker: Hallucination Correction for Multimodal Large Language Models
von: Yin, Shukang, et al.
Veröffentlicht: (2023)
von: Yin, Shukang, et al.
Veröffentlicht: (2023)
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
von: Deng, Ailin, et al.
Veröffentlicht: (2024)
von: Deng, Ailin, et al.
Veröffentlicht: (2024)
RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models
von: Xia, Peng, et al.
Veröffentlicht: (2024)
von: Xia, Peng, et al.
Veröffentlicht: (2024)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
von: Khayatan, Pegah, et al.
Veröffentlicht: (2026)
von: Khayatan, Pegah, et al.
Veröffentlicht: (2026)
Reducing Hallucinations in Vision-Language Models via Latent Space Steering
von: Liu, Sheng, et al.
Veröffentlicht: (2024)
von: Liu, Sheng, et al.
Veröffentlicht: (2024)
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
Generative Artificial Intelligence: A Systematic Review and Applications
von: Sengar, Sandeep Singh, et al.
Veröffentlicht: (2024)
von: Sengar, Sandeep Singh, et al.
Veröffentlicht: (2024)
VisionZip: Longer is Better but Not Necessary in Vision Language Models
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding
von: Wu, Chang-Hsun, et al.
Veröffentlicht: (2025)
von: Wu, Chang-Hsun, et al.
Veröffentlicht: (2025)
A Novel Framework for Automated Explain Vision Model Using Vision-Language Models
von: Nguyen, Phu-Vinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Phu-Vinh, et al.
Veröffentlicht: (2025)
Imperfect Vision Encoders: Efficient and Robust Tuning for Vision-Language Models
von: Panos, Aristeidis, et al.
Veröffentlicht: (2024)
von: Panos, Aristeidis, et al.
Veröffentlicht: (2024)
Foundations of Multisensory Artificial Intelligence
von: Liang, Paul Pu
Veröffentlicht: (2024)
von: Liang, Paul Pu
Veröffentlicht: (2024)
Vision-Language Model Based Handwriting Verification
von: Chauhan, Mihir, et al.
Veröffentlicht: (2024)
von: Chauhan, Mihir, et al.
Veröffentlicht: (2024)
Differentiable Prompt Learning for Vision Language Models
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
How Culturally Aware are Vision-Language Models?
von: Burda-Lassen, Olena, et al.
Veröffentlicht: (2024)
von: Burda-Lassen, Olena, et al.
Veröffentlicht: (2024)
ALOHa: A New Measure for Hallucination in Captioning Models
von: Petryk, Suzanne, et al.
Veröffentlicht: (2024)
von: Petryk, Suzanne, et al.
Veröffentlicht: (2024)
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
Mordal: Automated Pretrained Model Selection for Vision Language Models
von: He, Shiqi, et al.
Veröffentlicht: (2025)
von: He, Shiqi, et al.
Veröffentlicht: (2025)
A Vision for Multisensory Intelligence: Sensing, Science, and Synergy
von: Liang, Paul Pu
Veröffentlicht: (2026)
von: Liang, Paul Pu
Veröffentlicht: (2026)
Are Hallucinations Bad Estimations?
von: Liu, Hude, et al.
Veröffentlicht: (2025)
von: Liu, Hude, et al.
Veröffentlicht: (2025)
Efficient Architectures for High Resolution Vision-Language Models
von: Carvalho, Miguel, et al.
Veröffentlicht: (2025)
von: Carvalho, Miguel, et al.
Veröffentlicht: (2025)
Understanding the Effects of Distractors on Reasoning Vision-Language Models
von: Bae, Jiyun, et al.
Veröffentlicht: (2025)
von: Bae, Jiyun, et al.
Veröffentlicht: (2025)
Coordinated Robustness Evaluation Framework for Vision-Language Models
von: Babu, Ashwin Ramesh, et al.
Veröffentlicht: (2025)
von: Babu, Ashwin Ramesh, et al.
Veröffentlicht: (2025)
Modeling Caption Diversity in Contrastive Vision-Language Pretraining
von: Lavoie, Samuel, et al.
Veröffentlicht: (2024)
von: Lavoie, Samuel, et al.
Veröffentlicht: (2024)
RIV: Recursive Introspection Mask Diffusion Vision Language Model
von: Li, YuQian, et al.
Veröffentlicht: (2025)
von: Li, YuQian, et al.
Veröffentlicht: (2025)
Linear Spaces of Meanings: Compositional Structures in Vision-Language Models
von: Trager, Matthew, et al.
Veröffentlicht: (2023)
von: Trager, Matthew, et al.
Veröffentlicht: (2023)
MouSi: Poly-Visual-Expert Vision-Language Models
von: Fan, Xiaoran, et al.
Veröffentlicht: (2024)
von: Fan, Xiaoran, et al.
Veröffentlicht: (2024)
Nemesis: Normalizing the Soft-prompt Vectors of Vision-Language Models
von: Fu, Shuai, et al.
Veröffentlicht: (2024)
von: Fu, Shuai, et al.
Veröffentlicht: (2024)
Evaluating Large Vision-and-Language Models on Children's Mathematical Olympiads
von: Cherian, Anoop, et al.
Veröffentlicht: (2024)
von: Cherian, Anoop, et al.
Veröffentlicht: (2024)
Unified Vision-Language Modeling via Concept Space Alignment
von: Qiu, Yifu, et al.
Veröffentlicht: (2026)
von: Qiu, Yifu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models
von: Han, Zongbo, et al.
Veröffentlicht: (2024) -
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
von: Fieback, Laura, et al.
Veröffentlicht: (2025) -
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift
von: Xu, Qinwu
Veröffentlicht: (2026) -
Unified Triplet-Level Hallucination Evaluation for Large Vision-Language Models
von: Wu, Junjie, et al.
Veröffentlicht: (2024) -
SegSub: Evaluating Robustness to Knowledge Conflicts and Hallucinations in Vision-Language Models
von: Carragher, Peter, et al.
Veröffentlicht: (2025)