HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nath, Sujoy, Basu, Arkaprabha, Dasgupta, Sharanya, Das, Swagatam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2025)
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2025)
ARREST: Adversarial Resilient Regulation Enhancing Safety and Truth in Large Language Models
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2026)
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2026)
Revealing the Ancient Beauty: Digital Reconstruction of Temple Tiles using Computer Vision
von: Basu, Arkaprabha
Veröffentlicht: (2025)
von: Basu, Arkaprabha
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation
von: Wang, Junyang, et al.
Veröffentlicht: (2023)
von: Wang, Junyang, et al.
Veröffentlicht: (2023)
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift
von: Xu, Qinwu
Veröffentlicht: (2026)
von: Xu, Qinwu
Veröffentlicht: (2026)
Fortifying Fully Convolutional Generative Adversarial Networks for Image Super-Resolution Using Divergence Measures
von: Basu, Arkaprabha, et al.
Veröffentlicht: (2024)
von: Basu, Arkaprabha, et al.
Veröffentlicht: (2024)
HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection
von: Yang, Le, et al.
Veröffentlicht: (2024)
von: Yang, Le, et al.
Veröffentlicht: (2024)
HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding
von: Yuan, Fan, et al.
Veröffentlicht: (2024)
von: Yuan, Fan, et al.
Veröffentlicht: (2024)
Seeing is Believing: Rich-Context Hallucination Detection for MLLMs via Backward Visual Grounding
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
HiDrop: Hierarchical Vision Token Reduction in MLLMs via Late Injection, Concave Pyramid Pruning, and Early Exit
von: Wu, Hao, et al.
Veröffentlicht: (2026)
von: Wu, Hao, et al.
Veröffentlicht: (2026)
Analyzing Finetuning Representation Shift for Multimodal LLMs Steering
von: Khayatan, Pegah, et al.
Veröffentlicht: (2025)
von: Khayatan, Pegah, et al.
Veröffentlicht: (2025)
BridgeTower: Building Bridges Between Encoders in Vision-Language Representation Learning
von: Xu, Xiao, et al.
Veröffentlicht: (2022)
von: Xu, Xiao, et al.
Veröffentlicht: (2022)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
von: Wang, Weihang, et al.
Veröffentlicht: (2025)
von: Wang, Weihang, et al.
Veröffentlicht: (2025)
Mitigating Object Hallucinations in MLLMs via Multi-Frequency Perturbations
von: Li, Shuo, et al.
Veröffentlicht: (2025)
von: Li, Shuo, et al.
Veröffentlicht: (2025)
Exploring the Design Space of Visual Context Representation in Video MLLMs
von: Du, Yifan, et al.
Veröffentlicht: (2024)
von: Du, Yifan, et al.
Veröffentlicht: (2024)
Refining Skewed Perceptions in Vision-Language Contrastive Models through Visual Representations
von: Dai, Haocheng, et al.
Veröffentlicht: (2024)
von: Dai, Haocheng, et al.
Veröffentlicht: (2024)
OViP: Online Vision-Language Preference Learning for VLM Hallucination
von: Liu, Shujun, et al.
Veröffentlicht: (2025)
von: Liu, Shujun, et al.
Veröffentlicht: (2025)
AutoHallusion: Automatic Generation of Hallucination Benchmarks for Vision-Language Models
von: Wu, Xiyang, et al.
Veröffentlicht: (2024)
von: Wu, Xiyang, et al.
Veröffentlicht: (2024)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
von: Chang, Yue, et al.
Veröffentlicht: (2024)
von: Chang, Yue, et al.
Veröffentlicht: (2024)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
von: Kim, Minchan, et al.
Veröffentlicht: (2024)
von: Kim, Minchan, et al.
Veröffentlicht: (2024)
MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs
von: Gao, Yufei, et al.
Veröffentlicht: (2025)
von: Gao, Yufei, et al.
Veröffentlicht: (2025)
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2025)
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2025)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models
von: Jing, Liqiang, et al.
Veröffentlicht: (2025)
von: Jing, Liqiang, et al.
Veröffentlicht: (2025)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
von: Li, Bin, et al.
Veröffentlicht: (2025)
von: Li, Bin, et al.
Veröffentlicht: (2025)
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
von: Shang, Yuying, et al.
Veröffentlicht: (2024)
von: Shang, Yuying, et al.
Veröffentlicht: (2024)
Can Vision-Language Models Evaluate Handwritten Math?
von: Nath, Oikantik, et al.
Veröffentlicht: (2025)
von: Nath, Oikantik, et al.
Veröffentlicht: (2025)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
von: Wang, Zihu, et al.
Veröffentlicht: (2025)
von: Wang, Zihu, et al.
Veröffentlicht: (2025)
MLLMs-Augmented Visual-Language Representation Learning
von: Liu, Yanqing, et al.
Veröffentlicht: (2023)
von: Liu, Yanqing, et al.
Veröffentlicht: (2023)
Multi-Object Hallucination in Vision-Language Models
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
PromptSync: Bridging Domain Gaps in Vision-Language Models through Class-Aware Prototype Alignment and Discrimination
von: Khandelwal, Anant
Veröffentlicht: (2024)
von: Khandelwal, Anant
Veröffentlicht: (2024)
PAINT: Paying Attention to INformed Tokens to Mitigate Hallucination in Large Vision-Language Model
von: Arif, Kazi Hasan Ibn, et al.
Veröffentlicht: (2025)
von: Arif, Kazi Hasan Ibn, et al.
Veröffentlicht: (2025)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
von: Ma, Ruiqi, et al.
Veröffentlicht: (2025)
von: Ma, Ruiqi, et al.
Veröffentlicht: (2025)
Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models
von: Lovenia, Holy, et al.
Veröffentlicht: (2023)
von: Lovenia, Holy, et al.
Veröffentlicht: (2023)
ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models
von: Wan, Zifu, et al.
Veröffentlicht: (2025)
von: Wan, Zifu, et al.
Veröffentlicht: (2025)
Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2025) -
ARREST: Adversarial Resilient Regulation Enhancing Safety and Truth in Large Language Models
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2026) -
Revealing the Ancient Beauty: Digital Reconstruction of Temple Tiles using Computer Vision
von: Basu, Arkaprabha
Veröffentlicht: (2025) -
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
von: Wang, Chao, et al.
Veröffentlicht: (2025) -
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
von: Lu, Yifan, et al.
Veröffentlicht: (2025)