Attention to details, logits to truth: visual-aware attention and logits enhancement to mitigate hallucinations in LVLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jingyi, Li, Fei, Liu, Rujie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing
von: Huang, Yanbin, et al.
Veröffentlicht: (2026)
von: Huang, Yanbin, et al.
Veröffentlicht: (2026)
Stop learning it all to mitigate visual hallucination, Focus on the hallucination target
von: Yoon, Dokyoon, et al.
Veröffentlicht: (2025)
von: Yoon, Dokyoon, et al.
Veröffentlicht: (2025)
Technical report on label-informed logit redistribution for better domain generalization in low-shot classification with foundation models
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement
von: Qin, Zhenxin, et al.
Veröffentlicht: (2026)
von: Qin, Zhenxin, et al.
Veröffentlicht: (2026)
Make LVLMs Focus: Context-Aware Attention Modulation for Better Multimodal In-Context Learning
von: Li, Yanshu, et al.
Veröffentlicht: (2025)
von: Li, Yanshu, et al.
Veröffentlicht: (2025)
Attention Hijackers: Detect and Disentangle Attention Hijacking in LVLMs for Hallucination Mitigation
von: Chen, Beitao, et al.
Veröffentlicht: (2025)
von: Chen, Beitao, et al.
Veröffentlicht: (2025)
Evidence Packing for Cross-Domain Image Deepfake Detection with LVLMs
von: Liu, Yuxin, et al.
Veröffentlicht: (2026)
von: Liu, Yuxin, et al.
Veröffentlicht: (2026)
Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence
von: He, Jinghan, et al.
Veröffentlicht: (2024)
von: He, Jinghan, et al.
Veröffentlicht: (2024)
Mitigating Object Hallucinations in LVLMs via Attention Imbalance Rectification
von: Sun, Han, et al.
Veröffentlicht: (2026)
von: Sun, Han, et al.
Veröffentlicht: (2026)
Paying More Attention to Image: A Training-Free Method for Alleviating Hallucination in LVLMs
von: Liu, Shi, et al.
Veröffentlicht: (2024)
von: Liu, Shi, et al.
Veröffentlicht: (2024)
SAVAA: Mitigating Hallucinations in LVLMs via Step-wise Adaptive Visual Attention Amplification
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2026)
Revealing and Enhancing Core Visual Regions: Harnessing Internal Attention Dynamics for Hallucination Mitigation in LVLMs
von: Lyu, Guangtao, et al.
Veröffentlicht: (2026)
von: Lyu, Guangtao, et al.
Veröffentlicht: (2026)
Scalpel: Fine-Grained Alignment of Attention Activation Manifolds via Mixture Gaussian Bridges to Mitigate Multimodal Hallucination
von: Shi, Ziqiang, et al.
Veröffentlicht: (2026)
von: Shi, Ziqiang, et al.
Veröffentlicht: (2026)
MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation
von: Lu, Shuaiye, et al.
Veröffentlicht: (2025)
von: Lu, Shuaiye, et al.
Veröffentlicht: (2025)
Vocabulary Hijacking in LVLMs: Unveiling Critical Attention Heads by Excluding Inert Tokens to Mitigate Hallucination
von: Chen, Yangneng, et al.
Veröffentlicht: (2026)
von: Chen, Yangneng, et al.
Veröffentlicht: (2026)
EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy
von: Ge, Xuanyu, et al.
Veröffentlicht: (2026)
von: Ge, Xuanyu, et al.
Veröffentlicht: (2026)
HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding
von: Yuan, Fan, et al.
Veröffentlicht: (2024)
von: Yuan, Fan, et al.
Veröffentlicht: (2024)
Causally-Grounded Dual-Path Attention Intervention for Object Hallucination Mitigation in LVLMs
von: Yu, Liu, et al.
Veröffentlicht: (2025)
von: Yu, Liu, et al.
Veröffentlicht: (2025)
MHSA: A Lightweight Framework for Mitigating Hallucinations via Steered Attention in LVLMs
von: Ding, Wei, et al.
Veröffentlicht: (2026)
von: Ding, Wei, et al.
Veröffentlicht: (2026)
Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens
von: Zheng, Haohan, et al.
Veröffentlicht: (2025)
von: Zheng, Haohan, et al.
Veröffentlicht: (2025)
Density-aware global-local attention network for point cloud segmentation
von: Li, Chade, et al.
Veröffentlicht: (2024)
von: Li, Chade, et al.
Veröffentlicht: (2024)
TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
CoMemo: LVLMs Need Image Context with Image Memory
von: Liu, Shi, et al.
Veröffentlicht: (2025)
von: Liu, Shi, et al.
Veröffentlicht: (2025)
Attention to detail: inter-resolution knowledge distillation
von: del Amor, Rocío, et al.
Veröffentlicht: (2024)
von: del Amor, Rocío, et al.
Veröffentlicht: (2024)
When RAG Hurts: Diagnosing and Mitigating Attention Distraction in Retrieval-Augmented LVLMs
von: Zhao, Beidi, et al.
Veröffentlicht: (2026)
von: Zhao, Beidi, et al.
Veröffentlicht: (2026)
Neuromorphic visual attention for Sign-language recognition on SpiNNaker
von: Liskova, Sarka, et al.
Veröffentlicht: (2026)
von: Liskova, Sarka, et al.
Veröffentlicht: (2026)
Attention-aware Social Graph Transformer Networks for Stochastic Trajectory Prediction
von: Liu, Yao, et al.
Veröffentlicht: (2023)
von: Liu, Yao, et al.
Veröffentlicht: (2023)
Knowing the Answer Isn't Enough: Fixing Reasoning Path Failures in LVLMs
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025)
Self-Prophetic Decoding to Unlock Visual Search in LVLMs
von: He, Zhendong, et al.
Veröffentlicht: (2026)
von: He, Zhendong, et al.
Veröffentlicht: (2026)
Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2024)
Generative Modelling with High-Order Langevin Dynamics
von: Shi, Ziqiang, et al.
Veröffentlicht: (2024)
von: Shi, Ziqiang, et al.
Veröffentlicht: (2024)
See Different, Think Better: Visual Variations Mitigating Hallucinations in LVLMs
von: Dai, Ziyun, et al.
Veröffentlicht: (2025)
von: Dai, Ziyun, et al.
Veröffentlicht: (2025)
RTAT: A Robust Two-stage Association Tracker for Multi-Object Tracking
von: Guo, Song, et al.
Veröffentlicht: (2024)
von: Guo, Song, et al.
Veröffentlicht: (2024)
Getting to the Point: Pointing Improves LVLMs at Counting
von: Alghisi, Simone, et al.
Veröffentlicht: (2026)
von: Alghisi, Simone, et al.
Veröffentlicht: (2026)
Identify, Isolate, and Purge: Mitigating Hallucinations in LVLMs via Self-Evolving Distillation
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
Elevating Visual Question Answering through Implicitly Learned Reasoning Pathways in LVLMs
von: Jing, Liu, et al.
Veröffentlicht: (2025)
von: Jing, Liu, et al.
Veröffentlicht: (2025)
SHALE: A Scalable Benchmark for Fine-grained Hallucination Evaluation in LVLMs
von: Yan, Bei, et al.
Veröffentlicht: (2025)
von: Yan, Bei, et al.
Veröffentlicht: (2025)
FALFormer: Feature-aware Landmarks self-attention for Whole-slide Image Classification
von: Bui, Doanh C., et al.
Veröffentlicht: (2024)
von: Bui, Doanh C., et al.
Veröffentlicht: (2024)
RAUM-Net: Regional Attention and Uncertainty-aware Mamba Network
von: Liu, Mingquan
Veröffentlicht: (2025)
von: Liu, Mingquan
Veröffentlicht: (2025)
Flash Window Attention: speedup the attention computation for Swin Transformer
von: Zhang, Zhendong
Veröffentlicht: (2025)
von: Zhang, Zhendong
Veröffentlicht: (2025)
Ähnliche Einträge
-
VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing
von: Huang, Yanbin, et al.
Veröffentlicht: (2026) -
Stop learning it all to mitigate visual hallucination, Focus on the hallucination target
von: Yoon, Dokyoon, et al.
Veröffentlicht: (2025) -
Technical report on label-informed logit redistribution for better domain generalization in low-shot classification with foundation models
von: Khan, Behraj, et al.
Veröffentlicht: (2025) -
Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement
von: Qin, Zhenxin, et al.
Veröffentlicht: (2026) -
Make LVLMs Focus: Context-Aware Attention Modulation for Better Multimodal In-Context Learning
von: Li, Yanshu, et al.
Veröffentlicht: (2025)