Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Zou, Zhengtao, Gao, Ya, Guan, Jiarui, Li, Bin, Marttinen, Pekka |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
por: Zhang, Yuanhong, et al.
Publicado: (2026)
por: Zhang, Yuanhong, et al.
Publicado: (2026)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
por: Yin, Jianghao, et al.
Publicado: (2026)
por: Yin, Jianghao, et al.
Publicado: (2026)
Mitigating Hallucination in Vision-Language Models through Barrier-Regulated Adaptive Closed-form Steering
por: Jana, Soumyadeep, et al.
Publicado: (2026)
por: Jana, Soumyadeep, et al.
Publicado: (2026)
Residual Decoding: Mitigating Hallucinations in Large Vision-Language Models via History-Aware Residual Guidance
por: Chen, Xinrong, et al.
Publicado: (2026)
por: Chen, Xinrong, et al.
Publicado: (2026)
Vision-Language Introspection: Mitigating Overconfident Hallucinations in MLLMs via Interpretable Bi-Causal Steering
por: Liu, Shuliang, et al.
Publicado: (2026)
por: Liu, Shuliang, et al.
Publicado: (2026)
Latent-Compressed Variational Autoencoder for Video Diffusion Models
por: Guan, Jiarui, et al.
Publicado: (2026)
por: Guan, Jiarui, et al.
Publicado: (2026)
Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations
por: Chen, Boxu, et al.
Publicado: (2025)
por: Chen, Boxu, et al.
Publicado: (2025)
Reducing Hallucinations in Vision-Language Models via Latent Space Steering
por: Liu, Sheng, et al.
Publicado: (2024)
por: Liu, Sheng, et al.
Publicado: (2024)
Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models
por: Zhang, Chengsheng, et al.
Publicado: (2026)
por: Zhang, Chengsheng, et al.
Publicado: (2026)
Mitigating Multilingual Hallucination in Large Vision-Language Models
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
por: Wang, Zihu, et al.
Publicado: (2025)
por: Wang, Zihu, et al.
Publicado: (2025)
Conscious Gaze: Adaptive Attention Mechanisms for Hallucination Mitigation in Vision-Language Models
por: Bu, Weijue, et al.
Publicado: (2025)
por: Bu, Weijue, et al.
Publicado: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
por: Zhu, Younan, et al.
Publicado: (2025)
por: Zhu, Younan, et al.
Publicado: (2025)
INTER: Mitigating Hallucination in Large Vision-Language Models by Interaction Guidance Sampling
por: Dong, Xin, et al.
Publicado: (2025)
por: Dong, Xin, et al.
Publicado: (2025)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
por: Woo, Sangmin, et al.
Publicado: (2025)
por: Woo, Sangmin, et al.
Publicado: (2025)
SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models
por: Xia, Yuxuan, et al.
Publicado: (2026)
por: Xia, Yuxuan, et al.
Publicado: (2026)
MedHEval: Benchmarking Hallucinations and Mitigation Strategies in Medical Large Vision-Language Models
por: Chang, Aofei, et al.
Publicado: (2025)
por: Chang, Aofei, et al.
Publicado: (2025)
MAP: Mitigating Hallucinations in Large Vision-Language Models with Map-Level Attention Processing
por: Li, Chenxi, et al.
Publicado: (2025)
por: Li, Chenxi, et al.
Publicado: (2025)
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
por: Li, Zhuowei, et al.
Publicado: (2025)
por: Li, Zhuowei, et al.
Publicado: (2025)
Mixture of Decoding: An Attention-Inspired Adaptive Decoding Strategy to Mitigate Hallucinations in Large Vision-Language Models
por: Chen, Xinlong, et al.
Publicado: (2025)
por: Chen, Xinlong, et al.
Publicado: (2025)
A Low-Rank Method for Vision Language Model Hallucination Mitigation in Autonomous Driving
por: Long, Keke, et al.
Publicado: (2025)
por: Long, Keke, et al.
Publicado: (2025)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024)
por: Zhong, Weihong, et al.
Publicado: (2024)
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
por: Xie, Yuxi, et al.
Publicado: (2024)
por: Xie, Yuxi, et al.
Publicado: (2024)
Strategies for Robust Deep Learning Based Deformable Registration
por: Honkamaa, Joel, et al.
Publicado: (2025)
por: Honkamaa, Joel, et al.
Publicado: (2025)
New multimodal similarity measure for image registration via modeling local functional dependence with linear combination of learned basis functions
por: Honkamaa, Joel, et al.
Publicado: (2025)
por: Honkamaa, Joel, et al.
Publicado: (2025)
SITReg: Multi-resolution architecture for symmetric, inverse consistent, and topology preserving image registration
por: Honkamaa, Joel, et al.
Publicado: (2023)
por: Honkamaa, Joel, et al.
Publicado: (2023)
Investigating and Mitigating Object Hallucinations in Pretrained Vision-Language (CLIP) Models
por: Liu, Yufang, et al.
Publicado: (2024)
por: Liu, Yufang, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
por: Min, Kyungmin, et al.
Publicado: (2024)
por: Min, Kyungmin, et al.
Publicado: (2024)
Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models
por: Lee, Yi-Lun, et al.
Publicado: (2024)
por: Lee, Yi-Lun, et al.
Publicado: (2024)
Object-level Self-Distillation for Vision Pretraining
por: Hızlı, Çağlar, et al.
Publicado: (2025)
por: Hızlı, Çağlar, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
por: Wang, Xintong, et al.
Publicado: (2024)
por: Wang, Xintong, et al.
Publicado: (2024)
BIMA: Bijective Maximum Likelihood Learning Approach to Hallucination Prediction and Mitigation in Large Vision-Language Models
por: Tran, Huu-Thien, et al.
Publicado: (2025)
por: Tran, Huu-Thien, et al.
Publicado: (2025)
Prescribing the Right Remedy: Mitigating Hallucinations in Large Vision-Language Models via Targeted Instruction Tuning
por: Hu, Rui, et al.
Publicado: (2024)
por: Hu, Rui, et al.
Publicado: (2024)
Mitigating Hallucinations in Video Large Language Models via Spatiotemporal-Semantic Contrastive Decoding
por: Gao, Yuansheng, et al.
Publicado: (2026)
por: Gao, Yuansheng, et al.
Publicado: (2026)
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
por: Li, Qiming, et al.
Publicado: (2026)
por: Li, Qiming, et al.
Publicado: (2026)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
por: Lu, Yifan, et al.
Publicado: (2025)
por: Lu, Yifan, et al.
Publicado: (2025)
Review of Hallucination Understanding in Large Language and Vision Models
por: Ho, Zhengyi, et al.
Publicado: (2025)
por: Ho, Zhengyi, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)
por: Manevich, Avshalom, et al.
Publicado: (2024)
por: Manevich, Avshalom, et al.
Publicado: (2024)
Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models
por: Tang, Kai, et al.
Publicado: (2025)
por: Tang, Kai, et al.
Publicado: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
por: An, Wenbin, et al.
Publicado: (2024)
por: An, Wenbin, et al.
Publicado: (2024)
Ejemplares similares
-
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
por: Zhang, Yuanhong, et al.
Publicado: (2026) -
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
por: Yin, Jianghao, et al.
Publicado: (2026) -
Mitigating Hallucination in Vision-Language Models through Barrier-Regulated Adaptive Closed-form Steering
por: Jana, Soumyadeep, et al.
Publicado: (2026) -
Residual Decoding: Mitigating Hallucinations in Large Vision-Language Models via History-Aware Residual Guidance
por: Chen, Xinrong, et al.
Publicado: (2026) -
Vision-Language Introspection: Mitigating Overconfident Hallucinations in MLLMs via Interpretable Bi-Causal Steering
por: Liu, Shuliang, et al.
Publicado: (2026)