SECOND: Mitigating Perceptual Hallucination in Vision-Language Models via Selective and Contrastive Decoding
Fuente:
arXiv
Salvato in:
| Autori principali: | Park, Woohyeon, Kim, Woojin, Kim, Jaeik, Do, Jaeyoung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MMPB: It's Time for Multi-Modal Personalization
di: Kim, Jaeik, et al.
Pubblicazione: (2025)
di: Kim, Jaeik, et al.
Pubblicazione: (2025)
Exploring and Leveraging Class Vectors for Classifier Editing
di: Kim, Jaeik, et al.
Pubblicazione: (2025)
di: Kim, Jaeik, et al.
Pubblicazione: (2025)
MEDIC-AD: Towards Medical Vision-Language Model's Clinical Intelligence
di: Park, Woohyeon, et al.
Pubblicazione: (2026)
di: Park, Woohyeon, et al.
Pubblicazione: (2026)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
di: Min, Kyungmin, et al.
Pubblicazione: (2024)
di: Min, Kyungmin, et al.
Pubblicazione: (2024)
Mitigating Hallucination in Visual-Language Models via Re-Balancing Contrastive Decoding
di: Liang, Xiaoyu, et al.
Pubblicazione: (2024)
di: Liang, Xiaoyu, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
di: Wang, Chao, et al.
Pubblicazione: (2025)
di: Wang, Chao, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)
di: Manevich, Avshalom, et al.
Pubblicazione: (2024)
di: Manevich, Avshalom, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
di: Wang, Xintong, et al.
Pubblicazione: (2024)
di: Wang, Xintong, et al.
Pubblicazione: (2024)
Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models
di: Lee, Yi-Lun, et al.
Pubblicazione: (2024)
di: Lee, Yi-Lun, et al.
Pubblicazione: (2024)
Med-VCD: Mitigating Hallucination for Medical Large Vision Language Models through Visual Contrastive Decoding
di: Mahdavi, Zahra, et al.
Pubblicazione: (2025)
di: Mahdavi, Zahra, et al.
Pubblicazione: (2025)
SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models
di: Xia, Yuxuan, et al.
Pubblicazione: (2026)
di: Xia, Yuxuan, et al.
Pubblicazione: (2026)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
di: Kim, Minchan, et al.
Pubblicazione: (2024)
di: Kim, Minchan, et al.
Pubblicazione: (2024)
HalLoc: Token-level Localization of Hallucinations for Vision Language Models
di: Park, Eunkyu, et al.
Pubblicazione: (2025)
di: Park, Eunkyu, et al.
Pubblicazione: (2025)
VaLiD: Mitigating the Hallucination of Large Vision Language Models by Visual Layer Fusion Contrastive Decoding
di: Wang, Jiaqi, et al.
Pubblicazione: (2024)
di: Wang, Jiaqi, et al.
Pubblicazione: (2024)
ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models
di: Park, Yeji, et al.
Pubblicazione: (2024)
di: Park, Yeji, et al.
Pubblicazione: (2024)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
di: Fieback, Laura, et al.
Pubblicazione: (2025)
di: Fieback, Laura, et al.
Pubblicazione: (2025)
AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
di: Jung, Chaeyoung, et al.
Pubblicazione: (2025)
di: Jung, Chaeyoung, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Video Large Language Models via Spatiotemporal-Semantic Contrastive Decoding
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
ReCo: Reminder Composition Mitigates Hallucinations in Vision-Language Models
di: Chytas, Sotirios Panagiotis, et al.
Pubblicazione: (2025)
di: Chytas, Sotirios Panagiotis, et al.
Pubblicazione: (2025)
Mitigating Hallucination in Multimodal LLMs with Layer Contrastive Decoding
di: Tong, Bingkui, et al.
Pubblicazione: (2025)
di: Tong, Bingkui, et al.
Pubblicazione: (2025)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
di: Lee, Jihoon, et al.
Pubblicazione: (2025)
di: Lee, Jihoon, et al.
Pubblicazione: (2025)
Watermarking for Factuality: Guiding Vision-Language Models Toward Truth via Tri-layer Contrastive Decoding
di: Back, Kyungryul, et al.
Pubblicazione: (2025)
di: Back, Kyungryul, et al.
Pubblicazione: (2025)
SAKED: Mitigating Hallucination in Large Vision-Language Models via Stability-Aware Knowledge Enhanced Decoding
di: Li, Zhaoxu, et al.
Pubblicazione: (2026)
di: Li, Zhaoxu, et al.
Pubblicazione: (2026)
Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning
di: Park, Dongmin, et al.
Pubblicazione: (2024)
di: Park, Dongmin, et al.
Pubblicazione: (2024)
YARD: Y-Architecture Register Decoding for Efficient Hallucination Mitigation in Large Vision-Language Models
di: Chen, Ting, et al.
Pubblicazione: (2026)
di: Chen, Ting, et al.
Pubblicazione: (2026)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
di: Zhang, Ce, et al.
Pubblicazione: (2025)
di: Zhang, Ce, et al.
Pubblicazione: (2025)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
di: Ma, Ruiqi, et al.
Pubblicazione: (2025)
di: Ma, Ruiqi, et al.
Pubblicazione: (2025)
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
di: Kim, Namhee, et al.
Pubblicazione: (2025)
di: Kim, Namhee, et al.
Pubblicazione: (2025)
Residual Decoding: Mitigating Hallucinations in Large Vision-Language Models via History-Aware Residual Guidance
di: Chen, Xinrong, et al.
Pubblicazione: (2026)
di: Chen, Xinrong, et al.
Pubblicazione: (2026)
Mitigating Hallucinations in Large Vision-Language Models via Causal Route Gating
di: Cheng, Zhe, et al.
Pubblicazione: (2026)
di: Cheng, Zhe, et al.
Pubblicazione: (2026)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
di: Lu, Yifan, et al.
Pubblicazione: (2025)
di: Lu, Yifan, et al.
Pubblicazione: (2025)
Mixture of Decoding: An Attention-Inspired Adaptive Decoding Strategy to Mitigate Hallucinations in Large Vision-Language Models
di: Chen, Xinlong, et al.
Pubblicazione: (2025)
di: Chen, Xinlong, et al.
Pubblicazione: (2025)
EggHand: A Multimodal Foundation Model for Egocentric Hand Pose Forecasting
di: Choi, Jaeyoung, et al.
Pubblicazione: (2026)
di: Choi, Jaeyoung, et al.
Pubblicazione: (2026)
IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding
di: Zhu, Lanyun, et al.
Pubblicazione: (2024)
di: Zhu, Lanyun, et al.
Pubblicazione: (2024)
Exploring Causes and Mitigation of Hallucinations in Large Vision Language Models
di: Sun, Yaqi, et al.
Pubblicazione: (2025)
di: Sun, Yaqi, et al.
Pubblicazione: (2025)
Mask What Matters: Mitigating Object Hallucinations in Multimodal Large Language Models with Object-Aligned Visual Contrastive Decoding
di: Chen, Boqi, et al.
Pubblicazione: (2026)
di: Chen, Boqi, et al.
Pubblicazione: (2026)
Mitigating Image Captioning Hallucinations in Vision-Language Models
di: Zhao, Fei, et al.
Pubblicazione: (2025)
di: Zhao, Fei, et al.
Pubblicazione: (2025)
SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding
di: Wu, Chang-Hsun, et al.
Pubblicazione: (2025)
di: Wu, Chang-Hsun, et al.
Pubblicazione: (2025)
Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models
di: Zhang, Jinrui, et al.
Pubblicazione: (2024)
di: Zhang, Jinrui, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation
di: Zhu, Xingyu, et al.
Pubblicazione: (2026)
di: Zhu, Xingyu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
MMPB: It's Time for Multi-Modal Personalization
di: Kim, Jaeik, et al.
Pubblicazione: (2025) -
Exploring and Leveraging Class Vectors for Classifier Editing
di: Kim, Jaeik, et al.
Pubblicazione: (2025) -
MEDIC-AD: Towards Medical Vision-Language Model's Clinical Intelligence
di: Park, Woohyeon, et al.
Pubblicazione: (2026) -
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
di: Min, Kyungmin, et al.
Pubblicazione: (2024) -
Mitigating Hallucination in Visual-Language Models via Re-Balancing Contrastive Decoding
di: Liang, Xiaoyu, et al.
Pubblicazione: (2024)