CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
Fuente:
arXiv
Salvato in:
| Autori principali: | Ye, Zekai, Li, Qiming, Feng, Xiaocheng, Qin, Libo, Huang, Yichong, Li, Baohang, Jiang, Kui, Xiang, Yang, Zhang, Zhirui, Lu, Yunfei, Tang, Duyu, Tu, Dandan, Qin, Bing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
di: Li, Qiming, et al.
Pubblicazione: (2026)
di: Li, Qiming, et al.
Pubblicazione: (2026)
CC-Tuning: A Cross-Lingual Connection Mechanism for Improving Joint Multilingual Supervised Fine-Tuning
di: Ye, Yangfan, et al.
Pubblicazione: (2025)
di: Ye, Yangfan, et al.
Pubblicazione: (2025)
Ensuring Consistency for In-Image Translation
di: Fu, Chengpeng, et al.
Pubblicazione: (2024)
di: Fu, Chengpeng, et al.
Pubblicazione: (2024)
Enhancing Non-English Capabilities of English-Centric Large Language Models through Deep Supervision Fine-Tuning
di: Huo, Wenshuai, et al.
Pubblicazione: (2025)
di: Huo, Wenshuai, et al.
Pubblicazione: (2025)
Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges
di: Ye, Yangfan, et al.
Pubblicazione: (2024)
di: Ye, Yangfan, et al.
Pubblicazione: (2024)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation
di: Yuan, Zekun, et al.
Pubblicazione: (2026)
di: Yuan, Zekun, et al.
Pubblicazione: (2026)
LangGPS: Language Separability Guided Data Pre-Selection for Joint Multilingual Instruction Tuning
di: Ye, Yangfan, et al.
Pubblicazione: (2025)
di: Ye, Yangfan, et al.
Pubblicazione: (2025)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
di: Zhong, Weihong, et al.
Pubblicazione: (2024)
di: Zhong, Weihong, et al.
Pubblicazione: (2024)
Aligning Translation-Specific Understanding to General Understanding in Large Language Models
di: Huang, Yichong, et al.
Pubblicazione: (2024)
di: Huang, Yichong, et al.
Pubblicazione: (2024)
Ensemble Learning for Heterogeneous Large Language Models with Deep Parallel Collaboration
di: Huang, Yichong, et al.
Pubblicazione: (2024)
di: Huang, Yichong, et al.
Pubblicazione: (2024)
Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
Cross-Lingual Text-Rich Visual Comprehension: An Information Theory Perspective
di: Yu, Xinmiao, et al.
Pubblicazione: (2024)
di: Yu, Xinmiao, et al.
Pubblicazione: (2024)
Relay Decoding: Concatenating Large Language Models for Machine Translation
di: Fu, Chengpeng, et al.
Pubblicazione: (2024)
di: Fu, Chengpeng, et al.
Pubblicazione: (2024)
Mitigating Object Hallucinations in LVLMs via Attention Imbalance Rectification
di: Sun, Han, et al.
Pubblicazione: (2026)
di: Sun, Han, et al.
Pubblicazione: (2026)
x1: Learning to Think Adaptively Across Languages and Cultures
di: Ye, Yangfan, et al.
Pubblicazione: (2026)
di: Ye, Yangfan, et al.
Pubblicazione: (2026)
FroM: Frobenius Norm-Based Data-Free Adaptive Model Merging
di: Li, Zijian, et al.
Pubblicazione: (2025)
di: Li, Zijian, et al.
Pubblicazione: (2025)
One for All: Update Parameterized Knowledge Across Multiple Models
di: Ma, Weitao, et al.
Pubblicazione: (2025)
di: Ma, Weitao, et al.
Pubblicazione: (2025)
Mitigating Object Hallucination via Concentric Causal Attention
di: Xing, Yun, et al.
Pubblicazione: (2024)
di: Xing, Yun, et al.
Pubblicazione: (2024)
Not All Tokens See Equally: Perception-Grounded Policy Optimization for Large Vision-Language Models
di: Ye, Zekai, et al.
Pubblicazione: (2026)
di: Ye, Zekai, et al.
Pubblicazione: (2026)
HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding
di: Yuan, Fan, et al.
Pubblicazione: (2024)
di: Yuan, Fan, et al.
Pubblicazione: (2024)
Learning Fine-Grained Grounded Citations for Attributed Large Language Models
di: Huang, Lei, et al.
Pubblicazione: (2024)
di: Huang, Lei, et al.
Pubblicazione: (2024)
Mitigating Multilingual Hallucination in Large Vision-Language Models
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
Mitigating Object Hallucinations via Sentence-Level Early Intervention
di: Peng, Shangpin, et al.
Pubblicazione: (2025)
di: Peng, Shangpin, et al.
Pubblicazione: (2025)
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
di: Deng, Ailin, et al.
Pubblicazione: (2024)
di: Deng, Ailin, et al.
Pubblicazione: (2024)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
Causally-Grounded Dual-Path Attention Intervention for Object Hallucination Mitigation in LVLMs
di: Yu, Liu, et al.
Pubblicazione: (2025)
di: Yu, Liu, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
di: Lu, Yifan, et al.
Pubblicazione: (2025)
di: Lu, Yifan, et al.
Pubblicazione: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
di: An, Wenbin, et al.
Pubblicazione: (2024)
di: An, Wenbin, et al.
Pubblicazione: (2024)
CCHall: A Novel Benchmark for Joint Cross-Lingual and Cross-Modal Hallucinations Detection in Large Language Models
di: Zhang, Yongheng, et al.
Pubblicazione: (2025)
di: Zhang, Yongheng, et al.
Pubblicazione: (2025)
Breaking Language Barriers: Cross-Lingual Continual Pre-Training at Scale
di: Zheng, Wenzhen, et al.
Pubblicazione: (2024)
di: Zheng, Wenzhen, et al.
Pubblicazione: (2024)
CCL-XCoT: An Efficient Cross-Lingual Knowledge Transfer Method for Mitigating Hallucination Generation
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
di: Zhu, Younan, et al.
Pubblicazione: (2025)
di: Zhu, Younan, et al.
Pubblicazione: (2025)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
di: Li, Bin, et al.
Pubblicazione: (2025)
di: Li, Bin, et al.
Pubblicazione: (2025)
Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models
di: Zhang, Chengsheng, et al.
Pubblicazione: (2026)
di: Zhang, Chengsheng, et al.
Pubblicazione: (2026)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
di: Wang, Zihu, et al.
Pubblicazione: (2025)
di: Wang, Zihu, et al.
Pubblicazione: (2025)
Look Closer! An Adversarial Parametric Editing Framework for Hallucination Mitigation in VLMs
di: Hu, Jiayu, et al.
Pubblicazione: (2025)
di: Hu, Jiayu, et al.
Pubblicazione: (2025)
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
di: Zhou, Yiyang, et al.
Pubblicazione: (2023)
di: Zhou, Yiyang, et al.
Pubblicazione: (2023)
Do Vision Encoders Truly Explain Object Hallucination?: Mitigating Object Hallucination via Simple Fine-Grained CLIPScore
di: Oh, Hongseok, et al.
Pubblicazione: (2025)
di: Oh, Hongseok, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Li, Qiming, et al.
Pubblicazione: (2025) -
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
di: Li, Qiming, et al.
Pubblicazione: (2026) -
CC-Tuning: A Cross-Lingual Connection Mechanism for Improving Joint Multilingual Supervised Fine-Tuning
di: Ye, Yangfan, et al.
Pubblicazione: (2025) -
Ensuring Consistency for In-Image Translation
di: Fu, Chengpeng, et al.
Pubblicazione: (2024) -
Enhancing Non-English Capabilities of English-Centric Large Language Models through Deep Supervision Fine-Tuning
di: Huo, Wenshuai, et al.
Pubblicazione: (2025)