Mitigating Hallucinations in Large Vision-Language Models via Entity-Centric Multimodal Preference Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Jiulong, Shi, Zhengliang, Wang, Shuaiqiang, Huang, Jizhou, Yin, Dawei, Yan, Lingyong, Cao, Min, Zhang, Min |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning
di: Shen, Yucheng, et al.
Pubblicazione: (2026)
di: Shen, Yucheng, et al.
Pubblicazione: (2026)
GRAF: Multi-turn Jailbreaking via Global Refinement and Active Fabrication
di: Tang, Hua, et al.
Pubblicazione: (2025)
di: Tang, Hua, et al.
Pubblicazione: (2025)
Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models
di: Shi, Zhengliang, et al.
Pubblicazione: (2025)
di: Shi, Zhengliang, et al.
Pubblicazione: (2025)
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation
di: Zhu, Dongsheng, et al.
Pubblicazione: (2025)
di: Zhu, Dongsheng, et al.
Pubblicazione: (2025)
Facial-R1: Aligning Reasoning and Recognition for Facial Emotion Analysis
di: Wu, Jiulong, et al.
Pubblicazione: (2025)
di: Wu, Jiulong, et al.
Pubblicazione: (2025)
Reasoning-to-Defend: Safety-Aware Reasoning Can Defend Large Language Models from Jailbreaking
di: Zhu, Junda, et al.
Pubblicazione: (2025)
di: Zhu, Junda, et al.
Pubblicazione: (2025)
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization
di: Wu, Jiayi, et al.
Pubblicazione: (2024)
di: Wu, Jiayi, et al.
Pubblicazione: (2024)
Iterative Self-Incentivization Empowers Large Language Models as Agentic Searchers
di: Shi, Zhengliang, et al.
Pubblicazione: (2025)
di: Shi, Zhengliang, et al.
Pubblicazione: (2025)
MAO-ARAG: Multi-Agent Orchestration for Adaptive Retrieval-Augmented Generation
di: Chen, Yiqun, et al.
Pubblicazione: (2025)
di: Chen, Yiqun, et al.
Pubblicazione: (2025)
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
di: Xie, Yuxi, et al.
Pubblicazione: (2024)
di: Xie, Yuxi, et al.
Pubblicazione: (2024)
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
di: Fu, Yuhan, et al.
Pubblicazione: (2024)
di: Fu, Yuhan, et al.
Pubblicazione: (2024)
The Real, the Better: Aligning Large Language Models with Online Human Behaviors
di: Jiang, Guanying, et al.
Pubblicazione: (2024)
di: Jiang, Guanying, et al.
Pubblicazione: (2024)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
di: Yan, Lingyong, et al.
Pubblicazione: (2026)
di: Yan, Lingyong, et al.
Pubblicazione: (2026)
Improving the Robustness of Large Language Models via Consistency Alignment
di: Zhao, Yukun, et al.
Pubblicazione: (2024)
di: Zhao, Yukun, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
di: Min, Kyungmin, et al.
Pubblicazione: (2024)
di: Min, Kyungmin, et al.
Pubblicazione: (2024)
Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking Agents
di: Sun, Weiwei, et al.
Pubblicazione: (2023)
di: Sun, Weiwei, et al.
Pubblicazione: (2023)
Mitigating Entity-Level Hallucination in Large Language Models
di: Su, Weihang, et al.
Pubblicazione: (2024)
di: Su, Weihang, et al.
Pubblicazione: (2024)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
di: Lyu, Yougang, et al.
Pubblicazione: (2024)
di: Lyu, Yougang, et al.
Pubblicazione: (2024)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
di: Zhong, Weihong, et al.
Pubblicazione: (2024)
di: Zhong, Weihong, et al.
Pubblicazione: (2024)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
di: Yin, Jianghao, et al.
Pubblicazione: (2026)
di: Yin, Jianghao, et al.
Pubblicazione: (2026)
Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization
di: Compagnoni, Alberto, et al.
Pubblicazione: (2025)
di: Compagnoni, Alberto, et al.
Pubblicazione: (2025)
Debiasing Multimodal Large Language Models via Noise-Aware Preference Optimization
di: Zhang, Zefeng, et al.
Pubblicazione: (2025)
di: Zhang, Zefeng, et al.
Pubblicazione: (2025)
MACPO: Weak-to-Strong Alignment via Multi-Agent Contrastive Preference Optimization
di: Lyu, Yougang, et al.
Pubblicazione: (2024)
di: Lyu, Yougang, et al.
Pubblicazione: (2024)
MAIR: A Massive Benchmark for Evaluating Instructed Retrieval
di: Sun, Weiwei, et al.
Pubblicazione: (2024)
di: Sun, Weiwei, et al.
Pubblicazione: (2024)
Tool Learning in the Wild: Empowering Language Models as Automatic Tool Agents
di: Shi, Zhengliang, et al.
Pubblicazione: (2024)
di: Shi, Zhengliang, et al.
Pubblicazione: (2024)
Mitigating Hallucinated Translations in Large Language Models with Hallucination-focused Preference Optimization
di: Tang, Zilu, et al.
Pubblicazione: (2025)
di: Tang, Zilu, et al.
Pubblicazione: (2025)
Learning to Use Tools via Cooperative and Interactive Agents
di: Shi, Zhengliang, et al.
Pubblicazione: (2024)
di: Shi, Zhengliang, et al.
Pubblicazione: (2024)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
di: Lee, Jihoon, et al.
Pubblicazione: (2025)
di: Lee, Jihoon, et al.
Pubblicazione: (2025)
Mitigating Multimodal Hallucination via Phase-wise Self-reward
di: Zhang, Yu, et al.
Pubblicazione: (2026)
di: Zhang, Yu, et al.
Pubblicazione: (2026)
Joint Flashback Adaptation for Forgetting-Resistant Instruction Tuning
di: Zhao, Yukun, et al.
Pubblicazione: (2025)
di: Zhao, Yukun, et al.
Pubblicazione: (2025)
Watch Closely: Mitigating Object Hallucinations in Large Vision-Language Models with Disentangled Decoding
di: Ma, Ruiqi, et al.
Pubblicazione: (2025)
di: Ma, Ruiqi, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
di: Lu, Yifan, et al.
Pubblicazione: (2025)
di: Lu, Yifan, et al.
Pubblicazione: (2025)
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
di: Wang, Weihang, et al.
Pubblicazione: (2025)
di: Wang, Weihang, et al.
Pubblicazione: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
ATM: Adversarial Tuning Multi-agent System Makes a Robust Retrieval-Augmented Generator
di: Zhu, Junda, et al.
Pubblicazione: (2024)
di: Zhu, Junda, et al.
Pubblicazione: (2024)
MIHBench: Benchmarking and Mitigating Multi-Image Hallucinations in Multimodal Large Language Models
di: Li, Jiale, et al.
Pubblicazione: (2025)
di: Li, Jiale, et al.
Pubblicazione: (2025)
Mitigating Visual Hallucinations via Semantic Curriculum Preference Optimization in MLLMs
di: Li, Yuanshuai, et al.
Pubblicazione: (2025)
di: Li, Yuanshuai, et al.
Pubblicazione: (2025)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
di: Fazli, Mehrdad, et al.
Pubblicazione: (2025)
di: Fazli, Mehrdad, et al.
Pubblicazione: (2025)
Generative Multimodal Entity Linking
di: Shi, Senbao, et al.
Pubblicazione: (2023)
di: Shi, Senbao, et al.
Pubblicazione: (2023)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
di: Yeh, Min-Hsuan, et al.
Pubblicazione: (2025)
di: Yeh, Min-Hsuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning
di: Shen, Yucheng, et al.
Pubblicazione: (2026) -
GRAF: Multi-turn Jailbreaking via Global Refinement and Active Fabrication
di: Tang, Hua, et al.
Pubblicazione: (2025) -
Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models
di: Shi, Zhengliang, et al.
Pubblicazione: (2025) -
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation
di: Zhu, Dongsheng, et al.
Pubblicazione: (2025) -
Facial-R1: Aligning Reasoning and Recognition for Facial Emotion Analysis
di: Wu, Jiulong, et al.
Pubblicazione: (2025)