Explainable and Interpretable Multimodal Large Language Models: A Comprehensive Survey
Fuente:
arXiv
Salvato in:
| Autori principali: | Dang, Yunkai, Huang, Kaichen, Huo, Jiahao, Yan, Yibo, Huang, Sirui, Liu, Dongrui, Gao, Mengxi, Zhang, Jie, Qian, Chen, Wang, Kun, Liu, Yong, Shao, Jing, Xiong, Hui, Hu, Xuming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MINER: Mining the Underlying Pattern of Modality-Specific Neurons in Multimodal Large Language Models
di: Huang, Kaichen, et al.
Pubblicazione: (2024)
di: Huang, Kaichen, et al.
Pubblicazione: (2024)
MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model
di: Huo, Jiahao, et al.
Pubblicazione: (2024)
di: Huo, Jiahao, et al.
Pubblicazione: (2024)
The Tug of War Within: Mitigating the Fairness-Privacy Conflicts in Large Language Models
di: Qian, Chen, et al.
Pubblicazione: (2024)
di: Qian, Chen, et al.
Pubblicazione: (2024)
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis
di: Huang, Haoming, et al.
Pubblicazione: (2025)
di: Huang, Haoming, et al.
Pubblicazione: (2025)
Exploring Response Uncertainty in MLLMs: An Empirical Evaluation under Misleading Scenarios
di: Dang, Yunkai, et al.
Pubblicazione: (2024)
di: Dang, Yunkai, et al.
Pubblicazione: (2024)
VLSBench: Unveiling Visual Leakage in Multimodal Safety
di: Hu, Xuhao, et al.
Pubblicazione: (2024)
di: Hu, Xuhao, et al.
Pubblicazione: (2024)
MMUnlearner: Reformulating Multimodal Machine Unlearning in the Era of Multimodal Large Language Models
di: Huo, Jiahao, et al.
Pubblicazione: (2025)
di: Huo, Jiahao, et al.
Pubblicazione: (2025)
MathAgent: Leveraging a Mixture-of-Math-Agent Framework for Real-World Multimodal Mathematical Error Detection
di: Yan, Yibo, et al.
Pubblicazione: (2025)
di: Yan, Yibo, et al.
Pubblicazione: (2025)
Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering Capabilities
di: Li, Zhonghao, et al.
Pubblicazione: (2024)
di: Li, Zhonghao, et al.
Pubblicazione: (2024)
RiOSWorld: Benchmarking the Risk of Multimodal Computer-Use Agents
di: Yang, Jingyi, et al.
Pubblicazione: (2025)
di: Yang, Jingyi, et al.
Pubblicazione: (2025)
REEF: Representation Encoding Fingerprints for Large Language Models
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
The Better Angels of Machine Personality: How Personality Relates to LLM Safety
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
RvB: Automating AI System Hardening via Iterative Red-Blue Games
di: Huang, Lige, et al.
Pubblicazione: (2026)
di: Huang, Lige, et al.
Pubblicazione: (2026)
PACEbench: A Framework for Evaluating Practical AI Cyber-Exploitation Capabilities
di: Liu, Zicheng, et al.
Pubblicazione: (2025)
di: Liu, Zicheng, et al.
Pubblicazione: (2025)
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
di: Su, Jiamin, et al.
Pubblicazione: (2025)
di: Su, Jiamin, et al.
Pubblicazione: (2025)
Interpreting Emergent Extreme Events in Multi-Agent Systems
di: Tang, Ling, et al.
Pubblicazione: (2026)
di: Tang, Ling, et al.
Pubblicazione: (2026)
Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning
di: Qian, Chen, et al.
Pubblicazione: (2025)
di: Qian, Chen, et al.
Pubblicazione: (2025)
ReasonAny: Incorporating Reasoning Capability to Any Model via Simple and Effective Model Merging
di: Yang, Junyao, et al.
Pubblicazione: (2026)
di: Yang, Junyao, et al.
Pubblicazione: (2026)
Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations
di: Yan, Yibo, et al.
Pubblicazione: (2026)
di: Yan, Yibo, et al.
Pubblicazione: (2026)
Mitigating Modality Prior-Induced Hallucinations in Multimodal Large Language Models via Deciphering Attention Causality
di: Zhou, Guanyu, et al.
Pubblicazione: (2024)
di: Zhou, Guanyu, et al.
Pubblicazione: (2024)
PMark: Towards Robust and Distortion-free Semantic-level Watermarking with Channel Constraints
di: Huo, Jiahao, et al.
Pubblicazione: (2025)
di: Huo, Jiahao, et al.
Pubblicazione: (2025)
Loop as a Bridge: Can Looped Transformers Truly Link Representation Space and Natural Language Outputs?
di: Chen, Guanxu, et al.
Pubblicazione: (2026)
di: Chen, Guanxu, et al.
Pubblicazione: (2026)
Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
di: Zheng, Kening, et al.
Pubblicazione: (2024)
di: Zheng, Kening, et al.
Pubblicazione: (2024)
Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period of Large Language Models
di: Qian, Chen, et al.
Pubblicazione: (2024)
di: Qian, Chen, et al.
Pubblicazione: (2024)
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection
di: Yan, Yibo, et al.
Pubblicazione: (2024)
di: Yan, Yibo, et al.
Pubblicazione: (2024)
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval
di: Yan, Yibo, et al.
Pubblicazione: (2026)
di: Yan, Yibo, et al.
Pubblicazione: (2026)
Sculpting the Vector Space: Towards Efficient Multi-Vector Visual Document Retrieval via Prune-then-Merge Framework
di: Yan, Yibo, et al.
Pubblicazione: (2026)
di: Yan, Yibo, et al.
Pubblicazione: (2026)
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs
di: Fu, Chaoyou, et al.
Pubblicazione: (2024)
di: Fu, Chaoyou, et al.
Pubblicazione: (2024)
Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models
di: Zou, Xin, et al.
Pubblicazione: (2024)
di: Zou, Xin, et al.
Pubblicazione: (2024)
LED-Merging: Mitigating Safety-Utility Conflicts in Model Merging with Location-Election-Disjoint
di: Ma, Qianli, et al.
Pubblicazione: (2025)
di: Ma, Qianli, et al.
Pubblicazione: (2025)
Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation
di: Zheng, Kening, et al.
Pubblicazione: (2026)
di: Zheng, Kening, et al.
Pubblicazione: (2026)
Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning
di: Yan, Yibo, et al.
Pubblicazione: (2025)
di: Yan, Yibo, et al.
Pubblicazione: (2025)
TradeTrap: Are LLM-based Trading Agents Truly Reliable and Faithful?
di: Yan, Lewen, et al.
Pubblicazione: (2025)
di: Yan, Lewen, et al.
Pubblicazione: (2025)
CausalEmbed: Auto-Regressive Multi-Vector Generation in Latent Space for Visual Document Embedding
di: Huo, Jiahao, et al.
Pubblicazione: (2026)
di: Huo, Jiahao, et al.
Pubblicazione: (2026)
Optimized Cost Per Click in Online Advertising: A Theoretical Analysis
di: Zhang, Kaichen, et al.
Pubblicazione: (2024)
di: Zhang, Kaichen, et al.
Pubblicazione: (2024)
The Impact of Generative Artificial Intelligence on Market Equilibrium: Evidence from a Natural Experiment
di: Zhang, Kaichen, et al.
Pubblicazione: (2023)
di: Zhang, Kaichen, et al.
Pubblicazione: (2023)
A Comprehensive Survey on Self-Interpretable Neural Networks
di: Ji, Yang, et al.
Pubblicazione: (2025)
di: Ji, Yang, et al.
Pubblicazione: (2025)
A Survey of AIOps in the Era of Large Language Models
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
LLMs Deceive Unintentionally: Emergent Misalignment in Dishonesty from Misaligned Samples to Biased Human-AI Interactions
di: Hu, Xuhao, et al.
Pubblicazione: (2025)
di: Hu, Xuhao, et al.
Pubblicazione: (2025)
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
di: Yan, Yibo, et al.
Pubblicazione: (2024)
di: Yan, Yibo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MINER: Mining the Underlying Pattern of Modality-Specific Neurons in Multimodal Large Language Models
di: Huang, Kaichen, et al.
Pubblicazione: (2024) -
MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model
di: Huo, Jiahao, et al.
Pubblicazione: (2024) -
The Tug of War Within: Mitigating the Fairness-Privacy Conflicts in Large Language Models
di: Qian, Chen, et al.
Pubblicazione: (2024) -
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis
di: Huang, Haoming, et al.
Pubblicazione: (2025) -
Exploring Response Uncertainty in MLLMs: An Empirical Evaluation under Misleading Scenarios
di: Dang, Yunkai, et al.
Pubblicazione: (2024)