A Concept-Based Explainability Framework for Large Multimodal Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Parekh, Jayneel, Khayatan, Pegah, Shukor, Mustafa, Newson, Alasdair, Cord, Matthieu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning to Steer: Input-dependent Steering for Multimodal LLMs
di: Parekh, Jayneel, et al.
Pubblicazione: (2025)
di: Parekh, Jayneel, et al.
Pubblicazione: (2025)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
di: Khayatan, Pegah, et al.
Pubblicazione: (2026)
di: Khayatan, Pegah, et al.
Pubblicazione: (2026)
Analyzing Finetuning Representation Shift for Multimodal LLMs Steering
di: Khayatan, Pegah, et al.
Pubblicazione: (2025)
di: Khayatan, Pegah, et al.
Pubblicazione: (2025)
Restyling Unsupervised Concept Based Interpretable Networks with Generative Models
di: Parekh, Jayneel, et al.
Pubblicazione: (2024)
di: Parekh, Jayneel, et al.
Pubblicazione: (2024)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
Skipping Computations in Multimodal LLMs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
What Makes Multimodal In-Context Learning Work?
di: Baldassini, Folco Bertini, et al.
Pubblicazione: (2024)
di: Baldassini, Folco Bertini, et al.
Pubblicazione: (2024)
Zero-Shot Refinement of Buildings' Segmentation Models using SAM
di: Mayladan, Ali, et al.
Pubblicazione: (2023)
di: Mayladan, Ali, et al.
Pubblicazione: (2023)
Beyond Task Performance: Evaluating and Reducing the Flaws of Large Multimodal Models with In-Context Learning
di: Shukor, Mustafa, et al.
Pubblicazione: (2023)
di: Shukor, Mustafa, et al.
Pubblicazione: (2023)
ReGentS: Real-World Safety-Critical Driving Scenario Generation Made Stable
di: Yin, Yuan, et al.
Pubblicazione: (2024)
di: Yin, Yuan, et al.
Pubblicazione: (2024)
Explainable Multimodal Sentiment Analysis on Bengali Memes
di: Elahi, Kazi Toufique, et al.
Pubblicazione: (2023)
di: Elahi, Kazi Toufique, et al.
Pubblicazione: (2023)
CBVLM: Training-free Explainable Concept-based Large Vision Language Models for Medical Image Classification
di: Patrício, Cristiano, et al.
Pubblicazione: (2025)
di: Patrício, Cristiano, et al.
Pubblicazione: (2025)
A Concept-based Interpretable Model for the Diagnosis of Choroid Neoplasias using Multimodal Data
di: Wu, Yifan, et al.
Pubblicazione: (2024)
di: Wu, Yifan, et al.
Pubblicazione: (2024)
AiSciVision: A Framework for Specializing Large Multimodal Models in Scientific Image Classification
di: Hogan, Brendan, et al.
Pubblicazione: (2024)
di: Hogan, Brendan, et al.
Pubblicazione: (2024)
Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs
di: Luo, Jinqi, et al.
Pubblicazione: (2026)
di: Luo, Jinqi, et al.
Pubblicazione: (2026)
Improved Baselines for Data-efficient Perceptual Augmentation of LLMs
di: Vallaeys, Théophane, et al.
Pubblicazione: (2024)
di: Vallaeys, Théophane, et al.
Pubblicazione: (2024)
A Survey on Multimodal Large Language Models
di: Yin, Shukang, et al.
Pubblicazione: (2023)
di: Yin, Shukang, et al.
Pubblicazione: (2023)
FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model
di: Lee, Yebin, et al.
Pubblicazione: (2024)
di: Lee, Yebin, et al.
Pubblicazione: (2024)
Compositional Chain-of-Thought Prompting for Large Multimodal Models
di: Mitra, Chancharik, et al.
Pubblicazione: (2023)
di: Mitra, Chancharik, et al.
Pubblicazione: (2023)
Visual Question Decomposition on Multimodal Large Language Models
di: Zhang, Haowei, et al.
Pubblicazione: (2024)
di: Zhang, Haowei, et al.
Pubblicazione: (2024)
Woodpecker: Hallucination Correction for Multimodal Large Language Models
di: Yin, Shukang, et al.
Pubblicazione: (2023)
di: Yin, Shukang, et al.
Pubblicazione: (2023)
Promptception: How Sensitive Are Large Multimodal Models to Prompts?
di: Ismithdeen, Mohamed Insaf, et al.
Pubblicazione: (2025)
di: Ismithdeen, Mohamed Insaf, et al.
Pubblicazione: (2025)
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models
di: Wang, Hengyi, et al.
Pubblicazione: (2024)
di: Wang, Hengyi, et al.
Pubblicazione: (2024)
Ovis: Structural Embedding Alignment for Multimodal Large Language Model
di: Lu, Shiyin, et al.
Pubblicazione: (2024)
di: Lu, Shiyin, et al.
Pubblicazione: (2024)
MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
di: Yu, Weihao, et al.
Pubblicazione: (2023)
di: Yu, Weihao, et al.
Pubblicazione: (2023)
Towards Visual Text Grounding of Multimodal Large Language Model
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
Unsolvable Problem Detection: Robust Understanding Evaluation for Large Multimodal Models
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2024)
di: Miyai, Atsuyuki, et al.
Pubblicazione: (2024)
MM-Instruct: Generated Visual Instructions for Large Multimodal Model Alignment
di: Liu, Jihao, et al.
Pubblicazione: (2024)
di: Liu, Jihao, et al.
Pubblicazione: (2024)
Benchmarking the Thinking Mode of Multimodal Large Language Models in Clinical Tasks
di: Hong, Jindong, et al.
Pubblicazione: (2025)
di: Hong, Jindong, et al.
Pubblicazione: (2025)
Groma: Localized Visual Tokenization for Grounding Multimodal Large Language Models
di: Ma, Chuofan, et al.
Pubblicazione: (2024)
di: Ma, Chuofan, et al.
Pubblicazione: (2024)
Rethinking Visual Prompting for Multimodal Large Language Models with External Knowledge
di: Lin, Yuanze, et al.
Pubblicazione: (2024)
di: Lin, Yuanze, et al.
Pubblicazione: (2024)
mDPO: Conditional Preference Optimization for Multimodal Large Language Models
di: Wang, Fei, et al.
Pubblicazione: (2024)
di: Wang, Fei, et al.
Pubblicazione: (2024)
A Review of Multimodal Explainable Artificial Intelligence: Past, Present and Future
di: Sun, Shilin, et al.
Pubblicazione: (2024)
di: Sun, Shilin, et al.
Pubblicazione: (2024)
Uncertainty Quantification for Multimodal Large Language Models with Incoherence-adjusted Semantic Volume
di: Lau, Gregory Kang Ruey, et al.
Pubblicazione: (2026)
di: Lau, Gregory Kang Ruey, et al.
Pubblicazione: (2026)
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
di: Yiu, Eunice, et al.
Pubblicazione: (2024)
di: Yiu, Eunice, et al.
Pubblicazione: (2024)
SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions
di: Horawalavithana, Sameera, et al.
Pubblicazione: (2023)
di: Horawalavithana, Sameera, et al.
Pubblicazione: (2023)
Modality-Balancing Preference Optimization of Large Multimodal Models by Adversarial Negative Mining
di: Liu, Chenxi, et al.
Pubblicazione: (2025)
di: Liu, Chenxi, et al.
Pubblicazione: (2025)
Robust Adaptation of Large Multimodal Models for Retrieval Augmented Hateful Meme Detection
di: Mei, Jingbiao, et al.
Pubblicazione: (2025)
di: Mei, Jingbiao, et al.
Pubblicazione: (2025)
MapIQ: Evaluating Multimodal Large Language Models for Map Question Answering
di: Srivastava, Varun, et al.
Pubblicazione: (2025)
di: Srivastava, Varun, et al.
Pubblicazione: (2025)
Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
di: Huang, Wenxuan, et al.
Pubblicazione: (2025)
di: Huang, Wenxuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Learning to Steer: Input-dependent Steering for Multimodal LLMs
di: Parekh, Jayneel, et al.
Pubblicazione: (2025) -
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
di: Khayatan, Pegah, et al.
Pubblicazione: (2026) -
Analyzing Finetuning Representation Shift for Multimodal LLMs Steering
di: Khayatan, Pegah, et al.
Pubblicazione: (2025) -
Restyling Unsupervised Concept Based Interpretable Networks with Generative Models
di: Parekh, Jayneel, et al.
Pubblicazione: (2024) -
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)