MEGL: Multimodal Explanation-Guided Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yifei, Jiang, Tianxu, Pan, Bo, Wang, Jingyu, Bai, Guangji, Zhao, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
by: Zhu, Mengdan, et al.
Published: (2025)
by: Zhu, Mengdan, et al.
Published: (2025)
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
by: Zhang, Yifei, et al.
Published: (2023)
by: Zhang, Yifei, et al.
Published: (2023)
Generative Emotion Cause Explanation in Multimodal Conversations
by: Wang, Lin, et al.
Published: (2024)
by: Wang, Lin, et al.
Published: (2024)
Do Generated Data Always Help Contrastive Learning?
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
Studying How to Efficiently and Effectively Guide Models with Explanations
by: Rao, Sukrut, et al.
Published: (2023)
by: Rao, Sukrut, et al.
Published: (2023)
Non-negative Contrastive Learning
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
Explainable Cross-Disease Reasoning for Cardiovascular Risk Assessment from Low-Dose Computed Tomography
by: Zhang, Yifei, et al.
Published: (2025)
by: Zhang, Yifei, et al.
Published: (2025)
Advancing Comprehensive Aesthetic Insight with Multi-Scale Text-Guided Self-Supervised Learning
by: Liu, Yuti, et al.
Published: (2024)
by: Liu, Yuti, et al.
Published: (2024)
Invariant Representation Guided Multimodal Sentiment Decoding with Sequential Variation Regularization
by: Xu, Guoyang, et al.
Published: (2024)
by: Xu, Guoyang, et al.
Published: (2024)
Distilled Prompt Learning for Incomplete Multimodal Survival Prediction
by: Xu, Yingxue, et al.
Published: (2025)
by: Xu, Yingxue, et al.
Published: (2025)
Towards Multi-dimensional Explanation Alignment for Medical Classification
by: Hu, Lijie, et al.
Published: (2024)
by: Hu, Lijie, et al.
Published: (2024)
A Gray-box Attack against Latent Diffusion Model-based Image Editing by Posterior Collapse
by: Guo, Zhongliang, et al.
Published: (2024)
by: Guo, Zhongliang, et al.
Published: (2024)
Clinical Domain Knowledge-Derived Template Improves Post Hoc AI Explanations in Pneumothorax Classification
by: Yuan, Han, et al.
Published: (2024)
by: Yuan, Han, et al.
Published: (2024)
QG-CoC: Question-Guided Chain-of-Captions for Large Multimodal Models
by: Kao, Kuei-Chun, et al.
Published: (2025)
by: Kao, Kuei-Chun, et al.
Published: (2025)
AEGPO: Adaptive Entropy-Guided Policy Optimization for Diffusion Models
by: Li, Yuming, et al.
Published: (2026)
by: Li, Yuming, et al.
Published: (2026)
Explanation Bottleneck Models
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
Learning to Inference Adaptively for Multimodal Large Language Models
by: Xu, Zhuoyan, et al.
Published: (2025)
by: Xu, Zhuoyan, et al.
Published: (2025)
Uncovering and Mitigating Transient Blindness in Multimodal Model Editing
by: Han, Xiaoqi, et al.
Published: (2025)
by: Han, Xiaoqi, et al.
Published: (2025)
Revisiting Multimodal KV Cache Compression: A Frequency-Domain-Guided Outlier-KV-Aware Approach
by: Yang, Yaoxin, et al.
Published: (2025)
by: Yang, Yaoxin, et al.
Published: (2025)
With Limited Data for Multimodal Alignment, Let the STRUCTURE Guide You
by: Gröger, Fabian, et al.
Published: (2025)
by: Gröger, Fabian, et al.
Published: (2025)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
by: Wang, Hengyi, et al.
Published: (2024)
by: Wang, Hengyi, et al.
Published: (2024)
Learning Robust Convolutional Neural Networks with Relevant Feature Focusing via Explanations
by: Adachi, Kazuki, et al.
Published: (2022)
by: Adachi, Kazuki, et al.
Published: (2022)
FedAFD: Multimodal Federated Learning via Adversarial Fusion and Distillation
by: Tan, Min, et al.
Published: (2026)
by: Tan, Min, et al.
Published: (2026)
ID-like Prompt Learning for Few-Shot Out-of-Distribution Detection
by: Bai, Yichen, et al.
Published: (2023)
by: Bai, Yichen, et al.
Published: (2023)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
by: Li, Mingjie, et al.
Published: (2023)
by: Li, Mingjie, et al.
Published: (2023)
Contrastive Factor Analysis
by: Duan, Zhibin, et al.
Published: (2024)
by: Duan, Zhibin, et al.
Published: (2024)
Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models
by: Miao, Yanting, et al.
Published: (2026)
by: Miao, Yanting, et al.
Published: (2026)
InfMasking: Unleashing Synergistic Information by Contrastive Multimodal Interactions
by: Wen, Liangjian, et al.
Published: (2025)
by: Wen, Liangjian, et al.
Published: (2025)
Co-domain Symmetry for Complex-Valued Deep Learning
by: Singhal, Utkarsh, et al.
Published: (2021)
by: Singhal, Utkarsh, et al.
Published: (2021)
Explaining latent representations of generative models with large multimodal models
by: Zhu, Mengdan, et al.
Published: (2024)
by: Zhu, Mengdan, et al.
Published: (2024)
When Graph meets Multimodal: Benchmarking and Meditating on Multimodal Attributed Graphs Learning
by: Yan, Hao, et al.
Published: (2024)
by: Yan, Hao, et al.
Published: (2024)
EVO-LRP: Evolutionary Optimization of LRP for Interpretable Model Explanations
by: Zhang, Emerald, et al.
Published: (2025)
by: Zhang, Emerald, et al.
Published: (2025)
Emergent Visual Grounding in Large Multimodal Models Without Grounding Supervision
by: Cao, Shengcao, et al.
Published: (2024)
by: Cao, Shengcao, et al.
Published: (2024)
Leveraging Local Structure for Improving Model Explanations: An Information Propagation Approach
by: Yang, Ruo, et al.
Published: (2024)
by: Yang, Ruo, et al.
Published: (2024)
Explaining and Mitigating the Modality Gap in Contrastive Multimodal Learning
by: Yaras, Can, et al.
Published: (2024)
by: Yaras, Can, et al.
Published: (2024)
Guaranteed Optimal Compositional Explanations for Neurons
by: La Rosa, Biagio, et al.
Published: (2025)
by: La Rosa, Biagio, et al.
Published: (2025)
Transferability-Guided Cross-Domain Cross-Task Transfer Learning
by: Tan, Yang, et al.
Published: (2022)
by: Tan, Yang, et al.
Published: (2022)
Adaptive Interactive Segmentation for Multimodal Medical Imaging via Selection Engine
by: Li, Zhi, et al.
Published: (2024)
by: Li, Zhi, et al.
Published: (2024)
Deep Multimodal Learning with Missing Modality: A Survey
by: Wu, Renjie, et al.
Published: (2024)
by: Wu, Renjie, et al.
Published: (2024)
Reason--Imagine--Act: Closed-Loop LLM Decision Making with World Models for Autonomous Driving
by: Sun, Zhengqi, et al.
Published: (2026)
by: Sun, Zhengqi, et al.
Published: (2026)
Similar Items
-
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
by: Zhu, Mengdan, et al.
Published: (2025) -
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
by: Zhang, Yifei, et al.
Published: (2023) -
Generative Emotion Cause Explanation in Multimodal Conversations
by: Wang, Lin, et al.
Published: (2024) -
Do Generated Data Always Help Contrastive Learning?
by: Wang, Yifei, et al.
Published: (2024) -
Studying How to Efficiently and Effectively Guide Models with Explanations
by: Rao, Sukrut, et al.
Published: (2023)