GIA-MIC: Multimodal Emotion Recognition with Gated Interactive Attention and Modality-Invariant Learning Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | He, Jiajun, Mi, Jinyi, Toda, Tomoki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PARCO: Phoneme-Augmented Robust Contextual ASR via Contrastive Entity Disambiguation
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
Hardness-Aware Dynamic Curriculum Learning for Robust Multimodal Emotion Recognition with Missing Modalities
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
2DP-2MRC: 2-Dimensional Pointer-based Machine Reading Comprehension Method for Multimodal Moment Retrieval
by: He, Jiajun, et al.
Published: (2024)
by: He, Jiajun, et al.
Published: (2024)
PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
Masked Graph Learning with Recurrent Alignment for Multimodal Emotion Recognition in Conversation
by: Meng, Tao, et al.
Published: (2024)
by: Meng, Tao, et al.
Published: (2024)
Unimodal-driven Distillation in Multimodal Emotion Recognition with Dynamic Fusion
by: Li, Jiagen, et al.
Published: (2025)
by: Li, Jiagen, et al.
Published: (2025)
M4SER: Multimodal, Multirepresentation, Multitask, and Multistrategy Learning for Speech Emotion Recognition
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
Divide and Refine: Enhancing Multimodal Representation and Explainability for Emotion Recognition in Conversation
by: Mai, Anh-Tuan, et al.
Published: (2026)
by: Mai, Anh-Tuan, et al.
Published: (2026)
Feature Fusion Based on Mutual-Cross-Attention Mechanism for EEG Emotion Recognition
by: Zhao, Yimin, et al.
Published: (2024)
by: Zhao, Yimin, et al.
Published: (2024)
Gated Adaptation for Continual Learning in Human Activity Recognition
by: Azghan, Reza Rahimi, et al.
Published: (2026)
by: Azghan, Reza Rahimi, et al.
Published: (2026)
MiMIC: Multi-Modal Indian Earnings Calls Dataset to Predict Stock Prices
by: Ghosh, Sohom, et al.
Published: (2025)
by: Ghosh, Sohom, et al.
Published: (2025)
Bridging Modalities: Knowledge Distillation and Masked Training for Translating Multi-Modal Emotion Recognition to Uni-Modal, Speech-Only Emotion Recognition
by: Muaz, Muhammad, et al.
Published: (2024)
by: Muaz, Muhammad, et al.
Published: (2024)
Mosaic of Modalities: A Comprehensive Benchmark for Multimodal Graph Learning
by: Zhu, Jing, et al.
Published: (2024)
by: Zhu, Jing, et al.
Published: (2024)
Feature-level Interaction Explanations in Multimodal Transformers
by: Kim, Yeji, et al.
Published: (2026)
by: Kim, Yeji, et al.
Published: (2026)
CogniAlign: Word-Level Multimodal Speech Alignment with Gated Cross-Attention for Alzheimer's Detection
by: Ortiz-Perez, David, et al.
Published: (2025)
by: Ortiz-Perez, David, et al.
Published: (2025)
Multimodal Functional Maximum Correlation for Emotion Recognition
by: Zheng, Deyang, et al.
Published: (2025)
by: Zheng, Deyang, et al.
Published: (2025)
MultiMax: Sparse and Multi-Modal Attention Learning
by: Zhou, Yuxuan, et al.
Published: (2024)
by: Zhou, Yuxuan, et al.
Published: (2024)
MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion
by: Huang, Haofeng, et al.
Published: (2025)
by: Huang, Haofeng, et al.
Published: (2025)
Emotion and Intention Guided Multi-Modal Learning for Sticker Response Selection
by: Hu, Yuxuan, et al.
Published: (2025)
by: Hu, Yuxuan, et al.
Published: (2025)
Flash Invariant Point Attention
by: Liu, Andrew, et al.
Published: (2025)
by: Liu, Andrew, et al.
Published: (2025)
TED: Turn Emphasis with Dialogue Feature Attention for Emotion Recognition in Conversation
by: Ono, Junya, et al.
Published: (2025)
by: Ono, Junya, et al.
Published: (2025)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
by: Yan, Xueming, et al.
Published: (2025)
by: Yan, Xueming, et al.
Published: (2025)
Learning in Order! A Sequential Strategy to Learn Invariant Features for Multimodal Sentiment Analysis
by: Zhao, Xianbing, et al.
Published: (2024)
by: Zhao, Xianbing, et al.
Published: (2024)
Hierarchical Hypercomplex Network for Multimodal Emotion Recognition
by: Lopez, Eleonora, et al.
Published: (2024)
by: Lopez, Eleonora, et al.
Published: (2024)
GSINA: Improving Subgraph Extraction for Graph Invariant Learning via Graph Sinkhorn Attention
by: Yan, Junchi, et al.
Published: (2024)
by: Yan, Junchi, et al.
Published: (2024)
Generalizable Multimodal Large Language Model Editing via Invariant Trajectory Learning
by: Su, Jiajie, et al.
Published: (2026)
by: Su, Jiajie, et al.
Published: (2026)
Sparse-by-Design Cross-Modality Prediction: L0-Gated Representations for Reliable and Efficient Learning
by: Cenacchi, Filippo
Published: (2026)
by: Cenacchi, Filippo
Published: (2026)
Rethinking Gating Mechanism in Sparse MoE: Handling Arbitrary Modality Inputs with Confidence-Guided Gate
by: Zheng, Liangwei Nathan, et al.
Published: (2025)
by: Zheng, Liangwei Nathan, et al.
Published: (2025)
Forgetting Transformer: Softmax Attention with a Forget Gate
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Causal Debiasing Medical Multimodal Representation Learning with Missing Modalities
by: Zhu, Xiaoguang, et al.
Published: (2025)
by: Zhu, Xiaoguang, et al.
Published: (2025)
Learning Domain- and Class-Disentangled Prototypes for Domain-Generalized EEG Emotion Recognition
by: Li, Guangli, et al.
Published: (2025)
by: Li, Guangli, et al.
Published: (2025)
Magnitude and Rotation Invariant Detection of Transportation Modes with Missing Data Modalities
by: Van Der Donckt, Jeroen, et al.
Published: (2024)
by: Van Der Donckt, Jeroen, et al.
Published: (2024)
Robot-Gated Interactive Imitation Learning with Adaptive Intervention Mechanism
by: Cai, Haoyuan, et al.
Published: (2025)
by: Cai, Haoyuan, et al.
Published: (2025)
MultiModal-Learning for Predicting Molecular Properties: A Framework Based on Image and Graph Structures
by: Wang, Zhuoyuan, et al.
Published: (2023)
by: Wang, Zhuoyuan, et al.
Published: (2023)
Towards Robust Federated Multimodal Graph Learning under Modality Heterogeneity
by: Zhang, Sirui, et al.
Published: (2026)
by: Zhang, Sirui, et al.
Published: (2026)
Modality as Heterogeneity: Node Splitting and Graph Rewiring for Multimodal Graph Learning
by: Zhang, Yihan, et al.
Published: (2026)
by: Zhang, Yihan, et al.
Published: (2026)
Gating is Weighting: Understanding Gated Linear Attention through In-context Learning
by: Li, Yingcong, et al.
Published: (2025)
by: Li, Yingcong, et al.
Published: (2025)
Merge then Realign: Simple and Effective Modality-Incremental Continual Learning for Multimodal LLMs
by: Zhang, Dingkun, et al.
Published: (2025)
by: Zhang, Dingkun, et al.
Published: (2025)
Optimizing Sensory Neurons: Nonlinear Attention Mechanisms for Accelerated Convergence in Permutation-Invariant Neural Networks for Reinforcement Learning
by: Muzaffar, Junaid, et al.
Published: (2025)
by: Muzaffar, Junaid, et al.
Published: (2025)
Cross-Modal Bayesian Low-Rank Adaptation for Uncertainty-Aware Multimodal Learning
by: Naderi, Habibeh, et al.
Published: (2026)
by: Naderi, Habibeh, et al.
Published: (2026)
Similar Items
-
PARCO: Phoneme-Augmented Robust Contextual ASR via Contrastive Entity Disambiguation
by: He, Jiajun, et al.
Published: (2025) -
Hardness-Aware Dynamic Curriculum Learning for Robust Multimodal Emotion Recognition with Missing Modalities
by: Liu, Rui, et al.
Published: (2025) -
2DP-2MRC: 2-Dimensional Pointer-based Machine Reading Comprehension Method for Multimodal Moment Retrieval
by: He, Jiajun, et al.
Published: (2024) -
PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding
by: He, Jiajun, et al.
Published: (2025) -
Masked Graph Learning with Recurrent Alignment for Multimodal Emotion Recognition in Conversation
by: Meng, Tao, et al.
Published: (2024)