CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Cheng, Chen, Ke, Shou, Lidan, Chen, Gang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AMB-DSGDN: Adaptive Modality-Balanced Dynamic Semantic Graph Differential Network for Multimodal Emotion Recognition
von: Wang, Yunsheng, et al.
Veröffentlicht: (2026)
von: Wang, Yunsheng, et al.
Veröffentlicht: (2026)
HeLo: Heterogeneous Multi-Modal Fusion with Label Correlation for Emotion Distribution Learning
von: Zheng, Chuhang, et al.
Veröffentlicht: (2025)
von: Zheng, Chuhang, et al.
Veröffentlicht: (2025)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
Enhancing Modal Fusion by Alignment and Label Matching for Multimodal Emotion Recognition
von: Li, Qifei, et al.
Veröffentlicht: (2024)
von: Li, Qifei, et al.
Veröffentlicht: (2024)
M3TR: Temporal Retrieval Enhanced Multi-Modal Micro-video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)
Multimodal Emotion Recognition by Fusing Video Semantic in MOOC Learning Scenarios
von: Zhang, Yuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuan, et al.
Veröffentlicht: (2024)
TIP and Polish: Text-Image-Prototype Guided Multi-Modal Generation via Commonality-Discrepancy Modeling and Refinement
von: Ma, Zhiyong, et al.
Veröffentlicht: (2025)
von: Ma, Zhiyong, et al.
Veröffentlicht: (2025)
MM-HSD: Multi-Modal Hate Speech Detection in Videos
von: Céspedes-Sarrias, Berta, et al.
Veröffentlicht: (2025)
von: Céspedes-Sarrias, Berta, et al.
Veröffentlicht: (2025)
A Survey on Music Generation from Single-Modal, Cross-Modal, and Multi-Modal Perspectives
von: Li, Shuyu, et al.
Veröffentlicht: (2025)
von: Li, Shuyu, et al.
Veröffentlicht: (2025)
Memo2496: Expert-Annotated Dataset and Dual-View Adaptive Framework for Music Emotion Recognition
von: Li, Qilin, et al.
Veröffentlicht: (2025)
von: Li, Qilin, et al.
Veröffentlicht: (2025)
EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)
SimLabel: Similarity-Weighted Iterative Framework for Multi-annotator Learning with Missing Annotations
von: Zhang, Liyun, et al.
Veröffentlicht: (2025)
von: Zhang, Liyun, et al.
Veröffentlicht: (2025)
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)
Early Joint Learning of Emotion Information Makes MultiModal Model Understand You Better
von: Ge, Mengying, et al.
Veröffentlicht: (2024)
von: Ge, Mengying, et al.
Veröffentlicht: (2024)
Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs
von: Chen, Jianhao, et al.
Veröffentlicht: (2026)
von: Chen, Jianhao, et al.
Veröffentlicht: (2026)
FISHER: A Foundation Model for Multi-Modal Industrial Signal Comprehensive Representation
von: Fan, Pingyi, et al.
Veröffentlicht: (2025)
von: Fan, Pingyi, et al.
Veröffentlicht: (2025)
Unleashing the Power of Imbalanced Modality Information for Multi-modal Knowledge Graph Completion
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
Sync-TVA: A Graph-Attention Framework for Multimodal Emotion Recognition with Cross-Modal Fusion
von: Deng, Zeyu, et al.
Veröffentlicht: (2025)
von: Deng, Zeyu, et al.
Veröffentlicht: (2025)
Modality-Aware Contrastive and Uncertainty-Regularized Emotion Recognition
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
Emotion-Driven Personalized Recommendation for AI-Generated Content Using Multi-Modal Sentiment and Intent Analysis
von: Hu, Zheqi, et al.
Veröffentlicht: (2025)
von: Hu, Zheqi, et al.
Veröffentlicht: (2025)
Knowledge-enhanced Multi-perspective Video Representation Learning for Scene Recognition
von: Yu, Xuzheng, et al.
Veröffentlicht: (2024)
von: Yu, Xuzheng, et al.
Veröffentlicht: (2024)
KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News Detection
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
EigeNet: Geometry-Informed Multi-Modal Learning for Few-shot Novel View RIR Prediction
von: Jing, Chong, et al.
Veröffentlicht: (2026)
von: Jing, Chong, et al.
Veröffentlicht: (2026)
Towards Temporal-Aware Multi-Modal Retrieval Augmented Generation in Finance
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
Let the Model Learn to Feel: Mode-Guided Tonality Injection for Symbolic Music Emotion Recognition
von: Xia, Haiying, et al.
Veröffentlicht: (2025)
von: Xia, Haiying, et al.
Veröffentlicht: (2025)
AIMDiT: Modality Augmentation and Interaction via Multimodal Dimension Transformation for Emotion Recognition in Conversations
von: Wu, Sheng, et al.
Veröffentlicht: (2024)
von: Wu, Sheng, et al.
Veröffentlicht: (2024)
A Multi-modal Fusion Network for Terrain Perception Based on Illumination Aware
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Exposing Cross-Modal Consistency for Fake News Detection in Short-Form Videos
von: Tian, Chong, et al.
Veröffentlicht: (2026)
von: Tian, Chong, et al.
Veröffentlicht: (2026)
OmniDPO: A Preference Optimization Framework to Address Omni-Modal Hallucination
von: Chen, Junzhe, et al.
Veröffentlicht: (2025)
von: Chen, Junzhe, et al.
Veröffentlicht: (2025)
RMAdapter: Reconstruction-based Multi-Modal Adapter for Vision-Language Models
von: Lin, Xiang, et al.
Veröffentlicht: (2025)
von: Lin, Xiang, et al.
Veröffentlicht: (2025)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
von: Yan, Xueming, et al.
Veröffentlicht: (2025)
von: Yan, Xueming, et al.
Veröffentlicht: (2025)
Improving the Consistency in Cross-Lingual Cross-Modal Retrieval with 1-to-K Contrastive Learning
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
A Survey of Multi-sensor Fusion Perception for Embodied AI: Background, Methods, Challenges and Prospects
von: Ruan, Shulan, et al.
Veröffentlicht: (2025)
von: Ruan, Shulan, et al.
Veröffentlicht: (2025)
VCEMO: Multi-Modal Emotion Recognition for Chinese Voiceprints
von: Tang, Jinghua, et al.
Veröffentlicht: (2024)
von: Tang, Jinghua, et al.
Veröffentlicht: (2024)
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
von: Ouyang, Shuyi, et al.
Veröffentlicht: (2024)
von: Ouyang, Shuyi, et al.
Veröffentlicht: (2024)
Robust Fuzzy Multi-view Learning under View Conflict
von: Duan, Siyuan, et al.
Veröffentlicht: (2026)
von: Duan, Siyuan, et al.
Veröffentlicht: (2026)
DIRECT: Video Mashup Creation via Hierarchical Multi-Agent Planning and Intent-Guided Editing
von: Li, Ke, et al.
Veröffentlicht: (2026)
von: Li, Ke, et al.
Veröffentlicht: (2026)
Rethinking Prompting Strategies for Multi-Label Recognition with Partial Annotations
von: Rawlekar, Samyak, et al.
Veröffentlicht: (2024)
von: Rawlekar, Samyak, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AMB-DSGDN: Adaptive Modality-Balanced Dynamic Semantic Graph Differential Network for Multimodal Emotion Recognition
von: Wang, Yunsheng, et al.
Veröffentlicht: (2026) -
HeLo: Heterogeneous Multi-Modal Fusion with Label Correlation for Emotion Distribution Learning
von: Zheng, Chuhang, et al.
Veröffentlicht: (2025) -
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
Enhancing Modal Fusion by Alignment and Label Matching for Multimodal Emotion Recognition
von: Li, Qifei, et al.
Veröffentlicht: (2024) -
M3TR: Temporal Retrieval Enhanced Multi-Modal Micro-video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)