Multimodal Emotion Recognition by Fusing Video Semantic in MOOC Learning Scenarios
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yuan, Tao, Xiaomei, Ai, Hanxu, Chen, Tao, Gan, Yanling |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AMB-DSGDN: Adaptive Modality-Balanced Dynamic Semantic Graph Differential Network for Multimodal Emotion Recognition
di: Wang, Yunsheng, et al.
Pubblicazione: (2026)
di: Wang, Yunsheng, et al.
Pubblicazione: (2026)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
di: Cheng, Zebang, et al.
Pubblicazione: (2024)
di: Cheng, Zebang, et al.
Pubblicazione: (2024)
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025)
di: Zhu, Peican, et al.
Pubblicazione: (2025)
LLM-Guided Semantic Relational Reasoning for Multimodal Intent Recognition
di: Zhou, Qianrui, et al.
Pubblicazione: (2025)
di: Zhou, Qianrui, et al.
Pubblicazione: (2025)
Knowledge-enhanced Multi-perspective Video Representation Learning for Scene Recognition
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
di: Peng, Cheng, et al.
Pubblicazione: (2023)
di: Peng, Cheng, et al.
Pubblicazione: (2023)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
di: Yan, Xueming, et al.
Pubblicazione: (2025)
di: Yan, Xueming, et al.
Pubblicazione: (2025)
Towards Open-Vocabulary Video Semantic Segmentation
di: Li, Xinhao, et al.
Pubblicazione: (2024)
di: Li, Xinhao, et al.
Pubblicazione: (2024)
Memo2496: Expert-Annotated Dataset and Dual-View Adaptive Framework for Music Emotion Recognition
di: Li, Qilin, et al.
Pubblicazione: (2025)
di: Li, Qilin, et al.
Pubblicazione: (2025)
Semantic-Guided Unsupervised Video Summarization
di: Liu, Haizhou, et al.
Pubblicazione: (2026)
di: Liu, Haizhou, et al.
Pubblicazione: (2026)
CLCR: Cross-Level Semantic Collaborative Representation for Multimodal Learning
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
Let the Model Learn to Feel: Mode-Guided Tonality Injection for Symbolic Music Emotion Recognition
di: Xia, Haiying, et al.
Pubblicazione: (2025)
di: Xia, Haiying, et al.
Pubblicazione: (2025)
KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025)
di: Zhu, Peican, et al.
Pubblicazione: (2025)
Wireless Video Semantic Communication with Decoupled Diffusion Multi-frame Compensation
di: Xie, Bingyan, et al.
Pubblicazione: (2025)
di: Xie, Bingyan, et al.
Pubblicazione: (2025)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
di: Lin, Yuxiang, et al.
Pubblicazione: (2025)
di: Lin, Yuxiang, et al.
Pubblicazione: (2025)
Tri-Subspaces Disentanglement for Multimodal Sentiment Analysis
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
Unsupervised Multimodal Clustering for Semantics Discovery in Multimodal Utterances
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
PSA-MF: Personality-Sentiment Aligned Multi-Level Fusion for Multimodal Sentiment Analysis
di: Xie, Heng, et al.
Pubblicazione: (2025)
di: Xie, Heng, et al.
Pubblicazione: (2025)
XEmoGPT: An Explainable Multimodal Emotion Recognition Framework with Cue-Level Perception and Reasoning
di: Zhang, Hanwen, et al.
Pubblicazione: (2026)
di: Zhang, Hanwen, et al.
Pubblicazione: (2026)
QMAVIS: Long Video-Audio Understanding using Fusion of Large Multimodal Models
di: Lin, Zixing, et al.
Pubblicazione: (2026)
di: Lin, Zixing, et al.
Pubblicazione: (2026)
Semantic Item Graph Enhancement for Multimodal Recommendation
di: Zhang, Xiaoxiong, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoxiong, et al.
Pubblicazione: (2025)
MindFuse: Towards GenAI Explainability in Marketing Strategy Co-Creation
di: Farseev, Aleksandr, et al.
Pubblicazione: (2025)
di: Farseev, Aleksandr, et al.
Pubblicazione: (2025)
Enhancing Modal Fusion by Alignment and Label Matching for Multimodal Emotion Recognition
di: Li, Qifei, et al.
Pubblicazione: (2024)
di: Li, Qifei, et al.
Pubblicazione: (2024)
Sync-TVA: A Graph-Attention Framework for Multimodal Emotion Recognition with Cross-Modal Fusion
di: Deng, Zeyu, et al.
Pubblicazione: (2025)
di: Deng, Zeyu, et al.
Pubblicazione: (2025)
Temporal-Spatial Decouple before Act: Disentangled Representation Learning for Multimodal Sentiment Analysis
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
A Survey on Multimodal Benchmarks: In the Era of Large AI Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
di: Wang, Dali, et al.
Pubblicazione: (2026)
di: Wang, Dali, et al.
Pubblicazione: (2026)
AIMDiT: Modality Augmentation and Interaction via Multimodal Dimension Transformation for Emotion Recognition in Conversations
di: Wu, Sheng, et al.
Pubblicazione: (2024)
di: Wu, Sheng, et al.
Pubblicazione: (2024)
SemEval-2024 Task 3: Multimodal Emotion Cause Analysis in Conversations
di: Wang, Fanfan, et al.
Pubblicazione: (2024)
di: Wang, Fanfan, et al.
Pubblicazione: (2024)
The Dream Within Huang Long Cave: AI-Driven Interactive Narrative for Family Storytelling and Emotional Reflection
di: Huang, Jiayang, et al.
Pubblicazione: (2025)
di: Huang, Jiayang, et al.
Pubblicazione: (2025)
Has Multimodal Learning Delivered Universal Intelligence in Healthcare? A Comprehensive Survey
di: Lin, Qika, et al.
Pubblicazione: (2024)
di: Lin, Qika, et al.
Pubblicazione: (2024)
Conquering High Packet-Loss Erasure: MoE Swin Transformer-Based Video Semantic Communication
di: Teng, Lei, et al.
Pubblicazione: (2025)
di: Teng, Lei, et al.
Pubblicazione: (2025)
LL-GABR: Energy Efficient Live Video Streaming Using Reinforcement Learning
di: Raman, Adithya, et al.
Pubblicazione: (2024)
di: Raman, Adithya, et al.
Pubblicazione: (2024)
Bernini: Latent Semantic Planning for Video Diffusion
di: Bernini Team, et al.
Pubblicazione: (2026)
di: Bernini Team, et al.
Pubblicazione: (2026)
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
di: Lian, Niu, et al.
Pubblicazione: (2026)
di: Lian, Niu, et al.
Pubblicazione: (2026)
Differential Multimodal Transformers
di: Li, Jerry, et al.
Pubblicazione: (2025)
di: Li, Jerry, et al.
Pubblicazione: (2025)
DRKF: Decoupled Representations with Knowledge Fusion for Multimodal Emotion Recognition
di: Jiang, Peiyuan, et al.
Pubblicazione: (2025)
di: Jiang, Peiyuan, et al.
Pubblicazione: (2025)
Uncertainty-Aware 3D Emotional Talking Face Synthesis with Emotion Prior Distillation
di: Shen, Nanhan, et al.
Pubblicazione: (2026)
di: Shen, Nanhan, et al.
Pubblicazione: (2026)
Early Joint Learning of Emotion Information Makes MultiModal Model Understand You Better
di: Ge, Mengying, et al.
Pubblicazione: (2024)
di: Ge, Mengying, et al.
Pubblicazione: (2024)
End-to-End Learning-based Video Streaming Enhancement Pipeline: A Generative AI Approach
di: Artioli, Emanuele, et al.
Pubblicazione: (2025)
di: Artioli, Emanuele, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AMB-DSGDN: Adaptive Modality-Balanced Dynamic Semantic Graph Differential Network for Multimodal Emotion Recognition
di: Wang, Yunsheng, et al.
Pubblicazione: (2026) -
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
di: Cheng, Zebang, et al.
Pubblicazione: (2024) -
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025) -
LLM-Guided Semantic Relational Reasoning for Multimodal Intent Recognition
di: Zhou, Qianrui, et al.
Pubblicazione: (2025) -
Knowledge-enhanced Multi-perspective Video Representation Learning for Scene Recognition
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)