Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Zebang, Cheng, Zhi-Qi, He, Jun-Yan, Sun, Jingdong, Wang, Kai, Lin, Yuxiang, Lian, Zheng, Peng, Xiaojiang, Hauptmann, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024)
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024)
MIDI-LLaMA: An Instruction-Following Multimodal LLM for Symbolic Music Understanding
von: Yang, Meng, et al.
Veröffentlicht: (2026)
von: Yang, Meng, et al.
Veröffentlicht: (2026)
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
Emotion-LLaMAv2 and MMEVerse: A New Framework and Benchmark for Multimodal Emotion Understanding
von: Peng, Xiaojiang, et al.
Veröffentlicht: (2026)
von: Peng, Xiaojiang, et al.
Veröffentlicht: (2026)
MMS-LLaMA: Efficient LLM-based Audio-Visual Speech Recognition with Minimal Multimodal Speech Tokens
von: Yeo, Jeong Hun, et al.
Veröffentlicht: (2025)
von: Yeo, Jeong Hun, et al.
Veröffentlicht: (2025)
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
Whispering LLaMA: A Cross-Modal Generative Error Correction Framework for Speech Recognition
von: Radhakrishnan, Srijith, et al.
Veröffentlicht: (2023)
von: Radhakrishnan, Srijith, et al.
Veröffentlicht: (2023)
Multimodal Emotion Recognition with Large Language Models
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective Model
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
Emotional Cues Extraction and Fusion for Multi-modal Emotion Prediction and Recognition in Conversation
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
von: Liu, Shansong, et al.
Veröffentlicht: (2024)
von: Liu, Shansong, et al.
Veröffentlicht: (2024)
Angle-Optimized Partial Disentanglement for Multimodal Emotion Recognition in Conversation
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
von: Xing, Bohao, et al.
Veröffentlicht: (2024)
von: Xing, Bohao, et al.
Veröffentlicht: (2024)
LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
von: Zhang, Renrui, et al.
Veröffentlicht: (2023)
von: Zhang, Renrui, et al.
Veröffentlicht: (2023)
EmotionTalk: An Interactive Chinese Multimodal Emotion Dataset With Rich Annotations
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
Orthogonal Disentanglement with Projected Feature Alignment for Multimodal Emotion Recognition in Conversation
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
CMATH: Cross-Modality Augmented Transformer with Hierarchical Variational Distillation for Multimodal Emotion Recognition in Conversation
von: Zhu, Xiaofei, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaofei, et al.
Veröffentlicht: (2024)
XEmoGPT: An Explainable Multimodal Emotion Recognition Framework with Cue-Level Perception and Reasoning
von: Zhang, Hanwen, et al.
Veröffentlicht: (2026)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2026)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
ITEACH-Net: Inverted Teacher-studEnt seArCH Network for Emotion Recognition in Conversation
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
Multimodal Fusion via Hypergraph Autoencoder and Contrastive Learning for Emotion Recognition in Conversation
von: Yi, Zijian, et al.
Veröffentlicht: (2024)
von: Yi, Zijian, et al.
Veröffentlicht: (2024)
State-Anchored Complete-View Distillation for Robust Conversational Multimodal Emotion Recognition
von: Pan, Zhaoyan, et al.
Veröffentlicht: (2026)
von: Pan, Zhaoyan, et al.
Veröffentlicht: (2026)
Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion Recognition
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
TraveLLaMA: A Multimodal Travel Assistant with Large-Scale Dataset and Structured Reasoning
von: Chu, Meng, et al.
Veröffentlicht: (2025)
von: Chu, Meng, et al.
Veröffentlicht: (2025)
AV-EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Omni-modal LLMS with Audio-visual Cues
von: Zhou, Dingkun, et al.
Veröffentlicht: (2025)
von: Zhou, Dingkun, et al.
Veröffentlicht: (2025)
A Survey on Multimodal Music Emotion Recognition
von: Liyanarachchi, Rashini, et al.
Veröffentlicht: (2025)
von: Liyanarachchi, Rashini, et al.
Veröffentlicht: (2025)
Multimodal Emotion Recognition by Fusing Video Semantic in MOOC Learning Scenarios
von: Zhang, Yuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuan, et al.
Veröffentlicht: (2024)
Towards Multimodal Emotional Support Conversation Systems
von: Chu, Yuqi, et al.
Veröffentlicht: (2024)
von: Chu, Yuqi, et al.
Veröffentlicht: (2024)
Modality-Aware Contrastive and Uncertainty-Regularized Emotion Recognition
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News Detection
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
Cross-Space Synergy: A Unified Framework for Multimodal Emotion Recognition in Conversation
von: Lyu, Xiaosen, et al.
Veröffentlicht: (2025)
von: Lyu, Xiaosen, et al.
Veröffentlicht: (2025)
HADUA: Hierarchical Attention and Dynamic Uniform Alignment for Robust Cross-Subject Emotion Recognition
von: Tang, Jiahao, et al.
Veröffentlicht: (2026)
von: Tang, Jiahao, et al.
Veröffentlicht: (2026)
Multimodal Emotion Recognition from Raw Audio with Sinc-convolution
von: Zhang, Xiaohui, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaohui, et al.
Veröffentlicht: (2024)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
von: Yan, Xueming, et al.
Veröffentlicht: (2025)
von: Yan, Xueming, et al.
Veröffentlicht: (2025)
Calibrating Multimodal Consensus for Emotion Recognition
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025) -
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024) -
MIDI-LLaMA: An Instruction-Following Multimodal LLM for Symbolic Music Understanding
von: Yang, Meng, et al.
Veröffentlicht: (2026)