Emotion-LLaMAv2 and MMEVerse: A New Framework and Benchmark for Multimodal Emotion Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Xiaojiang, Chen, Jingyi, Cheng, Zebang, Peng, Bao, Wu, Fengyi, Dong, Yifei, Tu, Shuyuan, Hu, Qiyu, Huang, Huiting, Lin, Yuxiang, He, Jun-Yan, Wang, Kai, Lian, Zheng, Cheng, Zhi-Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)
AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs
von: Hu, He, et al.
Veröffentlicht: (2026)
von: Hu, He, et al.
Veröffentlicht: (2026)
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
MME-Emotion: A Holistic Evaluation Benchmark for Emotional Intelligence in Multimodal Large Language Models
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
MER 2026: From Discriminative Emotion Recognition to Generative Emotion Understanding
von: Lian, Zheng, et al.
Veröffentlicht: (2026)
von: Lian, Zheng, et al.
Veröffentlicht: (2026)
UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024)
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024)
EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models
von: Hu, He, et al.
Veröffentlicht: (2025)
von: Hu, He, et al.
Veröffentlicht: (2025)
DPDEdit: Detail-Preserved Diffusion Models for Multimodal Fashion Image Editing
von: Wang, Xiaolong, et al.
Veröffentlicht: (2024)
von: Wang, Xiaolong, et al.
Veröffentlicht: (2024)
HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions
von: Dong, Yifei, et al.
Veröffentlicht: (2025)
von: Dong, Yifei, et al.
Veröffentlicht: (2025)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective Model
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
GoViG: Goal-Conditioned Visual Navigation Instruction Generation via Multimodal Reasoning
von: Wu, Fengyi, et al.
Veröffentlicht: (2025)
von: Wu, Fengyi, et al.
Veröffentlicht: (2025)
Exploring the Relationships Among Display Rules, Emotional Job Demands, Emotional Labour and Kindergarten Teachers' Occupational Well‐Being
von: Xin Zheng, et al.
Veröffentlicht: (2024)
von: Xin Zheng, et al.
Veröffentlicht: (2024)
Beyond Emotion Recognition: A Multi-Turn Multimodal Emotion Understanding and Reasoning Benchmark
von: Hu, Jinpeng, et al.
Veröffentlicht: (2025)
von: Hu, Jinpeng, et al.
Veröffentlicht: (2025)
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
von: Xing, Bohao, et al.
Veröffentlicht: (2024)
von: Xing, Bohao, et al.
Veröffentlicht: (2024)
Benchmarking and Bridging Emotion Conflicts for Multimodal Emotion Reasoning
von: Han, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Han, Zhiyuan, et al.
Veröffentlicht: (2025)
Attribute-Grounded Selective Reasoning for Artwork Emotion Understanding with Multimodal Large Language Models
von: Zhang, Cheng, et al.
Veröffentlicht: (2026)
von: Zhang, Cheng, et al.
Veröffentlicht: (2026)
Mamba-Enhanced Text-Audio-Video Alignment Network for Emotion Recognition in Conversations
von: Li, Xinran, et al.
Veröffentlicht: (2024)
von: Li, Xinran, et al.
Veröffentlicht: (2024)
Emotion-o1: Adaptive Long Reasoning for Emotion Understanding in LLMs
von: Song, Changhao, et al.
Veröffentlicht: (2025)
von: Song, Changhao, et al.
Veröffentlicht: (2025)
MERBench: A Unified Evaluation Benchmark for Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
LEAF: Unveiling Two Sides of the Same Coin in Semi-supervised Facial Expression Recognition
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
When Tone and Words Disagree: Towards Robust Speech Emotion Recognition under Acoustic-Semantic Conflict
von: Huang, Dawei, et al.
Veröffentlicht: (2026)
von: Huang, Dawei, et al.
Veröffentlicht: (2026)
FlexEdit: Marrying Free-Shape Masks to VLLM for Flexible Image Editing
von: Yuan, Tianshuo, et al.
Veröffentlicht: (2024)
von: Yuan, Tianshuo, et al.
Veröffentlicht: (2024)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Omni-Emotion: Extending Video MLLM with Detailed Face and Audio Modeling for Multimodal Emotion Analysis
von: Yang, Qize, et al.
Veröffentlicht: (2025)
von: Yang, Qize, et al.
Veröffentlicht: (2025)
Invisible Gas Detection: An RGB-Thermal Cross Attention Network and A New Benchmark
von: Wang, Jue, et al.
Veröffentlicht: (2024)
von: Wang, Jue, et al.
Veröffentlicht: (2024)
CULEMO: Cultural Lenses on Emotion -- Benchmarking LLMs for Cross-Cultural Emotion Understanding
von: Belay, Tadesse Destaw, et al.
Veröffentlicht: (2025)
von: Belay, Tadesse Destaw, et al.
Veröffentlicht: (2025)
EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Spoken Dialogue Systems
von: Liu, Jingwen, et al.
Veröffentlicht: (2025)
von: Liu, Jingwen, et al.
Veröffentlicht: (2025)
Multi-Source EEG Emotion Recognition via Dynamic Contrastive Domain Adaptation
von: Xiao, Yun, et al.
Veröffentlicht: (2024)
von: Xiao, Yun, et al.
Veröffentlicht: (2024)
AffectGPT-RL: Revealing Roles of Reinforcement Learning in Open-Vocabulary Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2026)
von: Lian, Zheng, et al.
Veröffentlicht: (2026)
Generative Emotion Cause Explanation in Multimodal Conversations
von: Wang, Lin, et al.
Veröffentlicht: (2024)
von: Wang, Lin, et al.
Veröffentlicht: (2024)
EmoS: A High-Fidelity Multimodal Benchmark for Fine-grained Streaming Emotional Understanding
von: Guo, Pengze, et al.
Veröffentlicht: (2026)
von: Guo, Pengze, et al.
Veröffentlicht: (2026)
VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
von: Zhang, Boqiang, et al.
Veröffentlicht: (2025)
von: Zhang, Boqiang, et al.
Veröffentlicht: (2025)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025) -
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
von: Lin, Yuxiang, et al.
Veröffentlicht: (2025)