MME-Emotion: A Holistic Evaluation Benchmark for Emotional Intelligence in Multimodal Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Fan, Cheng, Zebang, Deng, Chong, Li, Haoxuan, Lian, Zheng, Chen, Qian, Liu, Huadai, Wang, Wen, Zhang, Yi-Fan, Zhang, Renrui, Guo, Ziyu, Zhu, Zhihong, Wu, Hao, Wang, Haixin, Zheng, Yefeng, Peng, Xiaojiang, Wu, Xian, Wang, Kun, Li, Xiangang, Ye, Jieping, Heng, Pheng-Ann |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
CellVerse: Do Large Language Models Really Understand Cell Biology?
par: Zhang, Fan, et autres
Publié: (2025)
par: Zhang, Fan, et autres
Publié: (2025)
Rethinking Facial Expression Recognition in the Era of Multimodal Large Language Models: Benchmark, Datasets, and Beyond
par: Zhang, Fan, et autres
Publié: (2025)
par: Zhang, Fan, et autres
Publié: (2025)
Are Video Models Ready as Zero-Shot Reasoners? An Empirical Study with the MME-CoF Benchmark
par: Guo, Ziyu, et autres
Publié: (2025)
par: Guo, Ziyu, et autres
Publié: (2025)
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
par: Lian, Zheng, et autres
Publié: (2025)
par: Lian, Zheng, et autres
Publié: (2025)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
par: Cheng, Zebang, et autres
Publié: (2024)
par: Cheng, Zebang, et autres
Publié: (2024)
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
par: Huang, Dawei, et autres
Publié: (2025)
par: Huang, Dawei, et autres
Publié: (2025)
MER 2026: From Discriminative Emotion Recognition to Generative Emotion Understanding
par: Lian, Zheng, et autres
Publié: (2026)
par: Lian, Zheng, et autres
Publié: (2026)
AffectGPT-RL: Revealing Roles of Reinforcement Learning in Open-Vocabulary Emotion Recognition
par: Lian, Zheng, et autres
Publié: (2026)
par: Lian, Zheng, et autres
Publié: (2026)
Emotion-LLaMAv2 and MMEVerse: A New Framework and Benchmark for Multimodal Emotion Understanding
par: Peng, Xiaojiang, et autres
Publié: (2026)
par: Peng, Xiaojiang, et autres
Publié: (2026)
Mamba-Enhanced Text-Audio-Video Alignment Network for Emotion Recognition in Conversations
par: Li, Xinran, et autres
Publié: (2024)
par: Li, Xinran, et autres
Publié: (2024)
MME-CoF-Pro: Evaluating Reasoning Coherence in Video Generative Models with Text and Visual Hints
par: Qi, Yu, et autres
Publié: (2026)
par: Qi, Yu, et autres
Publié: (2026)
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
par: Cheng, Zebang, et autres
Publié: (2024)
par: Cheng, Zebang, et autres
Publié: (2024)
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
par: Cheng, Zebang, et autres
Publié: (2024)
par: Cheng, Zebang, et autres
Publié: (2024)
EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs
par: Hu, He, et autres
Publié: (2026)
par: Hu, He, et autres
Publié: (2026)
MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency
par: Jiang, Dongzhi, et autres
Publié: (2025)
par: Jiang, Dongzhi, et autres
Publié: (2025)
Point Cloud Understanding via Attention-Driven Contrastive Learning
par: Wang, Yi, et autres
Publié: (2024)
par: Wang, Yi, et autres
Publié: (2024)
MM-Mixing: Multi-Modal Mixing Alignment for 3D Understanding
par: Wang, Jiaze, et autres
Publié: (2024)
par: Wang, Jiaze, et autres
Publié: (2024)
SAM2Point: Segment Any 3D as Videos in Zero-shot and Promptable Manners
par: Guo, Ziyu, et autres
Publié: (2024)
par: Guo, Ziyu, et autres
Publié: (2024)
Thinking-while-Generating: Interleaving Textual Reasoning throughout Visual Generation
par: Guo, Ziyu, et autres
Publié: (2025)
par: Guo, Ziyu, et autres
Publié: (2025)
UDDETTS: Unifying Discrete and Dimensional Emotions for Controllable Emotional Text-to-Speech
par: Liu, Jiaxuan, et autres
Publié: (2025)
par: Liu, Jiaxuan, et autres
Publié: (2025)
Biomedical Entity Linking as Multiple Choice Question Answering
par: Lin, Zhenxi, et autres
Publié: (2024)
par: Lin, Zhenxi, et autres
Publié: (2024)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
par: Lin, Yuxiang, et autres
Publié: (2025)
par: Lin, Yuxiang, et autres
Publié: (2025)
PrismAudio: Decomposed Chain-of-Thoughts and Multi-dimensional Rewards for Video-to-Audio Generation
par: Liu, Huadai, et autres
Publié: (2025)
par: Liu, Huadai, et autres
Publié: (2025)
Two in One Go: Single-stage Emotion Recognition with Decoupled Subject-context Transformer
par: Li, Xinpeng, et autres
Publié: (2024)
par: Li, Xinpeng, et autres
Publié: (2024)
AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models
par: Lian, Zheng, et autres
Publié: (2025)
par: Lian, Zheng, et autres
Publié: (2025)
AffectGPT-R1: Leveraging Reinforcement Learning for Open-Vocabulary Multimodal Emotion Recognition
par: Lian, Zheng, et autres
Publié: (2025)
par: Lian, Zheng, et autres
Publié: (2025)
MME-VideoOCR: Evaluating OCR-Based Capabilities of Multimodal LLMs in Video Scenarios
par: Shi, Yang, et autres
Publié: (2025)
par: Shi, Yang, et autres
Publié: (2025)
Are Emotion and Rhetoric Neurons in LLM? Neuron Recognition and Adaptive Masking for Emotion-Rhetoric Prediction Steering
par: Zheng, Li, et autres
Publié: (2026)
par: Zheng, Li, et autres
Publié: (2026)
EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models
par: Hu, He, et autres
Publié: (2025)
par: Hu, He, et autres
Publié: (2025)
PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video Prediction
par: Wu, Hao, et autres
Publié: (2023)
par: Wu, Hao, et autres
Publié: (2023)
Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models
par: Liu, Yuansen, et autres
Publié: (2025)
par: Liu, Yuansen, et autres
Publié: (2025)
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
par: Zhang, Yi-Fan, et autres
Publié: (2024)
par: Zhang, Yi-Fan, et autres
Publié: (2024)
Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO
par: Tong, Chengzhuo, et autres
Publié: (2025)
par: Tong, Chengzhuo, et autres
Publié: (2025)
FourierFlow: Frequency-aware Flow Matching for Generative Turbulence Modeling
par: Wang, Haixin, et autres
Publié: (2025)
par: Wang, Haixin, et autres
Publié: (2025)
FGGM: Fisher-Guided Gradient Masking for Continual Learning
par: Tan, Chao-Hong, et autres
Publié: (2026)
par: Tan, Chao-Hong, et autres
Publié: (2026)
DialogueLLM: Context and Emotion Knowledge-Tuned Large Language Models for Emotion Recognition in Conversations
par: Zhang, Yazhou, et autres
Publié: (2023)
par: Zhang, Yazhou, et autres
Publié: (2023)
MedKP: Medical Dialogue with Knowledge Enhancement and Clinical Pathway Encoding
par: Wu, Jiageng, et autres
Publié: (2024)
par: Wu, Jiageng, et autres
Publié: (2024)
EmoDiffGes: Emotion‐Aware Co‐Speech Holistic Gesture Generation with Progressive Synergistic Diffusion
par: Xinru Li, et autres
Publié: (2025)
par: Xinru Li, et autres
Publié: (2025)
T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT
par: Jiang, Dongzhi, et autres
Publié: (2025)
par: Jiang, Dongzhi, et autres
Publié: (2025)
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs
par: Yuan, Jiakang, et autres
Publié: (2025)
par: Yuan, Jiakang, et autres
Publié: (2025)
Documents similaires
-
CellVerse: Do Large Language Models Really Understand Cell Biology?
par: Zhang, Fan, et autres
Publié: (2025) -
Rethinking Facial Expression Recognition in the Era of Multimodal Large Language Models: Benchmark, Datasets, and Beyond
par: Zhang, Fan, et autres
Publié: (2025) -
Are Video Models Ready as Zero-Shot Reasoners? An Empirical Study with the MME-CoF Benchmark
par: Guo, Ziyu, et autres
Publié: (2025) -
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
par: Lian, Zheng, et autres
Publié: (2025) -
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
par: Cheng, Zebang, et autres
Publié: (2024)