Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Yuxiang, Sun, Jingdong, Cheng, Zhi-Qi, Wang, Jue, Liang, Haomin, Cheng, Zebang, Dong, Yifei, He, Jun-Yan, Peng, Xiaojiang, Hua, Xian-Sheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024)
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024)
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective Model
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
UCDR-Adapter: Exploring Adaptation of Pre-Trained Vision-Language Models for Universal Cross-Domain Retrieval
von: Jiang, Haoyu, et al.
Veröffentlicht: (2024)
von: Jiang, Haoyu, et al.
Veröffentlicht: (2024)
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
Voices, Faces, and Feelings: Multi-modal Emotion-Cognition Captioning for Mental Health Understanding
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2026)
MM-InstructEval: Zero-Shot Evaluation of (Multimodal) Large Language Models on Multimodal Reasoning Tasks
von: Yang, Xiaocui, et al.
Veröffentlicht: (2024)
von: Yang, Xiaocui, et al.
Veröffentlicht: (2024)
AV-EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Omni-modal LLMS with Audio-visual Cues
von: Zhou, Dingkun, et al.
Veröffentlicht: (2025)
von: Zhou, Dingkun, et al.
Veröffentlicht: (2025)
KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News Detection
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
von: Zhu, Peican, et al.
Veröffentlicht: (2025)
EmotionTalk: An Interactive Chinese Multimodal Emotion Dataset With Rich Annotations
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
Evolutionary Multimodal Reasoning via Hierarchical Semantic Representation for Intent Recognition
von: Zhou, Qianrui, et al.
Veröffentlicht: (2026)
von: Zhou, Qianrui, et al.
Veröffentlicht: (2026)
Towards Multimodal Emotional Support Conversation Systems
von: Chu, Yuqi, et al.
Veröffentlicht: (2024)
von: Chu, Yuqi, et al.
Veröffentlicht: (2024)
Multimodal Emotion Recognition with Large Language Models
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
CMATH: Cross-Modality Augmented Transformer with Hierarchical Variational Distillation for Multimodal Emotion Recognition in Conversation
von: Zhu, Xiaofei, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaofei, et al.
Veröffentlicht: (2024)
Virbo: Multimodal Multilingual Avatar Video Generation in Digital Marketing
von: Zhang, Juan, et al.
Veröffentlicht: (2024)
von: Zhang, Juan, et al.
Veröffentlicht: (2024)
Breaking the Curse of Knowledge: Towards Effective Multimodal Recommendation using Knowledge Soft Integration
von: Ouyang, Kai, et al.
Veröffentlicht: (2023)
von: Ouyang, Kai, et al.
Veröffentlicht: (2023)
SCI-Reason: A Dataset with Chain-of-Thought Rationales for Complex Multimodal Reasoning in Academic Areas
von: Ma, Chenghao, et al.
Veröffentlicht: (2025)
von: Ma, Chenghao, et al.
Veröffentlicht: (2025)
Emotional Cues Extraction and Fusion for Multi-modal Emotion Prediction and Recognition in Conversation
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
Let the Model Learn to Feel: Mode-Guided Tonality Injection for Symbolic Music Emotion Recognition
von: Xia, Haiying, et al.
Veröffentlicht: (2025)
von: Xia, Haiying, et al.
Veröffentlicht: (2025)
Angle-Optimized Partial Disentanglement for Multimodal Emotion Recognition in Conversation
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
LLM-Guided Semantic Relational Reasoning for Multimodal Intent Recognition
von: Zhou, Qianrui, et al.
Veröffentlicht: (2025)
von: Zhou, Qianrui, et al.
Veröffentlicht: (2025)
Multimodal LLM-based Query Paraphrasing for Video Search
von: Wu, Jiaxin, et al.
Veröffentlicht: (2024)
von: Wu, Jiaxin, et al.
Veröffentlicht: (2024)
MInD: Improving Multimodal Sentiment Analysis via Multimodal Information Disentanglement
von: Dai, Weichen, et al.
Veröffentlicht: (2024)
von: Dai, Weichen, et al.
Veröffentlicht: (2024)
Orthogonal Disentanglement with Projected Feature Alignment for Multimodal Emotion Recognition in Conversation
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
von: Che, Xinyi, et al.
Veröffentlicht: (2025)
Bridging Discrete and Continuous: A Multimodal Strategy for Complex Emotion Detection
von: Jia, Jiehui, et al.
Veröffentlicht: (2024)
von: Jia, Jiehui, et al.
Veröffentlicht: (2024)
MMC: Iterative Refinement of VLM Reasoning via MCTS-based Multimodal Critique
von: Liu, Shuhang, et al.
Veröffentlicht: (2025)
von: Liu, Shuhang, et al.
Veröffentlicht: (2025)
Hybrid CNN-Mamba Enhancement Network for Robust Multimodal Sentiment Analysis
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Multimodal Fusion via Hypergraph Autoencoder and Contrastive Learning for Emotion Recognition in Conversation
von: Yi, Zijian, et al.
Veröffentlicht: (2024)
von: Yi, Zijian, et al.
Veröffentlicht: (2024)
State-Anchored Complete-View Distillation for Robust Conversational Multimodal Emotion Recognition
von: Pan, Zhaoyan, et al.
Veröffentlicht: (2026)
von: Pan, Zhaoyan, et al.
Veröffentlicht: (2026)
Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion Recognition
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
Dark Side of Modalities: Reinforced Multimodal Distillation for Multimodal Knowledge Graph Reasoning
von: Zhao, Yu, et al.
Veröffentlicht: (2025)
von: Zhao, Yu, et al.
Veröffentlicht: (2025)
Robust Steganography with Boundary-Preserving Overflow Alleviation and Adaptive Error Correction
von: Cheng, Yu, et al.
Veröffentlicht: (2024)
von: Cheng, Yu, et al.
Veröffentlicht: (2024)
PRM-BAS: Enhancing Multimodal Reasoning through PRM-guided Beam Annealing Search
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
Multimodal Classification and Out-of-distribution Detection for Multimodal Intent Understanding
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
Divide and Conquer: Multimodal Video Deepfake Detection via Cross-Modal Fusion and Localization
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024) -
UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts
von: Cheng, Zhi-Qi, et al.
Veröffentlicht: (2024) -
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025)