VISAFF: Speaker-Centered Visual Affective Feature Learning for Emotion Recognition in Conversation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | ZHU, Linan, Zhai, Zihao, Han, Xiao, Fu, Yuqian, Chen, Xiangfan, Kong, Xiangjie, Shen, Guojiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Region Representation Learning from Urban Imagery with Noisy Long-Caption Supervision
von: Zhang, Yimei, et al.
Veröffentlicht: (2025)
von: Zhang, Yimei, et al.
Veröffentlicht: (2025)
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching
von: Han, Xiao, et al.
Veröffentlicht: (2024)
von: Han, Xiao, et al.
Veröffentlicht: (2024)
Affective Flow Language Model for Emotional Support Conversation
von: Zou, Chenghui, et al.
Veröffentlicht: (2026)
von: Zou, Chenghui, et al.
Veröffentlicht: (2026)
Socratic RL: A Novel Framework for Efficient Knowledge Acquisition through Iterative Reflection and Viewpoint Distillation
von: Wu, Xiangfan
Veröffentlicht: (2025)
von: Wu, Xiangfan
Veröffentlicht: (2025)
Multimodal Trajectory Representation Learning for Travel Time Estimation
von: Liu, Zhi, et al.
Veröffentlicht: (2025)
von: Liu, Zhi, et al.
Veröffentlicht: (2025)
ML-SAN: Multi-Level Speaker-Adaptive Network for Emotion Recognition in Conversations
von: Wang, Kexue, et al.
Veröffentlicht: (2026)
von: Wang, Kexue, et al.
Veröffentlicht: (2026)
A-MBER: Affective Memory Benchmark for Emotion Recognition
von: Wen, Deliang, et al.
Veröffentlicht: (2026)
von: Wen, Deliang, et al.
Veröffentlicht: (2026)
A Comprehensive Survey on Multi-modal Conversational Emotion Recognition with Deep Learning
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
von: Shou, Yuntao, et al.
Veröffentlicht: (2023)
A Dataset for Spatiotemporal-Sensitive POI Question Answering
von: Han, Xiao, et al.
Veröffentlicht: (2025)
von: Han, Xiao, et al.
Veröffentlicht: (2025)
Centering Emotion Hotspots: Multimodal Local-Global Fusion and Cross-Modal Alignment for Emotion Recognition in Conversations
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
TED: Turn Emphasis with Dialogue Feature Attention for Emotion Recognition in Conversation
von: Ono, Junya, et al.
Veröffentlicht: (2025)
von: Ono, Junya, et al.
Veröffentlicht: (2025)
Dynamic Demonstration Retrieval and Cognitive Understanding for Emotional Support Conversation
von: Xu, Zhe, et al.
Veröffentlicht: (2024)
von: Xu, Zhe, et al.
Veröffentlicht: (2024)
Emotion Recognition in Multi-Speaker Conversations through Speaker Identification, Knowledge Distillation, and Hierarchical Fusion
von: Li, Xiao, et al.
Veröffentlicht: (2025)
von: Li, Xiao, et al.
Veröffentlicht: (2025)
SPOT-Trip: Dual-Preference Driven Out-of-Town Trip Recommendation
von: Liu, Yinghui, et al.
Veröffentlicht: (2025)
von: Liu, Yinghui, et al.
Veröffentlicht: (2025)
Disentangled Dual-Branch Graph Learning for Conversational Emotion Recognition
von: Guo, Chengling, et al.
Veröffentlicht: (2026)
von: Guo, Chengling, et al.
Veröffentlicht: (2026)
Masked Graph Learning with Recurrent Alignment for Multimodal Emotion Recognition in Conversation
von: Meng, Tao, et al.
Veröffentlicht: (2024)
von: Meng, Tao, et al.
Veröffentlicht: (2024)
SpeakerLM: End-to-End Versatile Speaker Diarization and Recognition with Multimodal Large Language Models
von: Yin, Han, et al.
Veröffentlicht: (2025)
von: Yin, Han, et al.
Veröffentlicht: (2025)
Qieemo: Speech Is All You Need in the Emotion Recognition in Conversations
von: Chen, Jinming, et al.
Veröffentlicht: (2025)
von: Chen, Jinming, et al.
Veröffentlicht: (2025)
Deep Dive Into Music Videos: Hierarchical Emotion Recognition With Rich Audio and Visual Features
von: Yagya Raj Pandeya, et al.
Veröffentlicht: (2025)
von: Yagya Raj Pandeya, et al.
Veröffentlicht: (2025)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
Exploring Multilingual Unseen Speaker Emotion Recognition: Leveraging Co-Attention Cues in Multitask Learning
von: Goel, Arnav, et al.
Veröffentlicht: (2024)
von: Goel, Arnav, et al.
Veröffentlicht: (2024)
Continuous Adversarial Text Representation Learning for Affective Recognition
von: Son, Seungah, et al.
Veröffentlicht: (2025)
von: Son, Seungah, et al.
Veröffentlicht: (2025)
Deep Emotion Recognition in Textual Conversations: A Survey
von: Pereira, Patrícia, et al.
Veröffentlicht: (2022)
von: Pereira, Patrícia, et al.
Veröffentlicht: (2022)
ECRC: Emotion-Causality Recognition in Korean Conversation for GCN
von: Lee, J. K., et al.
Veröffentlicht: (2024)
von: Lee, J. K., et al.
Veröffentlicht: (2024)
LaERC-S: Improving LLM-based Emotion Recognition in Conversation with Speaker Characteristics
von: Fu, Yumeng, et al.
Veröffentlicht: (2024)
von: Fu, Yumeng, et al.
Veröffentlicht: (2024)
Speaker Emotion Recognition: Leveraging Self-Supervised Models for Feature Extraction Using Wav2Vec2 and HuBERT
von: Jafarzadeh, Pourya, et al.
Veröffentlicht: (2024)
von: Jafarzadeh, Pourya, et al.
Veröffentlicht: (2024)
Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2024)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2024)
Layer-aware TDNN: Speaker Recognition Using Multi-Layer Features from Pre-Trained Models
von: Kim, Jin Sob, et al.
Veröffentlicht: (2024)
von: Kim, Jin Sob, et al.
Veröffentlicht: (2024)
Triple Disentangled Representation Learning for Multimodal Affective Analysis
von: Zhou, Ying, et al.
Veröffentlicht: (2024)
von: Zhou, Ying, et al.
Veröffentlicht: (2024)
Affective Multimodal Agents with Proactive Knowledge Grounding for Emotionally Aligned Marketing Dialogue
von: Yu, Lin, et al.
Veröffentlicht: (2025)
von: Yu, Lin, et al.
Veröffentlicht: (2025)
TrackAny3D: Transferring Pretrained 3D Models for Category-unified 3D Point Cloud Tracking
von: Wang, Mengmeng, et al.
Veröffentlicht: (2025)
von: Wang, Mengmeng, et al.
Veröffentlicht: (2025)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing
von: Sun, Xiao
Veröffentlicht: (2025)
von: Sun, Xiao
Veröffentlicht: (2025)
Investigating Safety Vulnerabilities of Large Audio-Language Models Under Speaker Emotional Variations
von: Feng, Bo-Han, et al.
Veröffentlicht: (2025)
von: Feng, Bo-Han, et al.
Veröffentlicht: (2025)
Divide and Refine: Enhancing Multimodal Representation and Explainability for Emotion Recognition in Conversation
von: Mai, Anh-Tuan, et al.
Veröffentlicht: (2026)
von: Mai, Anh-Tuan, et al.
Veröffentlicht: (2026)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
von: Deng, Zhaoyuan, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoyuan, et al.
Veröffentlicht: (2024)
Multi-dataset Joint Pre-training of Emotional EEG Enables Generalizable Affective Computing
von: Zhang, Qingzhu, et al.
Veröffentlicht: (2025)
von: Zhang, Qingzhu, et al.
Veröffentlicht: (2025)
Affective Visual Dialog: A Large-Scale Benchmark for Emotional Reasoning Based on Visually Grounded Conversations
von: Haydarov, Kilichbek, et al.
Veröffentlicht: (2023)
von: Haydarov, Kilichbek, et al.
Veröffentlicht: (2023)
CARE: Cognitive-reasoning Augmented Reinforcement for Emotional Support Conversation
von: Zhu, Jie, et al.
Veröffentlicht: (2025)
von: Zhu, Jie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Improving Region Representation Learning from Urban Imagery with Noisy Long-Caption Supervision
von: Zhang, Yimei, et al.
Veröffentlicht: (2025) -
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching
von: Han, Xiao, et al.
Veröffentlicht: (2024) -
Affective Flow Language Model for Emotional Support Conversation
von: Zou, Chenghui, et al.
Veröffentlicht: (2026) -
Socratic RL: A Novel Framework for Efficient Knowledge Acquisition through Iterative Reflection and Viewpoint Distillation
von: Wu, Xiangfan
Veröffentlicht: (2025) -
Multimodal Trajectory Representation Learning for Travel Time Estimation
von: Liu, Zhi, et al.
Veröffentlicht: (2025)