VISAFF: Speaker-Centered Visual Affective Feature Learning for Emotion Recognition in Conversation
Fuente:
arXiv
Saved in:
| Main Authors: | ZHU, Linan, Zhai, Zihao, Han, Xiao, Fu, Yuqian, Chen, Xiangfan, Kong, Xiangjie, Shen, Guojiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Region Representation Learning from Urban Imagery with Noisy Long-Caption Supervision
by: Zhang, Yimei, et al.
Published: (2025)
by: Zhang, Yimei, et al.
Published: (2025)
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching
by: Han, Xiao, et al.
Published: (2024)
by: Han, Xiao, et al.
Published: (2024)
Affective Flow Language Model for Emotional Support Conversation
by: Zou, Chenghui, et al.
Published: (2026)
by: Zou, Chenghui, et al.
Published: (2026)
Socratic RL: A Novel Framework for Efficient Knowledge Acquisition through Iterative Reflection and Viewpoint Distillation
by: Wu, Xiangfan
Published: (2025)
by: Wu, Xiangfan
Published: (2025)
Multimodal Trajectory Representation Learning for Travel Time Estimation
by: Liu, Zhi, et al.
Published: (2025)
by: Liu, Zhi, et al.
Published: (2025)
ML-SAN: Multi-Level Speaker-Adaptive Network for Emotion Recognition in Conversations
by: Wang, Kexue, et al.
Published: (2026)
by: Wang, Kexue, et al.
Published: (2026)
A-MBER: Affective Memory Benchmark for Emotion Recognition
by: Wen, Deliang, et al.
Published: (2026)
by: Wen, Deliang, et al.
Published: (2026)
A Comprehensive Survey on Multi-modal Conversational Emotion Recognition with Deep Learning
by: Shou, Yuntao, et al.
Published: (2023)
by: Shou, Yuntao, et al.
Published: (2023)
A Dataset for Spatiotemporal-Sensitive POI Question Answering
by: Han, Xiao, et al.
Published: (2025)
by: Han, Xiao, et al.
Published: (2025)
Centering Emotion Hotspots: Multimodal Local-Global Fusion and Cross-Modal Alignment for Emotion Recognition in Conversations
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
by: Wen, Zhiyuan, et al.
Published: (2024)
by: Wen, Zhiyuan, et al.
Published: (2024)
TED: Turn Emphasis with Dialogue Feature Attention for Emotion Recognition in Conversation
by: Ono, Junya, et al.
Published: (2025)
by: Ono, Junya, et al.
Published: (2025)
Dynamic Demonstration Retrieval and Cognitive Understanding for Emotional Support Conversation
by: Xu, Zhe, et al.
Published: (2024)
by: Xu, Zhe, et al.
Published: (2024)
Emotion Recognition in Multi-Speaker Conversations through Speaker Identification, Knowledge Distillation, and Hierarchical Fusion
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
SPOT-Trip: Dual-Preference Driven Out-of-Town Trip Recommendation
by: Liu, Yinghui, et al.
Published: (2025)
by: Liu, Yinghui, et al.
Published: (2025)
Disentangled Dual-Branch Graph Learning for Conversational Emotion Recognition
by: Guo, Chengling, et al.
Published: (2026)
by: Guo, Chengling, et al.
Published: (2026)
Masked Graph Learning with Recurrent Alignment for Multimodal Emotion Recognition in Conversation
by: Meng, Tao, et al.
Published: (2024)
by: Meng, Tao, et al.
Published: (2024)
SpeakerLM: End-to-End Versatile Speaker Diarization and Recognition with Multimodal Large Language Models
by: Yin, Han, et al.
Published: (2025)
by: Yin, Han, et al.
Published: (2025)
Qieemo: Speech Is All You Need in the Emotion Recognition in Conversations
by: Chen, Jinming, et al.
Published: (2025)
by: Chen, Jinming, et al.
Published: (2025)
Deep Dive Into Music Videos: Hierarchical Emotion Recognition With Rich Audio and Visual Features
by: Yagya Raj Pandeya, et al.
Published: (2025)
by: Yagya Raj Pandeya, et al.
Published: (2025)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
by: Patel, Shubham, et al.
Published: (2024)
by: Patel, Shubham, et al.
Published: (2024)
Exploring Multilingual Unseen Speaker Emotion Recognition: Leveraging Co-Attention Cues in Multitask Learning
by: Goel, Arnav, et al.
Published: (2024)
by: Goel, Arnav, et al.
Published: (2024)
Continuous Adversarial Text Representation Learning for Affective Recognition
by: Son, Seungah, et al.
Published: (2025)
by: Son, Seungah, et al.
Published: (2025)
Deep Emotion Recognition in Textual Conversations: A Survey
by: Pereira, Patrícia, et al.
Published: (2022)
by: Pereira, Patrícia, et al.
Published: (2022)
ECRC: Emotion-Causality Recognition in Korean Conversation for GCN
by: Lee, J. K., et al.
Published: (2024)
by: Lee, J. K., et al.
Published: (2024)
LaERC-S: Improving LLM-based Emotion Recognition in Conversation with Speaker Characteristics
by: Fu, Yumeng, et al.
Published: (2024)
by: Fu, Yumeng, et al.
Published: (2024)
Speaker Emotion Recognition: Leveraging Self-Supervised Models for Feature Extraction Using Wav2Vec2 and HuBERT
by: Jafarzadeh, Pourya, et al.
Published: (2024)
by: Jafarzadeh, Pourya, et al.
Published: (2024)
Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
by: Yang, Chao-Han Huck, et al.
Published: (2024)
by: Yang, Chao-Han Huck, et al.
Published: (2024)
Layer-aware TDNN: Speaker Recognition Using Multi-Layer Features from Pre-Trained Models
by: Kim, Jin Sob, et al.
Published: (2024)
by: Kim, Jin Sob, et al.
Published: (2024)
Triple Disentangled Representation Learning for Multimodal Affective Analysis
by: Zhou, Ying, et al.
Published: (2024)
by: Zhou, Ying, et al.
Published: (2024)
Affective Multimodal Agents with Proactive Knowledge Grounding for Emotionally Aligned Marketing Dialogue
by: Yu, Lin, et al.
Published: (2025)
by: Yu, Lin, et al.
Published: (2025)
TrackAny3D: Transferring Pretrained 3D Models for Category-unified 3D Point Cloud Tracking
by: Wang, Mengmeng, et al.
Published: (2025)
by: Wang, Mengmeng, et al.
Published: (2025)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
by: Peng, Cheng, et al.
Published: (2023)
by: Peng, Cheng, et al.
Published: (2023)
WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing
by: Sun, Xiao
Published: (2025)
by: Sun, Xiao
Published: (2025)
Investigating Safety Vulnerabilities of Large Audio-Language Models Under Speaker Emotional Variations
by: Feng, Bo-Han, et al.
Published: (2025)
by: Feng, Bo-Han, et al.
Published: (2025)
Divide and Refine: Enhancing Multimodal Representation and Explainability for Emotion Recognition in Conversation
by: Mai, Anh-Tuan, et al.
Published: (2026)
by: Mai, Anh-Tuan, et al.
Published: (2026)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
by: Deng, Zhaoyuan, et al.
Published: (2024)
by: Deng, Zhaoyuan, et al.
Published: (2024)
Multi-dataset Joint Pre-training of Emotional EEG Enables Generalizable Affective Computing
by: Zhang, Qingzhu, et al.
Published: (2025)
by: Zhang, Qingzhu, et al.
Published: (2025)
Affective Visual Dialog: A Large-Scale Benchmark for Emotional Reasoning Based on Visually Grounded Conversations
by: Haydarov, Kilichbek, et al.
Published: (2023)
by: Haydarov, Kilichbek, et al.
Published: (2023)
CARE: Cognitive-reasoning Augmented Reinforcement for Emotional Support Conversation
by: Zhu, Jie, et al.
Published: (2025)
by: Zhu, Jie, et al.
Published: (2025)
Similar Items
-
Improving Region Representation Learning from Urban Imagery with Noisy Long-Caption Supervision
by: Zhang, Yimei, et al.
Published: (2025) -
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching
by: Han, Xiao, et al.
Published: (2024) -
Affective Flow Language Model for Emotional Support Conversation
by: Zou, Chenghui, et al.
Published: (2026) -
Socratic RL: A Novel Framework for Efficient Knowledge Acquisition through Iterative Reflection and Viewpoint Distillation
by: Wu, Xiangfan
Published: (2025) -
Multimodal Trajectory Representation Learning for Travel Time Estimation
by: Liu, Zhi, et al.
Published: (2025)