Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions
Fuente:
arXiv
Guardado en:
| Autores principales: | Sun, Licai, Jiang, Xingxun, Chen, Haoyu, Li, Yante, Lian, Zheng, Liu, Biu, Zong, Yuan, Zheng, Wenming, Leppänen, Jukka M., Zhao, Guoying |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Is Micro-expression Ethnic Leaning?
por: Khor, Huai-Qian, et al.
Publicado: (2025)
por: Khor, Huai-Qian, et al.
Publicado: (2025)
Infused Suppression Of Magnification Artefacts For Micro-AU Detection
por: Khor, Huai-Qian, et al.
Publicado: (2025)
por: Khor, Huai-Qian, et al.
Publicado: (2025)
Towards Consistent and Controllable Image Synthesis for Face Editing
por: Wei, Mengting, et al.
Publicado: (2025)
por: Wei, Mengting, et al.
Publicado: (2025)
MagicFace: High-Fidelity Facial Expression Editing with Action-Unit Control
por: Wei, Mengting, et al.
Publicado: (2025)
por: Wei, Mengting, et al.
Publicado: (2025)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
por: Qi, Tianhua, et al.
Publicado: (2024)
por: Qi, Tianhua, et al.
Publicado: (2024)
Deep Change Monitoring: A Hyperbolic Representative Learning Framework and a Dataset for Long-term Fine-grained Tree Change Detection
por: Li, Yante, et al.
Publicado: (2025)
por: Li, Yante, et al.
Publicado: (2025)
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
por: Sun, Licai, et al.
Publicado: (2024)
por: Sun, Licai, et al.
Publicado: (2024)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
por: Lu, Cheng, et al.
Publicado: (2024)
por: Lu, Cheng, et al.
Publicado: (2024)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
por: Qi, Tianhua, et al.
Publicado: (2026)
por: Qi, Tianhua, et al.
Publicado: (2026)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
por: Qi, Tianhua, et al.
Publicado: (2024)
por: Qi, Tianhua, et al.
Publicado: (2024)
Towards Localized Fine-Grained Control for Facial Expression Generation
por: Varanka, Tuomas, et al.
Publicado: (2024)
por: Varanka, Tuomas, et al.
Publicado: (2024)
Micro-AU CLIP: Fine-Grained Contrastive Learning from Local Independence to Global Dependency for Micro-Expression Action Unit Detection
por: Wei, Jinsheng, et al.
Publicado: (2026)
por: Wei, Jinsheng, et al.
Publicado: (2026)
Speech Swin-Transformer: Exploring a Hierarchical Transformer with Shifted Windows for Speech Emotion Recognition
por: Wang, Yong, et al.
Publicado: (2024)
por: Wang, Yong, et al.
Publicado: (2024)
AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition
por: Lian, Zheng, et al.
Publicado: (2024)
por: Lian, Zheng, et al.
Publicado: (2024)
MagicPortrait: Temporally Consistent Face Reenactment with 3D Geometric Guidance
por: Wei, Mengting, et al.
Publicado: (2025)
por: Wei, Mengting, et al.
Publicado: (2025)
Temporal Label Hierachical Network for Compound Emotion Recognition
por: Li, Sunan, et al.
Publicado: (2024)
por: Li, Sunan, et al.
Publicado: (2024)
SVFAP: Self-supervised Video Facial Affect Perceiver
por: Sun, Licai, et al.
Publicado: (2023)
por: Sun, Licai, et al.
Publicado: (2023)
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
por: Lian, Zheng, et al.
Publicado: (2025)
por: Lian, Zheng, et al.
Publicado: (2025)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
por: Lian, Zheng, et al.
Publicado: (2023)
por: Lian, Zheng, et al.
Publicado: (2023)
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
por: Zhao, Yan, et al.
Publicado: (2024)
por: Zhao, Yan, et al.
Publicado: (2024)
AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models
por: Lian, Zheng, et al.
Publicado: (2025)
por: Lian, Zheng, et al.
Publicado: (2025)
ITEACH-Net: Inverted Teacher-studEnt seArCH Network for Emotion Recognition in Conversation
por: Sun, Haiyang, et al.
Publicado: (2023)
por: Sun, Haiyang, et al.
Publicado: (2023)
MERBench: A Unified Evaluation Benchmark for Multimodal Emotion Recognition
por: Lian, Zheng, et al.
Publicado: (2024)
por: Lian, Zheng, et al.
Publicado: (2024)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
por: Hu, Zhuozhao, et al.
Publicado: (2025)
por: Hu, Zhuozhao, et al.
Publicado: (2025)
Decoupled Doubly Contrastive Learning for Cross Domain Facial Action Unit Detection
por: Li, Yong, et al.
Publicado: (2025)
por: Li, Yong, et al.
Publicado: (2025)
Axisymmetric Coil Winding Surfaces for Non-Axisymmetric Fusion Devices
por: Biu, J., et al.
Publicado: (2025)
por: Biu, J., et al.
Publicado: (2025)
Biometric Authentication Based on Enhanced Remote Photoplethysmography Signal Morphology
por: Sun, Zhaodong, et al.
Publicado: (2024)
por: Sun, Zhaodong, et al.
Publicado: (2024)
Towards Robust 3D Pose Transfer with Adversarial Learning
por: Chen, Haoyu, et al.
Publicado: (2024)
por: Chen, Haoyu, et al.
Publicado: (2024)
Multimodal Functional Maximum Correlation for Emotion Recognition
por: Zheng, Deyang, et al.
Publicado: (2025)
por: Zheng, Deyang, et al.
Publicado: (2025)
Hybrid-supervised Hypergraph-enhanced Transformer for Micro-gesture Based Emotion Recognition
por: Xia, Zhaoqiang, et al.
Publicado: (2025)
por: Xia, Zhaoqiang, et al.
Publicado: (2025)
EmoTransCap: Dataset and Pipeline for Emotion Transition-Aware Speech Captioning in Discourses
por: Xu, Shuhao, et al.
Publicado: (2026)
por: Xu, Shuhao, et al.
Publicado: (2026)
PC-MNet: Dual-Level Congruity Modeling for Multimodal Sarcasm Detection via Polarity-Modulated Attention
por: Li, Maoheng, et al.
Publicado: (2026)
por: Li, Maoheng, et al.
Publicado: (2026)
Incorporating Scene Context and Semantic Labels for Enhanced Group-level Emotion Recognition
por: Zhu, Qing, et al.
Publicado: (2025)
por: Zhu, Qing, et al.
Publicado: (2025)
Self-Supervised Facial Representation Learning with Facial Region Awareness
por: Gao, Zheng, et al.
Publicado: (2024)
por: Gao, Zheng, et al.
Publicado: (2024)
MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition
por: Lian, Zheng, et al.
Publicado: (2024)
por: Lian, Zheng, et al.
Publicado: (2024)
SPOLRE: Semantic Preserving Object Layout Reconstruction for Image Captioning System Testing
por: Liu, Yi, et al.
Publicado: (2024)
por: Liu, Yi, et al.
Publicado: (2024)
How Large Language Models Are Changing MOOC Essay Answers: A Comparison of Pre- and Post-LLM Responses
por: Leppänen, Leo, et al.
Publicado: (2025)
por: Leppänen, Leo, et al.
Publicado: (2025)
Capturing Rich Behavior Representations: A Dynamic Action Semantic-Aware Graph Transformer for Video Captioning
por: Liu, Caihua, et al.
Publicado: (2025)
por: Liu, Caihua, et al.
Publicado: (2025)
OV-MER: Towards Open-Vocabulary Multimodal Emotion Recognition
por: Lian, Zheng, et al.
Publicado: (2024)
por: Lian, Zheng, et al.
Publicado: (2024)
Enhanced Algorithmic Perfect State Transfer on IBM Quantum Computers
por: Ge, Zong-Yuan, et al.
Publicado: (2025)
por: Ge, Zong-Yuan, et al.
Publicado: (2025)
Ejemplares similares
-
Is Micro-expression Ethnic Leaning?
por: Khor, Huai-Qian, et al.
Publicado: (2025) -
Infused Suppression Of Magnification Artefacts For Micro-AU Detection
por: Khor, Huai-Qian, et al.
Publicado: (2025) -
Towards Consistent and Controllable Image Synthesis for Face Editing
por: Wei, Mengting, et al.
Publicado: (2025) -
MagicFace: High-Fidelity Facial Expression Editing with Action-Unit Control
por: Wei, Mengting, et al.
Publicado: (2025) -
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
por: Qi, Tianhua, et al.
Publicado: (2024)