Spatiotemporal Emotional Synchrony in Dyadic Interactions: The Role of Speech Conditions in Facial and Vocal Affective Alignment
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Herbuela, Von Ralph Dane Marquez, Nagai, Yukie |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Multimodal Emotion Coupling via Speech-to-Facial and Bodily Gestures in Dyadic Interaction
par: Herbuela, Von Ralph Dane Marquez, et autres
Publié: (2025)
par: Herbuela, Von Ralph Dane Marquez, et autres
Publié: (2025)
Learning Physiology-Informed Vocal Spectrotemporal Representations for Speech Emotion Recognition
par: Zhang, Xu, et autres
Publié: (2026)
par: Zhang, Xu, et autres
Publié: (2026)
Facial Expression-Enhanced TTS: Combining Face Representation and Emotion Intensity for Adaptive Speech
par: Chu, Yunji, et autres
Publié: (2024)
par: Chu, Yunji, et autres
Publié: (2024)
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
par: Dutta, Soumya, et autres
Publié: (2025)
par: Dutta, Soumya, et autres
Publié: (2025)
NV-Bench: Benchmark of Nonverbal Vocalization Synthesis for Expressive Text-to-Speech Generation
par: Ni, Qinke, et autres
Publié: (2026)
par: Ni, Qinke, et autres
Publié: (2026)
Egocentric Speaker Classification in Child-Adult Dyadic Interactions: From Sensing to Computational Modeling
par: Feng, Tiantian, et autres
Publié: (2024)
par: Feng, Tiantian, et autres
Publié: (2024)
Qieemo: Speech Is All You Need in the Emotion Recognition in Conversations
par: Chen, Jinming, et autres
Publié: (2025)
par: Chen, Jinming, et autres
Publié: (2025)
Color-based Emotion Representation for Speech Emotion Recognition
par: Nagase, Ryotaro, et autres
Publié: (2026)
par: Nagase, Ryotaro, et autres
Publié: (2026)
MNV-17: A High-Quality Performative Mandarin Dataset for Nonverbal Vocalization Recognition in Speech
par: Mai, Jialong, et autres
Publié: (2025)
par: Mai, Jialong, et autres
Publié: (2025)
ES4R: Speech Encoding Based on Prepositive Affective Modeling for Empathetic Response Generation
par: Gao, Zhuoyue, et autres
Publié: (2026)
par: Gao, Zhuoyue, et autres
Publié: (2026)
Coding Speech through Vocal Tract Kinematics
par: Cho, Cheol Jun, et autres
Publié: (2024)
par: Cho, Cheol Jun, et autres
Publié: (2024)
Emotional Text-To-Speech Based on Mutual-Information-Guided Emotion-Timbre Disentanglement
par: Yang, Jianing, et autres
Publié: (2025)
par: Yang, Jianing, et autres
Publié: (2025)
VocalAgent: Large Language Models for Vocal Health Diagnostics with Safety-Aware Evaluation
par: Kim, Yubin, et autres
Publié: (2025)
par: Kim, Yubin, et autres
Publié: (2025)
EmoSpeech: A Corpus of Emotionally Rich and Contextually Detailed Speech Annotations
par: Bian, Weizhen, et autres
Publié: (2024)
par: Bian, Weizhen, et autres
Publié: (2024)
Joint Learning using Mixture-of-Expert-Based Representation for Speech Enhancement and Robust Emotion Recognition
par: Tzeng, Jing-Tong, et autres
Publié: (2025)
par: Tzeng, Jing-Tong, et autres
Publié: (2025)
Text-to-Song: Towards Controllable Music Generation Incorporating Vocals and Accompaniment
par: Hong, Zhiqing, et autres
Publié: (2024)
par: Hong, Zhiqing, et autres
Publié: (2024)
TASU2: Controllable CTC Simulation for Alignment and Low-Resource Adaptation of Speech LLMs
par: Peng, Jing, et autres
Publié: (2026)
par: Peng, Jing, et autres
Publié: (2026)
EmoSphere-TTS: Emotional Style and Intensity Modeling via Spherical Emotion Vector for Controllable Emotional Text-to-Speech
par: Cho, Deok-Hyeon, et autres
Publié: (2024)
par: Cho, Deok-Hyeon, et autres
Publié: (2024)
Persian Speech Emotion Recognition by Fine-Tuning Transformers
par: Shayaninasab, Minoo, et autres
Publié: (2024)
par: Shayaninasab, Minoo, et autres
Publié: (2024)
Adaptive Duration Model for Text Speech Alignment
par: Cao, Junjie
Publié: (2025)
par: Cao, Junjie
Publié: (2025)
ASR for Affective Speech: Investigating Impact of Emotion and Speech Generative Strategy
par: Wu, Ya-Tse, et autres
Publié: (2026)
par: Wu, Ya-Tse, et autres
Publié: (2026)
EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vector
par: Cho, Deok-Hyeon, et autres
Publié: (2024)
par: Cho, Deok-Hyeon, et autres
Publié: (2024)
Efficient Finetuning for Dimensional Speech Emotion Recognition in the Age of Transformers
par: Sampath, Aneesha, et autres
Publié: (2025)
par: Sampath, Aneesha, et autres
Publié: (2025)
EmoAttack: Utilizing Emotional Voice Conversion for Speech Backdoor Attacks on Deep Speech Classification Models
par: Yao, Wenhan, et autres
Publié: (2024)
par: Yao, Wenhan, et autres
Publié: (2024)
Toward Efficient Speech Emotion Recognition via Spectral Learning and Attention
par: Lee, HyeYoung, et autres
Publié: (2025)
par: Lee, HyeYoung, et autres
Publié: (2025)
Breaking Resource Barriers in Speech Emotion Recognition via Data Distillation
par: Chang, Yi, et autres
Publié: (2024)
par: Chang, Yi, et autres
Publié: (2024)
Active Learning with Task Adaptation Pre-training for Speech Emotion Recognition
par: Li, Dongyuan, et autres
Publié: (2024)
par: Li, Dongyuan, et autres
Publié: (2024)
MoVE: Translating Laughter and Tears via Mixture of Vocalization Experts in Speech-to-Speech Translation
par: Chen, Szu-Chi, et autres
Publié: (2026)
par: Chen, Szu-Chi, et autres
Publié: (2026)
Speech-DRAME: A Framework for Human-Aligned Benchmarks in Speech Role-Play
par: Shi, Jiatong, et autres
Publié: (2025)
par: Shi, Jiatong, et autres
Publié: (2025)
DiEmo-TTS: Disentangled Emotion Representations via Self-Supervised Distillation for Cross-Speaker Emotion Transfer in Text-to-Speech
par: Cho, Deok-Hyeon, et autres
Publié: (2025)
par: Cho, Deok-Hyeon, et autres
Publié: (2025)
Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
par: Neekhara, Paarth, et autres
Publié: (2024)
par: Neekhara, Paarth, et autres
Publié: (2024)
Improved Intelligibility of Dysarthric Speech using Conditional Flow Matching
par: Das, Shoutrik, et autres
Publié: (2025)
par: Das, Shoutrik, et autres
Publié: (2025)
Construction and Evaluation of Mandarin Multimodal Emotional Speech Database
par: Ting, Zhu, et autres
Publié: (2024)
par: Ting, Zhu, et autres
Publié: (2024)
Improvement and Implementation of a Speech Emotion Recognition Model Based on Dual-Layer LSTM
par: Yang, Xiaoran, et autres
Publié: (2024)
par: Yang, Xiaoran, et autres
Publié: (2024)
Are you sure? Analysing Uncertainty Quantification Approaches for Real-world Speech Emotion Recognition
par: Schrüfer, Oliver, et autres
Publié: (2024)
par: Schrüfer, Oliver, et autres
Publié: (2024)
Speech Emotion Recognition Using MFCC Features and LSTM-Based Deep Learning Model
par: Oluwademilade, Adelekun, et autres
Publié: (2026)
par: Oluwademilade, Adelekun, et autres
Publié: (2026)
Explaining Deep Learning Embeddings for Speech Emotion Recognition by Predicting Interpretable Acoustic Features
par: Dixit, Satvik, et autres
Publié: (2024)
par: Dixit, Satvik, et autres
Publié: (2024)
MSAC: Multiple Speech Attribute Control Method for Reliable Speech Emotion Recognition
par: Pan, Yu, et autres
Publié: (2023)
par: Pan, Yu, et autres
Publié: (2023)
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
par: Ma, Ziyang, et autres
Publié: (2023)
par: Ma, Ziyang, et autres
Publié: (2023)
MPE-TTS: Customized Emotion Zero-Shot Text-To-Speech Using Multi-Modal Prompt
par: Wu, Zhichao, et autres
Publié: (2025)
par: Wu, Zhichao, et autres
Publié: (2025)
Documents similaires
-
Multimodal Emotion Coupling via Speech-to-Facial and Bodily Gestures in Dyadic Interaction
par: Herbuela, Von Ralph Dane Marquez, et autres
Publié: (2025) -
Learning Physiology-Informed Vocal Spectrotemporal Representations for Speech Emotion Recognition
par: Zhang, Xu, et autres
Publié: (2026) -
Facial Expression-Enhanced TTS: Combining Face Representation and Emotion Intensity for Adaptive Speech
par: Chu, Yunji, et autres
Publié: (2024) -
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
par: Dutta, Soumya, et autres
Publié: (2025) -
NV-Bench: Benchmark of Nonverbal Vocalization Synthesis for Expressive Text-to-Speech Generation
par: Ni, Qinke, et autres
Publié: (2026)