Expressivity and Speech Synthesis
Fuente:
arXiv
Guardado en:
| Autores principales: | Triantafyllopoulos, Andreas, Schuller, Björn W. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
INTERSPEECH 2009 Emotion Challenge Revisited: Benchmarking 15 Years of Progress in Speech Emotion Recognition
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models
por: Tsangko, Iosif, et al.
Publicado: (2026)
por: Tsangko, Iosif, et al.
Publicado: (2026)
ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis
por: He, Xiangheng, et al.
Publicado: (2024)
por: He, Xiangheng, et al.
Publicado: (2024)
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
Sustained Vowels for Pre- vs Post-Treatment COPD Classification
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
Explainable Speech Emotion Recognition: Weighted Attribute Fairness to Model Demographic Contributions to Social Bias
por: Ogunnubi, Tomisin, et al.
Publicado: (2026)
por: Ogunnubi, Tomisin, et al.
Publicado: (2026)
Generative Expressive Conversational Speech Synthesis
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
On Prompt Sensitivity of ChatGPT in Affective Computing
por: Amin, Mostafa M., et al.
Publicado: (2024)
por: Amin, Mostafa M., et al.
Publicado: (2024)
Speech Swin-Transformer: Exploring a Hierarchical Transformer with Shifted Windows for Speech Emotion Recognition
por: Wang, Yong, et al.
Publicado: (2024)
por: Wang, Yong, et al.
Publicado: (2024)
Abusive Speech Detection in Indic Languages Using Acoustic Features
por: Spiesberger, Anika A., et al.
Publicado: (2024)
por: Spiesberger, Anika A., et al.
Publicado: (2024)
Retrieval-Augmented Dialogue Knowledge Aggregation for Expressive Conversational Speech Synthesis
por: Liu, Rui, et al.
Publicado: (2025)
por: Liu, Rui, et al.
Publicado: (2025)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Task Selection and Assignment for Multi-modal Multi-task Dialogue Act Classification with Non-stationary Multi-armed Bandits
por: He, Xiangheng, et al.
Publicado: (2023)
por: He, Xiangheng, et al.
Publicado: (2023)
DFingerNet: Noise-Adaptive Speech Enhancement for Hearing Aids
por: Tsangko, Iosif, et al.
Publicado: (2025)
por: Tsangko, Iosif, et al.
Publicado: (2025)
Detecting COPD Through Speech Analysis: A Dataset of Danish Speech and Machine Learning Approach
por: Sankey-Olsen, Cuno, et al.
Publicado: (2025)
por: Sankey-Olsen, Cuno, et al.
Publicado: (2025)
Discourse Features Enhance Detection of Document-Level Machine-Generated Content
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
Neuron-Level Emotion Control in Speech-Generative Large Audio-Language Models
por: Zhao, Xiutian, et al.
Publicado: (2026)
por: Zhao, Xiutian, et al.
Publicado: (2026)
Style Mixture of Experts for Expressive Text-To-Speech Synthesis
por: Jawaid, Ahad, et al.
Publicado: (2024)
por: Jawaid, Ahad, et al.
Publicado: (2024)
HAFFormer: A Hierarchical Attention-Free Framework for Alzheimer's Disease Detection From Spontaneous Speech
por: Dong, Zhongren, et al.
Publicado: (2024)
por: Dong, Zhongren, et al.
Publicado: (2024)
Large Language Models for Depression Recognition in Spoken Language Integrating Psychological Knowledge
por: Li, Yupei, et al.
Publicado: (2025)
por: Li, Yupei, et al.
Publicado: (2025)
Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generators
por: Mady, Mohamed, et al.
Publicado: (2026)
por: Mady, Mohamed, et al.
Publicado: (2026)
Speech-Based Depressive Mood Detection in the Presence of Multiple Sclerosis: A Cross-Corpus and Cross-Lingual Study
por: Gonzalez-Machorro, Monica, et al.
Publicado: (2025)
por: Gonzalez-Machorro, Monica, et al.
Publicado: (2025)
Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy
por: Sheikh, Shakeel, et al.
Publicado: (2026)
por: Sheikh, Shakeel, et al.
Publicado: (2026)
Discovering and Causally Validating Emotion-Sensitive Neurons in Large Audio-Language Models
por: Zhao, Xiutian, et al.
Publicado: (2026)
por: Zhao, Xiutian, et al.
Publicado: (2026)
Automatic Emotion Modelling in Written Stories
por: Christ, Lukas, et al.
Publicado: (2022)
por: Christ, Lukas, et al.
Publicado: (2022)
SeamlessExpressiveLM: Speech Language Model for Expressive Speech-to-Speech Translation with Chain-of-Thought
por: Gong, Hongyu, et al.
Publicado: (2024)
por: Gong, Hongyu, et al.
Publicado: (2024)
From Divergence to Consensus: Evaluating the Role of Large Language Models in Facilitating Agreement through Adaptive Strategies
por: Triantafyllopoulos, Loukas, et al.
Publicado: (2025)
por: Triantafyllopoulos, Loukas, et al.
Publicado: (2025)
SpeechCraft: A Fine-grained Expressive Speech Dataset with Natural Language Description
por: Jin, Zeyu, et al.
Publicado: (2024)
por: Jin, Zeyu, et al.
Publicado: (2024)
ExHuBERT: Enhancing HuBERT Through Block Extension and Fine-Tuning on 37 Emotion Datasets
por: Amiriparian, Shahin, et al.
Publicado: (2024)
por: Amiriparian, Shahin, et al.
Publicado: (2024)
Does the Definition of Difficulty Matter? Scoring Functions and their Role for Curriculum Learning
por: Rampp, Simon, et al.
Publicado: (2024)
por: Rampp, Simon, et al.
Publicado: (2024)
autrainer: A Modular and Extensible Deep Learning Toolkit for Computer Audition Tasks
por: Rampp, Simon, et al.
Publicado: (2024)
por: Rampp, Simon, et al.
Publicado: (2024)
Non-Invasive Suicide Risk Prediction Through Speech Analysis
por: Amiriparian, Shahin, et al.
Publicado: (2024)
por: Amiriparian, Shahin, et al.
Publicado: (2024)
Audio-based Step-count Estimation for Running -- Windowing and Neural Network Baselines
por: Wagner, Philipp, et al.
Publicado: (2024)
por: Wagner, Philipp, et al.
Publicado: (2024)
Modeling Emotional Trajectories in Written Stories Utilizing Transformers and Weakly-Supervised Learning
por: Christ, Lukas, et al.
Publicado: (2024)
por: Christ, Lukas, et al.
Publicado: (2024)
GTR-Voice: Articulatory Phonetics Informed Controllable Expressive Speech Synthesis
por: Li, Zehua Kcriss, et al.
Publicado: (2024)
por: Li, Zehua Kcriss, et al.
Publicado: (2024)
Semantic Differentiation in Speech Emotion Recognition: Insights from Descriptive and Expressive Speech Roles
por: Guo, Rongchen, et al.
Publicado: (2025)
por: Guo, Rongchen, et al.
Publicado: (2025)
Computational Narrative Understanding for Expressive Text-to-Speech
por: Michel, Gaspard, et al.
Publicado: (2025)
por: Michel, Gaspard, et al.
Publicado: (2025)
Ejemplares similares
-
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024) -
INTERSPEECH 2009 Emotion Challenge Revisited: Benchmarking 15 Years of Progress in Speech Emotion Recognition
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024) -
A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models
por: Tsangko, Iosif, et al.
Publicado: (2026) -
ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis
por: He, Xiangheng, et al.
Publicado: (2024) -
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)