Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
Fuente:
arXiv
Guardado en:
| Autores principales: | Narain, Jaya, Kowtha, Vasudha, Lea, Colin, Tooley, Lauren, Yee, Dianna, Mitra, Vikramjit, Huang, Zifang, Marques, Miquel Espi, Huang, Jon, Avendano, Carlos, Ren, Shirley |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Affect Models Have Weak Generalizability to Atypical Speech
por: Narain, Jaya, et al.
Publicado: (2025)
por: Narain, Jaya, et al.
Publicado: (2025)
Foundation Model Hidden Representations for Heart Rate Estimation from Auscultation
por: Nie, Jingping, et al.
Publicado: (2025)
por: Nie, Jingping, et al.
Publicado: (2025)
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
por: Romana, Amrit, et al.
Publicado: (2025)
por: Romana, Amrit, et al.
Publicado: (2025)
Hypernetworks for Personalizing ASR to Atypical Speech
por: Müller-Eberstein, Max, et al.
Publicado: (2024)
por: Müller-Eberstein, Max, et al.
Publicado: (2024)
Speech Foundation Models Generalize to Time Series Tasks from Wearable Sensor Data
por: Narain, Jaya, et al.
Publicado: (2025)
por: Narain, Jaya, et al.
Publicado: (2025)
Using LLMs for Late Multimodal Sensor Fusion for Activity Recognition
por: Demirel, Ilker, et al.
Publicado: (2025)
por: Demirel, Ilker, et al.
Publicado: (2025)
The global education industry : lessons from private education in developing countries / James Tooley
por: Tooley, James
Publicado: (1999)
por: Tooley, James
Publicado: (1999)
Big6 Turbo Tools and Location & Access
por: Tooley, Melinda
Publicado: (2005)
por: Tooley, Melinda
Publicado: (2005)
Big6 Turbotools and Synthesis
por: Tooley, Melinda
Publicado: (2005)
por: Tooley, Melinda
Publicado: (2005)
An experiment in market‐led higher education: The case of the Buckingham ‘licence’
por: James Tooley
Publicado: (2024)
por: James Tooley
Publicado: (2024)
Modeling speech emotion with label variance and analyzing performance across speakers and unseen acoustic conditions
por: Mitra, Vikramjit, et al.
Publicado: (2025)
por: Mitra, Vikramjit, et al.
Publicado: (2025)
Voice Conversion for Lombard Speaking Style with Implicit and Explicit Acoustic Feature Conditioning
por: Woszczyk, Dominika, et al.
Publicado: (2025)
por: Woszczyk, Dominika, et al.
Publicado: (2025)
Progress in Finnish Librarianship
por: Tooley, J. B.
Publicado: (1971)
por: Tooley, J. B.
Publicado: (1971)
Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect
por: Presa, Joao Paulo Cavalcante, et al.
Publicado: (2026)
por: Presa, Joao Paulo Cavalcante, et al.
Publicado: (2026)
¿Tiene solución la ciudad?
por: Mariano Vázquez Espí
Publicado: (2008)
por: Mariano Vázquez Espí
Publicado: (2008)
PERCEPCIÓN AFECTIVA DEL ALUMNADO EN EDUCACIÓN FÍSICA PRIMARIA EN LA PROVINCIA DE ALICANTE
por: Andrea Jordá-Espi
Publicado: (2019)
por: Andrea Jordá-Espi
Publicado: (2019)
Construcciones utópicas: tres tesis y una regla práctica
por: Mariano Vázquez Espí
Publicado: (2003)
por: Mariano Vázquez Espí
Publicado: (2003)
Do LLMs estimate uncertainty well in instruction-following?
por: Heo, Juyeon, et al.
Publicado: (2024)
por: Heo, Juyeon, et al.
Publicado: (2024)
Model-driven Heart Rate Estimation and Heart Murmur Detection based on Phonocardiogram
por: Nie, Jingping, et al.
Publicado: (2024)
por: Nie, Jingping, et al.
Publicado: (2024)
Sustainable processing and nutritional enhancement of dried figs ( Ficus Carica L. cv. Brown Turkey) by different drying technologies
por: Vikramjit Singh, et al.
Publicado: (2025)
por: Vikramjit Singh, et al.
Publicado: (2025)
S$^2$Voice: Style-Aware Autoregressive Modeling with Enhanced Conditioning for Singing Style Conversion
por: Wang, Ziqian, et al.
Publicado: (2026)
por: Wang, Ziqian, et al.
Publicado: (2026)
Voice and Speech in Atypical Parkinsonian Disorders
por: Federico Rodriguez‐Porcel, et al.
Publicado: (2026)
por: Federico Rodriguez‐Porcel, et al.
Publicado: (2026)
StyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis
por: Zhang, Yu, et al.
Publicado: (2023)
por: Zhang, Yu, et al.
Publicado: (2023)
ANÁLISIS DE LAS FALACIAS EN TORNO A LA TEORÍA DE LA SOBERANÍA NACIONAL (O POPULAR)
por: Luis-Tomás Zapater Espí
Publicado: (2017)
por: Luis-Tomás Zapater Espí
Publicado: (2017)
La visión del estudiante respecto a la Diplomatura de Fisioterapia en la Universitat de València: un estudio descriptivo
por: Gemma Victoria Espí López
Publicado: (2012)
por: Gemma Victoria Espí López
Publicado: (2012)
Factor-Conditioned Speaking-Style Captioning
por: Ando, Atsushi, et al.
Publicado: (2024)
por: Ando, Atsushi, et al.
Publicado: (2024)
A Statewide Instructional Television Program via Satellite for RN-to-BSN Students.
por: Shomaker, Dianna
Publicado: (1993)
por: Shomaker, Dianna
Publicado: (1993)
Looking back: Selected contributions by C. R. Rao to multivariate analysis
por: Dianna Smith
Publicado: (2024)
por: Dianna Smith
Publicado: (2024)
Gassner and Burau representations over $\mathbb{Z}_p$-modules
por: Bharathram, Vasudha
Publicado: (2022)
por: Bharathram, Vasudha
Publicado: (2022)
TCSinger: Zero-Shot Singing Voice Synthesis with Style Transfer and Multi-Level Style Control
por: Zhang, Yu, et al.
Publicado: (2024)
por: Zhang, Yu, et al.
Publicado: (2024)
Distribution of Primitive Lattice Points in Large Dimensions
por: Han, Jiyoung
Publicado: (2024)
por: Han, Jiyoung
Publicado: (2024)
Do LLMs "know" internally when they follow instructions?
por: Heo, Juyeon, et al.
Publicado: (2024)
por: Heo, Juyeon, et al.
Publicado: (2024)
Voice "Cloning" is Style Transfer
por: Zhou, Kaitlyn, et al.
Publicado: (2026)
por: Zhou, Kaitlyn, et al.
Publicado: (2026)
August Wilson: The Unrestrained Voice of Black America
por: Arpita Mitra
Publicado: (2021)
por: Arpita Mitra
Publicado: (2021)
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
por: Wang, Suzhen, et al.
Publicado: (2024)
por: Wang, Suzhen, et al.
Publicado: (2024)
StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control
por: Zhao, Haishu, et al.
Publicado: (2026)
por: Zhao, Haishu, et al.
Publicado: (2026)
XBRL Filing in the Indonesian Capital Market: An Exploration of Firms' Responses
por: Fitri Amalia, et al.
Publicado: (2026)
por: Fitri Amalia, et al.
Publicado: (2026)
Prompting Whisper for Improved Verbatim Transcription and End-to-end Miscue Detection
por: Smith, Griffin Dietz, et al.
Publicado: (2025)
por: Smith, Griffin Dietz, et al.
Publicado: (2025)
El federalismo de la India está plagado de altercados por las vías fluviales
por: Narain, A
Publicado: (2009)
por: Narain, A
Publicado: (2009)
Water Security, Conflict and Cooperation in Peri-Urban South Asia Flows across Boundaries
por: Vishal Narain
por: Vishal Narain
Ejemplares similares
-
Affect Models Have Weak Generalizability to Atypical Speech
por: Narain, Jaya, et al.
Publicado: (2025) -
Foundation Model Hidden Representations for Heart Rate Estimation from Auscultation
por: Nie, Jingping, et al.
Publicado: (2025) -
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
por: Romana, Amrit, et al.
Publicado: (2025) -
Hypernetworks for Personalizing ASR to Atypical Speech
por: Müller-Eberstein, Max, et al.
Publicado: (2024) -
Speech Foundation Models Generalize to Time Series Tasks from Wearable Sensor Data
por: Narain, Jaya, et al.
Publicado: (2025)