Salvato in:
| Autori principali: | Siriwardena, Yashish M., Swedlow, Nathan, Howard, Audrey, Gitterman, Evan, Darcy, Dan, Espy-Wilson, Carol, Fanelli, Andrea |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.05947 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Speech Inversion Through Self-Supervised Embeddings and Enhanced Tract Variables
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2023)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2023)
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
A multi-modal approach for identifying schizophrenia using cross-modal attention
di: Premananth, Gowtham, et al.
Pubblicazione: (2023)
di: Premananth, Gowtham, et al.
Pubblicazione: (2023)
A Multimodal Framework for the Assessment of the Schizophrenia Spectrum
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
Quantifying Articulatory Coordination as a Biomarker for Schizophrenia
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
Acoustic to Articulatory Speech Inversion for Children with Velopharyngeal Insufficiency
di: Tabatabaee, Saba, et al.
Pubblicazione: (2025)
di: Tabatabaee, Saba, et al.
Pubblicazione: (2025)
Enhancing Acoustic-to-Articulatory Speech Inversion by Incorporating Nasality
di: Tabatabaee, Saba, et al.
Pubblicazione: (2025)
di: Tabatabaee, Saba, et al.
Pubblicazione: (2025)
Perceptual Ratings Predict Speech Inversion Articulatory Kinematics in Childhood Speech Sound Disorders
di: Benway, Nina R., et al.
Pubblicazione: (2025)
di: Benway, Nina R., et al.
Pubblicazione: (2025)
Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
Articulation-Informed ASR: Integrating Articulatory Features into ASR via Auxiliary Speech Inversion and Cross-Attention Fusion
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2025)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2025)
Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
di: Tabatabaee, Saba, et al.
Pubblicazione: (2026)
di: Tabatabaee, Saba, et al.
Pubblicazione: (2026)
FT-Boosted SV: Towards Noise Robust Speaker Verification for English Speaking Classroom Environments
di: Tabatabaee, Saba, et al.
Pubblicazione: (2025)
di: Tabatabaee, Saba, et al.
Pubblicazione: (2025)
A Computational Approach to Analyzing Disrupted Language in Schizophrenia: Integrating Surprisal and Coherence Measures
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
Reverse Attention for Lightweight Speech Enhancement on Edge Devices
di: Ojha, Shuubham, et al.
Pubblicazione: (2025)
di: Ojha, Shuubham, et al.
Pubblicazione: (2025)
On the Relationship between Accent Strength and Articulatory Features
di: Huang, Kevin, et al.
Pubblicazione: (2025)
di: Huang, Kevin, et al.
Pubblicazione: (2025)
RealClass: A Framework for Classroom Speech Simulation with Public Datasets and Game Engines
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2025)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2025)
Speech-Based Prioritization for Schizophrenia Intervention
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
MacST: Multi-Accent Speech Synthesis via Text Transliteration for Accent Conversion
di: Inoue, Sho, et al.
Pubblicazione: (2024)
di: Inoue, Sho, et al.
Pubblicazione: (2024)
Deep Speech Synthesis from Multimodal Articulatory Representations
di: Wu, Peter, et al.
Pubblicazione: (2024)
di: Wu, Peter, et al.
Pubblicazione: (2024)
From Weak Labels to Strong Results: Utilizing 5,000 Hours of Noisy Classroom Transcripts with Minimal Accurate Data
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2025)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2025)
Teaching Machines to Speak Using Articulatory Control
di: Anand, Akshay, et al.
Pubblicazione: (2025)
di: Anand, Akshay, et al.
Pubblicazione: (2025)
Kid-Whisper: Towards Bridging the Performance Gap in Automatic Speech Recognition for Children VS. Adults
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2023)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2023)
Scalable Controllable Accented TTS
di: Xinyuan, Henry Li, et al.
Pubblicazione: (2025)
di: Xinyuan, Henry Li, et al.
Pubblicazione: (2025)
Towards a Quantitative Analysis of Coarticulation with a Phoneme-to-Articulatory Model
di: Fan, Chaofei, et al.
Pubblicazione: (2024)
di: Fan, Chaofei, et al.
Pubblicazione: (2024)
RT-VC: Real-Time Zero-Shot Voice Conversion with Speech Articulatory Coding
di: Liu, Yisi, et al.
Pubblicazione: (2025)
di: Liu, Yisi, et al.
Pubblicazione: (2025)
Continued Pretraining for Domain Adaptation of Wav2vec2.0 in Automatic Speech Recognition for Elementary Math Classroom Settings
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2024)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2024)
Multimodal Biomarkers for Schizophrenia: Towards Individual Symptom Severity Estimation
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)
Acoustic-to-Articulatory Inversion of Clean Speech Using an MRI-Trained Model
di: Azzouz, Sofiane, et al.
Pubblicazione: (2026)
di: Azzouz, Sofiane, et al.
Pubblicazione: (2026)
SpeechAccentLLM: A Unified Framework for Foreign Accent Conversion and Text to Speech
di: Cheng, Zhuangfei, et al.
Pubblicazione: (2025)
di: Cheng, Zhuangfei, et al.
Pubblicazione: (2025)
Multi-Scale Accent Modeling and Disentangling for Multi-Speaker Multi-Accent Text-to-Speech Synthesis
di: Zhou, Xuehao, et al.
Pubblicazione: (2024)
di: Zhou, Xuehao, et al.
Pubblicazione: (2024)
Convert and Speak: Zero-shot Accent Conversion with Minimum Supervision
di: Jia, Zhijun, et al.
Pubblicazione: (2024)
di: Jia, Zhijun, et al.
Pubblicazione: (2024)
Streaming Non-Autoregressive Model for Accent Conversion and Pronunciation Improvement
di: Nguyen, Tuan-Nam, et al.
Pubblicazione: (2025)
di: Nguyen, Tuan-Nam, et al.
Pubblicazione: (2025)
Pairwise Evaluation of Accent Similarity in Speech Synthesis
di: Zhong, Jinzuomu, et al.
Pubblicazione: (2025)
di: Zhong, Jinzuomu, et al.
Pubblicazione: (2025)
Activation Steering for Accent Adaptation in Speech Foundation Models
di: Sun, Jinuo, et al.
Pubblicazione: (2026)
di: Sun, Jinuo, et al.
Pubblicazione: (2026)
CPT-Boosted Wav2vec2.0: Towards Noise Robust Speech Recognition for Classroom Environments
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2024)
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2024)
Rethinking Discrete Speech Representation Tokens for Accent Generation
di: Zhong, Jinzuomu, et al.
Pubblicazione: (2026)
di: Zhong, Jinzuomu, et al.
Pubblicazione: (2026)
Simulating Articulatory Trajectories with Phonological Feature Interpolation
di: Tandazo, Angelo Ortiz, et al.
Pubblicazione: (2024)
di: Tandazo, Angelo Ortiz, et al.
Pubblicazione: (2024)
Activation Steering for Accent-Neutralized Zero-Shot Text-To-Speech
di: Yang, Mu, et al.
Pubblicazione: (2026)
di: Yang, Mu, et al.
Pubblicazione: (2026)
Non-autoregressive real-time Accent Conversion model with voice cloning
di: Nechaev, Vladimir, et al.
Pubblicazione: (2024)
di: Nechaev, Vladimir, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Improving Speech Inversion Through Self-Supervised Embeddings and Enhanced Tract Variables
di: Attia, Ahmed Adel, et al.
Pubblicazione: (2023) -
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
di: Premananth, Gowtham, et al.
Pubblicazione: (2025) -
A multi-modal approach for identifying schizophrenia using cross-modal attention
di: Premananth, Gowtham, et al.
Pubblicazione: (2023) -
A Multimodal Framework for the Assessment of the Schizophrenia Spectrum
di: Premananth, Gowtham, et al.
Pubblicazione: (2024) -
Quantifying Articulatory Coordination as a Biomarker for Schizophrenia
di: Premananth, Gowtham, et al.
Pubblicazione: (2025)