Enhancing AAC Software for Dysarthric Speakers in e-Health Settings: An Evaluation Using TORGO
Fuente:
arXiv
Salvato in:
| Autori principali: | Hui, Macarious, Zhang, Jinda, Mohan, Aanchan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
di: Park, Seohyun, et al.
Pubblicazione: (2025)
di: Park, Seohyun, et al.
Pubblicazione: (2025)
Homogeneous Speaker Features for On-the-Fly Dysarthric and Elderly Speaker Adaptation
di: Geng, Mengzhe, et al.
Pubblicazione: (2024)
di: Geng, Mengzhe, et al.
Pubblicazione: (2024)
UltrasonicSpheres: Localized, Multi-Channel Sound Spheres Using Off-the-Shelf Speakers and Earables
di: Küttner, Michael, et al.
Pubblicazione: (2025)
di: Küttner, Michael, et al.
Pubblicazione: (2025)
FeatureSense: Protecting Speaker Attributes in Always-On Audio Sensing System
di: Chhaglani, Bhawana, et al.
Pubblicazione: (2025)
di: Chhaglani, Bhawana, et al.
Pubblicazione: (2025)
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
di: Cheng, Xize, et al.
Pubblicazione: (2025)
di: Cheng, Xize, et al.
Pubblicazione: (2025)
Real-time and Continuous Turn-taking Prediction Using Voice Activity Projection
di: Inoue, Koji, et al.
Pubblicazione: (2024)
di: Inoue, Koji, et al.
Pubblicazione: (2024)
The ICASSP 2026 HumDial Challenge: Benchmarking Human-like Spoken Dialogue Systems in the LLM Era
di: Zhao, Zhixian, et al.
Pubblicazione: (2026)
di: Zhao, Zhixian, et al.
Pubblicazione: (2026)
Enhancing Dysarthric Speech Recognition for Unseen Speakers via Prototype-Based Adaptation
di: Wang, Shiyao, et al.
Pubblicazione: (2024)
di: Wang, Shiyao, et al.
Pubblicazione: (2024)
Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection
di: Inoue, Koji, et al.
Pubblicazione: (2024)
di: Inoue, Koji, et al.
Pubblicazione: (2024)
Lla-VAP: LSTM Ensemble of Llama and VAP for Turn-Taking Prediction
di: Jeon, Hyunbae, et al.
Pubblicazione: (2024)
di: Jeon, Hyunbae, et al.
Pubblicazione: (2024)
Are Expressions for Music Emotions the Same Across Cultures?
di: Celen, Elif, et al.
Pubblicazione: (2025)
di: Celen, Elif, et al.
Pubblicazione: (2025)
Speech vs. Transcript: Does It Matter for Human Annotators in Speech Summarization?
di: Sharma, Roshan, et al.
Pubblicazione: (2024)
di: Sharma, Roshan, et al.
Pubblicazione: (2024)
InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training
di: Wang, Dingdong, et al.
Pubblicazione: (2025)
di: Wang, Dingdong, et al.
Pubblicazione: (2025)
MCMChaos: Improvising Rap Music with MCMC Methods and Chaos Theory
di: Kimelman, Robert G.
Pubblicazione: (2024)
di: Kimelman, Robert G.
Pubblicazione: (2024)
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
di: Uro, Rémi, et al.
Pubblicazione: (2024)
di: Uro, Rémi, et al.
Pubblicazione: (2024)
Investigating the Effects of Large-Scale Pseudo-Stereo Data and Different Speech Foundation Model on Dialogue Generative Spoken Language Model
di: Fu, Yu-Kuan, et al.
Pubblicazione: (2024)
di: Fu, Yu-Kuan, et al.
Pubblicazione: (2024)
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
di: Yan, Hui, et al.
Pubblicazione: (2024)
di: Yan, Hui, et al.
Pubblicazione: (2024)
Interactive Sonification for Health and Energy using ChucK and Unity
di: Zhao, Yichun, et al.
Pubblicazione: (2024)
di: Zhao, Yichun, et al.
Pubblicazione: (2024)
Evaluating Spatialized Auditory Cues for Rapid Attention Capture in XR
di: Kim, Yoonsang, et al.
Pubblicazione: (2026)
di: Kim, Yoonsang, et al.
Pubblicazione: (2026)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
di: Yu, Luca Jiang-Tao, et al.
Pubblicazione: (2024)
di: Yu, Luca Jiang-Tao, et al.
Pubblicazione: (2024)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
di: Zheng, Shuoyang, et al.
Pubblicazione: (2024)
di: Zheng, Shuoyang, et al.
Pubblicazione: (2024)
Using Confidence Scores to Improve Eyes-free Detection of Speech Recognition Errors
di: Nowrin, Sadia, et al.
Pubblicazione: (2024)
di: Nowrin, Sadia, et al.
Pubblicazione: (2024)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
di: Manukalpa, J. M. Chan Sri, et al.
Pubblicazione: (2025)
di: Manukalpa, J. M. Chan Sri, et al.
Pubblicazione: (2025)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
di: Mishra, Ruchik, et al.
Pubblicazione: (2024)
di: Mishra, Ruchik, et al.
Pubblicazione: (2024)
Towards Reliable Large Audio Language Model
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection
di: Li, Lu, et al.
Pubblicazione: (2025)
di: Li, Lu, et al.
Pubblicazione: (2025)
EmoKnob: Enhance Voice Cloning with Fine-Grained Emotion Control
di: Chen, Haozhe, et al.
Pubblicazione: (2024)
di: Chen, Haozhe, et al.
Pubblicazione: (2024)
Beyond-Voice: Towards Continuous 3D Hand Pose Tracking on Commercial Home Assistant Devices
di: Li, Yin, et al.
Pubblicazione: (2023)
di: Li, Yin, et al.
Pubblicazione: (2023)
SingVisio: Visual Analytics of Diffusion Model for Singing Voice Conversion
di: Xue, Liumeng, et al.
Pubblicazione: (2024)
di: Xue, Liumeng, et al.
Pubblicazione: (2024)
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
di: Han, Zhichen, et al.
Pubblicazione: (2024)
di: Han, Zhichen, et al.
Pubblicazione: (2024)
AIx Speed: Playback Speed Optimization Using Listening Comprehension of Speech Recognition Models
di: Kawamura, Kazuki, et al.
Pubblicazione: (2024)
di: Kawamura, Kazuki, et al.
Pubblicazione: (2024)
Advancing User-Voice Interaction: Exploring Emotion-Aware Voice Assistants Through a Role-Swapping Approach
di: Ma, Yong, et al.
Pubblicazione: (2025)
di: Ma, Yong, et al.
Pubblicazione: (2025)
WSCoach: Wearable Real-time Auditory Feedback for Reducing Unwanted Words in Daily Communication
di: Youpeng, Zhang, et al.
Pubblicazione: (2025)
di: Youpeng, Zhang, et al.
Pubblicazione: (2025)
ListenNet: A Lightweight Spatio-Temporal Enhancement Nested Network for Auditory Attention Detection
di: Fan, Cunhang, et al.
Pubblicazione: (2025)
di: Fan, Cunhang, et al.
Pubblicazione: (2025)
Artificial Neural Networks to Recognize Speakers Division from Continuous Bengali Speech
di: Ali, Hasmot, et al.
Pubblicazione: (2024)
di: Ali, Hasmot, et al.
Pubblicazione: (2024)
Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models
di: Dietrich, Juergen
Pubblicazione: (2026)
di: Dietrich, Juergen
Pubblicazione: (2026)
Gammatonegram Representation for End-to-End Dysarthric Speech Processing Tasks: Speech Recognition, Speaker Identification, and Intelligibility Assessment
di: Farhadipour, Aref, et al.
Pubblicazione: (2023)
di: Farhadipour, Aref, et al.
Pubblicazione: (2023)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
di: Park, Seohyun, et al.
Pubblicazione: (2025) -
Homogeneous Speaker Features for On-the-Fly Dysarthric and Elderly Speaker Adaptation
di: Geng, Mengzhe, et al.
Pubblicazione: (2024) -
UltrasonicSpheres: Localized, Multi-Channel Sound Spheres Using Off-the-Shelf Speakers and Earables
di: Küttner, Michael, et al.
Pubblicazione: (2025) -
FeatureSense: Protecting Speaker Attributes in Always-On Audio Sensing System
di: Chhaglani, Bhawana, et al.
Pubblicazione: (2025) -
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
di: Cheng, Xize, et al.
Pubblicazione: (2025)