Demonstration of Adapt4Me: An Uncertainty-Aware Authoring Environment for Personalizing Automatic Speech Recognition to Non-normative Speech
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pokel, Niclas, Zhao, Yiming, Moure, Pehuén, Gao, Yingqiang, Böhringer, Roman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Data-Efficient ASR Personalization for Non-Normative Speech Using an Uncertainty-Based Phoneme Difficulty Score for Guided Sampling
von: Pokel, Niclas, et al.
Veröffentlicht: (2025)
von: Pokel, Niclas, et al.
Veröffentlicht: (2025)
Adapting Foundation Speech Recognition Models to Impaired Speech: A Semantic Re-chaining Approach for Personalization of German Speech
von: Pokel, Niclas, et al.
Veröffentlicht: (2025)
von: Pokel, Niclas, et al.
Veröffentlicht: (2025)
Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition
von: Pokel, Niclas, et al.
Veröffentlicht: (2025)
von: Pokel, Niclas, et al.
Veröffentlicht: (2025)
When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition
von: Moure, Pehuén, et al.
Veröffentlicht: (2026)
von: Moure, Pehuén, et al.
Veröffentlicht: (2026)
Challenges in Automatic Speech Recognition for Adults with Cognitive Impairment
von: Cohn, Michelle, et al.
Veröffentlicht: (2026)
von: Cohn, Michelle, et al.
Veröffentlicht: (2026)
Adapting Whisper for Lightweight and Efficient Automatic Speech Recognition of Children for On-device Edge Applications
von: Dutta, Satwik, et al.
Veröffentlicht: (2025)
von: Dutta, Satwik, et al.
Veröffentlicht: (2025)
Lost in Transcription: Subtitle Errors in Automatic Speech Recognition Reduce Speaker and Content Evaluations
von: Kadoma, Kowe, et al.
Veröffentlicht: (2026)
von: Kadoma, Kowe, et al.
Veröffentlicht: (2026)
Hardware-Aware Federated Learning for Speech Emotion Recognition
von: Yuksel, Beyazit Bestami, et al.
Veröffentlicht: (2026)
von: Yuksel, Beyazit Bestami, et al.
Veröffentlicht: (2026)
Feel my Speech: Automatic Speech Emotion Conversion for Tangible, Haptic, or Proxemic Interaction Design
von: Aslan, Ilhan
Veröffentlicht: (2024)
von: Aslan, Ilhan
Veröffentlicht: (2024)
Tap-to-Adapt: Learning User-Aligned Response Timing for Speech Agents
von: He, Zihong, et al.
Veröffentlicht: (2026)
von: He, Zihong, et al.
Veröffentlicht: (2026)
Interactive Cycle Model: The Linkage Combination among Automatic Speech Recognition, Large Language Models and Smart Glasses
von: Wang, Libo
Veröffentlicht: (2024)
von: Wang, Libo
Veröffentlicht: (2024)
CapTune: Adapting Non-Speech Captions With Anchored Generative Models
von: Huang, Jeremy Zhengqi, et al.
Veröffentlicht: (2025)
von: Huang, Jeremy Zhengqi, et al.
Veröffentlicht: (2025)
M4SER: Multimodal, Multirepresentation, Multitask, and Multistrategy Learning for Speech Emotion Recognition
von: He, Jiajun, et al.
Veröffentlicht: (2025)
von: He, Jiajun, et al.
Veröffentlicht: (2025)
PersonaMail: Learning and Adapting Personal Communication Preferences for Context-Aware Email Writing
von: Yao, Rui, et al.
Veröffentlicht: (2026)
von: Yao, Rui, et al.
Veröffentlicht: (2026)
From Delays to Densities: Exploring Data Uncertainty through Speech, Text, and Visualization
von: Stokes, Chase, et al.
Veröffentlicht: (2024)
von: Stokes, Chase, et al.
Veröffentlicht: (2024)
Speech Command + Speech Emotion: Exploring Emotional Speech Commands as a Compound and Playful Modality
von: Aslan, Ilhan, et al.
Veröffentlicht: (2025)
von: Aslan, Ilhan, et al.
Veröffentlicht: (2025)
Talk Me Through It: Developing Effective Systems for Chart Authoring
von: Ponochevnyi, Nazar, et al.
Veröffentlicht: (2026)
von: Ponochevnyi, Nazar, et al.
Veröffentlicht: (2026)
Improved Dysarthric Speech to Text Conversion via TTS Personalization
von: Mihajlik, Péter, et al.
Veröffentlicht: (2025)
von: Mihajlik, Péter, et al.
Veröffentlicht: (2025)
Confides: A Visual Analytics Solution for Automated Speech Recognition Analysis and Exploration
von: Ha, Sunwoo, et al.
Veröffentlicht: (2024)
von: Ha, Sunwoo, et al.
Veröffentlicht: (2024)
Speejis: Enhancing User Experience of Mobile Voice Messaging with Automatic Visual Speech Emotion Cues
von: Aslan, Ilhan, et al.
Veröffentlicht: (2025)
von: Aslan, Ilhan, et al.
Veröffentlicht: (2025)
Mixing Modes: Active and Passive Integration of Speech, Text, and Visualization for Communicating Data Uncertainty
von: Stokes, Chase, et al.
Veröffentlicht: (2024)
von: Stokes, Chase, et al.
Veröffentlicht: (2024)
Take the Power Back: Screen-Based Personal Moderation Against Hate Speech on Instagram
von: Luther, Anna Ricarda, et al.
Veröffentlicht: (2026)
von: Luther, Anna Ricarda, et al.
Veröffentlicht: (2026)
A Personalized and Adaptable User Interface for a Speech and Cursor Brain-Computer Interface
von: Peracha, Hamza, et al.
Veröffentlicht: (2026)
von: Peracha, Hamza, et al.
Veröffentlicht: (2026)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
von: Mishra, Ruchik, et al.
Veröffentlicht: (2024)
von: Mishra, Ruchik, et al.
Veröffentlicht: (2024)
HandProxy: Expanding the Affordances of Speech Interfaces in Immersive Environments with a Virtual Proxy Hand
von: Liang, Chen, et al.
Veröffentlicht: (2025)
von: Liang, Chen, et al.
Veröffentlicht: (2025)
Voicing Uncertainty: How Speech, Text, and Visualizations Influence Decisions with Data Uncertainty
von: Stokes, Chase, et al.
Veröffentlicht: (2024)
von: Stokes, Chase, et al.
Veröffentlicht: (2024)
TalkLess: Blending Extractive and Abstractive Speech Summarization for Editing Speech to Preserve Content and Style
von: Benharrak, Karim, et al.
Veröffentlicht: (2025)
von: Benharrak, Karim, et al.
Veröffentlicht: (2025)
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)
Adapting to the User: A Systematic Review of Personalized Interaction in VR
von: Li, Tangyao, et al.
Veröffentlicht: (2025)
von: Li, Tangyao, et al.
Veröffentlicht: (2025)
DG Comics: Semi-Automatically Authoring Graph Comics for Dynamic Graphs
von: Kim, Joohee, et al.
Veröffentlicht: (2024)
von: Kim, Joohee, et al.
Veröffentlicht: (2024)
Automatic Authoring of Physical and Perceptual/Affective Motion Effects for Virtual Reality
von: Lee, Jiwan, et al.
Veröffentlicht: (2024)
von: Lee, Jiwan, et al.
Veröffentlicht: (2024)
"Help Me, But Don't Track Me": Intervention Timing and Privacy Boundaries for Process-Aware AI Tutors
von: Li, Jane Hanqi, et al.
Veröffentlicht: (2026)
von: Li, Jane Hanqi, et al.
Veröffentlicht: (2026)
Speech-based Mark for Data Sonification
von: Zhao, Yichun, et al.
Veröffentlicht: (2024)
von: Zhao, Yichun, et al.
Veröffentlicht: (2024)
Child Speech Recognition in Human-Robot Interaction: Problem Solved?
von: Janssens, Ruben, et al.
Veröffentlicht: (2024)
von: Janssens, Ruben, et al.
Veröffentlicht: (2024)
Communication Access Real-Time Translation Through Collaborative Correction of Automatic Speech Recognition
von: Kuhn, Korbinian, et al.
Veröffentlicht: (2025)
von: Kuhn, Korbinian, et al.
Veröffentlicht: (2025)
A Unified Editing Method for Co-Speech Gesture Generation via Diffusion Inversion
von: Zhao, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhao, Zeyu, et al.
Veröffentlicht: (2024)
"How to Explore Biases in Speech Emotion AI with Users?" A Speech-Emotion-Acting Study Exploring Age and Language Biases
von: Borre, Josephine Beatrice Skovbo, et al.
Veröffentlicht: (2025)
von: Borre, Josephine Beatrice Skovbo, et al.
Veröffentlicht: (2025)
Speech-driven Personalized Gesture Synthetics: Harnessing Automatic Fuzzy Feature Inference
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
Toward using Speech to Sense Student Emotion in Remote Learning Environments
von: Vyas, Sargam, et al.
Veröffentlicht: (2026)
von: Vyas, Sargam, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Scarf Plots
von: Pathmanathan, Nelusa, et al.
Veröffentlicht: (2025)
von: Pathmanathan, Nelusa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Data-Efficient ASR Personalization for Non-Normative Speech Using an Uncertainty-Based Phoneme Difficulty Score for Guided Sampling
von: Pokel, Niclas, et al.
Veröffentlicht: (2025) -
Adapting Foundation Speech Recognition Models to Impaired Speech: A Semantic Re-chaining Approach for Personalization of German Speech
von: Pokel, Niclas, et al.
Veröffentlicht: (2025) -
Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition
von: Pokel, Niclas, et al.
Veröffentlicht: (2025) -
When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition
von: Moure, Pehuén, et al.
Veröffentlicht: (2026) -
Challenges in Automatic Speech Recognition for Adults with Cognitive Impairment
von: Cohn, Michelle, et al.
Veröffentlicht: (2026)