The Interspeech 2025 Speech Accessibility Project Challenge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Xiuwen, Phukon, Bornali, Na, Jonghwan, Cutrell, Ed, Han, Kyu, Hasegawa-Johnson, Mark, Jiang, Pan-Pan, Kuila, Aadhrik, Lea, Colin, MacDonald, Bob, Mantena, Gautam, Ravichandran, Venkatesh, Sari, Leda, Tomanek, Katrin, Yoo, Chang D., Zwilling, Chris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fine-Tuning Automatic Speech Recognition for People with Parkinson's: An Effective Strategy for Enhancing Speech Technology Accessibility
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2024)
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2024)
Aligning ASR Evaluation with Human and LLM Judgments: Intelligibility Metrics Using Phonetic, Semantic, and NLI Approaches
von: Phukon, Bornali, et al.
Veröffentlicht: (2025)
von: Phukon, Bornali, et al.
Veröffentlicht: (2025)
Towards Robust Dysarthric Speech Recognition: LLM-Agent Post-ASR Correction Beyond WER
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2026)
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2026)
Towards a Single ASR Model That Generalizes to Disordered Speech
von: Tobin, Jimmy, et al.
Veröffentlicht: (2024)
von: Tobin, Jimmy, et al.
Veröffentlicht: (2024)
Hypernetworks for Personalizing ASR to Atypical Speech
von: Müller-Eberstein, Max, et al.
Veröffentlicht: (2024)
von: Müller-Eberstein, Max, et al.
Veröffentlicht: (2024)
Learnings from curating a trustworthy, well-annotated, and useful dataset of disordered English speech
von: Jiang, Pan-Pan, et al.
Veröffentlicht: (2024)
von: Jiang, Pan-Pan, et al.
Veröffentlicht: (2024)
Reassessing Active Learning Adoption in Contemporary NLP: A Community Survey
von: Romberg, Julia, et al.
Veröffentlicht: (2025)
von: Romberg, Julia, et al.
Veröffentlicht: (2025)
LaurenLWhite/IS25: Interspeech 2025 Repo
von: Lauren White
Veröffentlicht: (2025)
von: Lauren White
Veröffentlicht: (2025)
Interspeech 2025 URGENT Speech Enhancement Challenge
von: Saijo, Kohei, et al.
Veröffentlicht: (2025)
von: Saijo, Kohei, et al.
Veröffentlicht: (2025)
Something from Nothing: Data Augmentation for Robust Severity Level Estimation of Dysarthric Speech
von: Bae, Jaesung, et al.
Veröffentlicht: (2026)
von: Bae, Jaesung, et al.
Veröffentlicht: (2026)
Topological markers for a one-dimensional fermionic chain coupled to a single-mode cavity
von: Ritz-Zwilling, Anna, et al.
Veröffentlicht: (2026)
von: Ritz-Zwilling, Anna, et al.
Veröffentlicht: (2026)
Speech Recognition With LLMs Adapted to Disordered Speech Using Reinforcement Learning
von: Nagpal, Chirag, et al.
Veröffentlicht: (2024)
von: Nagpal, Chirag, et al.
Veröffentlicht: (2024)
The Interspeech 2024 Challenge on Speech Processing Using Discrete Units
von: Chang, Xuankai, et al.
Veröffentlicht: (2024)
von: Chang, Xuankai, et al.
Veröffentlicht: (2024)
Hearing Between the Lines: Unlocking the Reasoning Power of LLMs for Speech Evaluation
von: Chandra, Arjun, et al.
Veröffentlicht: (2026)
von: Chandra, Arjun, et al.
Veröffentlicht: (2026)
ChartA11y: Designing Accessible Touch Experiences of Visualizations with Blind Smartphone Users
von: Zhang, Zhuohao Jerry, et al.
Veröffentlicht: (2024)
von: Zhang, Zhuohao Jerry, et al.
Veröffentlicht: (2024)
Actual issues of modern science. ESeJ, 24 (2023)
von: Buychik, Alexander, et al.
Veröffentlicht: (2023)
von: Buychik, Alexander, et al.
Veröffentlicht: (2023)
lormaechea/interspeech-2025: Supplementary materials for Interspeech 2025 paper submission
von: Lucía Ormaechea, et al.
Veröffentlicht: (2025)
von: Lucía Ormaechea, et al.
Veröffentlicht: (2025)
From Text to Context: An Entailment Approach for News Stakeholder Classification
von: Kuila, Alapan, et al.
Veröffentlicht: (2024)
von: Kuila, Alapan, et al.
Veröffentlicht: (2024)
Deciphering Political Entity Sentiment in News with Large Language Models: Zero-Shot and Few-Shot Strategies
von: Kuila, Alapan, et al.
Veröffentlicht: (2024)
von: Kuila, Alapan, et al.
Veröffentlicht: (2024)
Cysteine‐Based Dynamic Self‐Assembly and Their Importance in the Origins of Life
von: Soumen Kuila, et al.
Veröffentlicht: (2024)
von: Soumen Kuila, et al.
Veröffentlicht: (2024)
Mimetic Alignment with ASPECT: Evaluation of AI-inferred Personal Profiles
von: Shang, Ruoxi, et al.
Veröffentlicht: (2026)
von: Shang, Ruoxi, et al.
Veröffentlicht: (2026)
TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge
von: Alumäe, Tanel, et al.
Veröffentlicht: (2025)
von: Alumäe, Tanel, et al.
Veröffentlicht: (2025)
The Interspeech 2026 Audio Encoder Capability Challenge for Large Audio Language Models
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2026)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2026)
Delayed Pattern Formation in Two-Dimensional Domains
von: Das, Nirmali Prabha, et al.
Veröffentlicht: (2026)
von: Das, Nirmali Prabha, et al.
Veröffentlicht: (2026)
IQRA 2026: Interspeech Challenge on Automatic Pronunciation Assessment for Modern Standard Arabic (MSA)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2026)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2026)
The influence of Chinese culture and customs on the beliefs and health‐related behaviours of Chinese women with gestational diabetes mellitus: A qualitative study
von: Xiuwen Luo, et al.
Veröffentlicht: (2024)
von: Xiuwen Luo, et al.
Veröffentlicht: (2024)
CJST: CTC Compressor based Joint Speech and Text Training for Decoder-Only ASR
von: Zhou, Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Wei, et al.
Veröffentlicht: (2024)
NTU Speechlab LLM-Based Multilingual ASR System for Interspeech MLC-SLM Challenge 2025
von: Peng, Yizhou, et al.
Veröffentlicht: (2025)
von: Peng, Yizhou, et al.
Veröffentlicht: (2025)
Extended string-net models with all anyons at finite temperature
von: Soares, André O., et al.
Veröffentlicht: (2025)
von: Soares, André O., et al.
Veröffentlicht: (2025)
Alkali‐assisted functionalization of hexagonal boron nitride reinforced epoxy composite to improve the thermomechanical and anticorrosion performance for advanced thermal management
von: Chinmoy Kuila, et al.
Veröffentlicht: (2024)
von: Chinmoy Kuila, et al.
Veröffentlicht: (2024)
Mechanically Strong and Thermally Conductive Zr–BN Hybrid Filler‐Embedded Carbon Fiber‐Reinforced Epoxy Composite for Multifunctional Applications
von: Chinmoy Kuila, et al.
Veröffentlicht: (2025)
von: Chinmoy Kuila, et al.
Veröffentlicht: (2025)
Visible‐light Catalysed Trifluoromethylthiolation and Related Dearomative Spirocyclizations
von: Barnali Roy, et al.
Veröffentlicht: (2024)
von: Barnali Roy, et al.
Veröffentlicht: (2024)
High Yield Synthesis of Spirocyclic Dienones from Phenols Employing Tribromide Catalysed Dearomatization
von: Puspendu Kuila, et al.
Veröffentlicht: (2024)
von: Puspendu Kuila, et al.
Veröffentlicht: (2024)
The X-LANCE Technical Report for Interspeech 2024 Speech Processing Using Discrete Speech Unit Challenge
von: Guo, Yiwei, et al.
Veröffentlicht: (2024)
von: Guo, Yiwei, et al.
Veröffentlicht: (2024)
The Interspeech 2026 Audio Reasoning Challenge: Evaluating Reasoning Process Quality for Audio Reasoning Models and Agents
von: Ma, Ziyang, et al.
Veröffentlicht: (2026)
von: Ma, Ziyang, et al.
Veröffentlicht: (2026)
Detecting Hallucination and Coverage Errors in Retrieval Augmented Generation for Controversial Topics
von: Chang, Tyler A., et al.
Veröffentlicht: (2024)
von: Chang, Tyler A., et al.
Veröffentlicht: (2024)
Pat-DEVAL: Chain-of-Legal-Thought Evaluation for Patent Description
von: Yoo, Yongmin, et al.
Veröffentlicht: (2026)
von: Yoo, Yongmin, et al.
Veröffentlicht: (2026)
UTDUSS: UTokyo-SaruLab System for Interspeech2024 Speech Processing Using Discrete Speech Unit Challenge
von: Nakata, Wataru, et al.
Veröffentlicht: (2024)
von: Nakata, Wataru, et al.
Veröffentlicht: (2024)
WhatsAI: Transforming Meta Ray-Bans into an Extensible Generative AI Platform for Accessibility
von: Zaman, Nasif, et al.
Veröffentlicht: (2025)
von: Zaman, Nasif, et al.
Veröffentlicht: (2025)
RAVEN: Realtime Accessibility in Virtual ENvironments for Blind and Low-Vision People
von: Cao, Xinyun, et al.
Veröffentlicht: (2025)
von: Cao, Xinyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Fine-Tuning Automatic Speech Recognition for People with Parkinson's: An Effective Strategy for Enhancing Speech Technology Accessibility
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2024) -
Aligning ASR Evaluation with Human and LLM Judgments: Intelligibility Metrics Using Phonetic, Semantic, and NLI Approaches
von: Phukon, Bornali, et al.
Veröffentlicht: (2025) -
Towards Robust Dysarthric Speech Recognition: LLM-Agent Post-ASR Correction Beyond WER
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2026) -
Towards a Single ASR Model That Generalizes to Disordered Speech
von: Tobin, Jimmy, et al.
Veröffentlicht: (2024) -
Hypernetworks for Personalizing ASR to Atypical Speech
von: Müller-Eberstein, Max, et al.
Veröffentlicht: (2024)