PROFASR-BENCH: A Benchmark for Context-Conditioned ASR in High-Stakes Professional Speech
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Piskala, Deepak Babu |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark
par: Wang, He, et autres
Publié: (2025)
par: Wang, He, et autres
Publié: (2025)
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
par: Piskala, Deepak Babu
Publié: (2026)
par: Piskala, Deepak Babu
Publié: (2026)
Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark
par: Turetzky, Arnon, et autres
Publié: (2026)
par: Turetzky, Arnon, et autres
Publié: (2026)
Nwāchā Munā: A Devanagari Speech Corpus and Proximal Transfer Benchmark for Nepal Bhasha ASR
par: Sharma, Rishikesh Kumar, et autres
Publié: (2026)
par: Sharma, Rishikesh Kumar, et autres
Publié: (2026)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
par: Rai, Anand, et autres
Publié: (2025)
par: Rai, Anand, et autres
Publié: (2025)
Benchmarking Children's ASR with Supervised and Self-supervised Speech Foundation Models
par: Fan, Ruchao, et autres
Publié: (2024)
par: Fan, Ruchao, et autres
Publié: (2024)
Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR
par: Lee, Minsik, et autres
Publié: (2026)
par: Lee, Minsik, et autres
Publié: (2026)
RO-N3WS: Enhancing Generalization in Low-Resource ASR with Diverse Romanian Speech Benchmarks
par: Diaconu, Alexandra, et autres
Publié: (2026)
par: Diaconu, Alexandra, et autres
Publié: (2026)
SW-ASR: A Context-Aware Hybrid ASR Pipeline for Robust Single Word Speech Recognition
par: Sharma, Manali, et autres
Publié: (2026)
par: Sharma, Manali, et autres
Publié: (2026)
Uni-ASR: Unified LLM-Based Architecture for Non-Streaming and Streaming Automatic Speech Recognition
par: Xia, Yinfeng, et autres
Publié: (2026)
par: Xia, Yinfeng, et autres
Publié: (2026)
CLiFT-ASR: A Cross-Lingual Fine-Tuning Framework for Low-Resource Taiwanese Hokkien Speech Recognition
par: Sung, Hung-Yang, et autres
Publié: (2025)
par: Sung, Hung-Yang, et autres
Publié: (2025)
Configurable Multilingual ASR with Speech Summary Representations
par: Zhu, Harrison, et autres
Publié: (2024)
par: Zhu, Harrison, et autres
Publié: (2024)
Benchmarking Japanese Speech Recognition on ASR-LLM Setups with Multi-Pass Augmented Generative Error Correction
par: Ko, Yuka, et autres
Publié: (2024)
par: Ko, Yuka, et autres
Publié: (2024)
CantoASR: Prosody-Aware ASR-LALM Collaboration for Low-Resource Cantonese
par: Chen, Dazhong, et autres
Publié: (2025)
par: Chen, Dazhong, et autres
Publié: (2025)
Semi-Autoregressive Streaming ASR With Label Context
par: Arora, Siddhant, et autres
Publié: (2023)
par: Arora, Siddhant, et autres
Publié: (2023)
Performant ASR Models for Medical Entities in Accented Speech
par: Afonja, Tejumade, et autres
Publié: (2024)
par: Afonja, Tejumade, et autres
Publié: (2024)
Crossmodal ASR Error Correction with Discrete Speech Units
par: Li, Yuanchao, et autres
Publié: (2024)
par: Li, Yuanchao, et autres
Publié: (2024)
ASR-EC Benchmark: Evaluating Large Language Models on Chinese ASR Error Correction
par: Wei, Victor Junqiu, et autres
Publié: (2024)
par: Wei, Victor Junqiu, et autres
Publié: (2024)
Vedavani: A Benchmark Corpus for ASR on Vedic Sanskrit Poetry
par: Kumar, Sujeet, et autres
Publié: (2025)
par: Kumar, Sujeet, et autres
Publié: (2025)
PRiSM: Benchmarking Phone Realization in Speech Models
par: Bharadwaj, Shikhar, et autres
Publié: (2026)
par: Bharadwaj, Shikhar, et autres
Publié: (2026)
Moonshine v2: Ergodic Streaming Encoder ASR for Latency-Critical Speech Applications
par: Kudlur, Manjunath, et autres
Publié: (2026)
par: Kudlur, Manjunath, et autres
Publié: (2026)
WER We Stand: Benchmarking Urdu ASR Models
par: Arif, Samee, et autres
Publié: (2024)
par: Arif, Samee, et autres
Publié: (2024)
AsyncSwitch: Asynchronous Text-Speech Adaptation for Code-Switched ASR
par: Nguyen, Tuan, et autres
Publié: (2025)
par: Nguyen, Tuan, et autres
Publié: (2025)
Exploring SSL Discrete Speech Features for Zipformer-based Contextual ASR
par: Cui, Mingyu, et autres
Publié: (2024)
par: Cui, Mingyu, et autres
Publié: (2024)
Interventional Speech Noise Injection for ASR Generalizable Spoken Language Understanding
par: Jung, Yeonjoon, et autres
Publié: (2024)
par: Jung, Yeonjoon, et autres
Publié: (2024)
Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition
par: Wang, Peng, et autres
Publié: (2026)
par: Wang, Peng, et autres
Publié: (2026)
A Comprehensive Study on the Effectiveness of ASR Representations for Noise-Robust Speech Emotion Recognition
par: Shi, Xiaohan, et autres
Publié: (2023)
par: Shi, Xiaohan, et autres
Publié: (2023)
Echotune: A Modular Extractor Leveraging the Variable-Length Nature of Speech in ASR Tasks
par: Chen, Sizhou, et autres
Publié: (2023)
par: Chen, Sizhou, et autres
Publié: (2023)
ASR Benchmarking: Need for a More Representative Conversational Dataset
par: Maheshwari, Gaurav, et autres
Publié: (2024)
par: Maheshwari, Gaurav, et autres
Publié: (2024)
MLMA: Towards Multilingual ASR With Mamba-based Architectures
par: Ali, Mohamed Nabih, et autres
Publié: (2025)
par: Ali, Mohamed Nabih, et autres
Publié: (2025)
PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech
par: Menta, Venkata Pushpak Teja
Publié: (2026)
par: Menta, Venkata Pushpak Teja
Publié: (2026)
Mind the Gap: Entity-Preserved Context-Aware ASR Structured Transcriptions
par: Altinok, Duygu
Publié: (2025)
par: Altinok, Duygu
Publié: (2025)
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
par: Mdhaffar, Salima, et autres
Publié: (2024)
par: Mdhaffar, Salima, et autres
Publié: (2024)
Bridging Speech and Text: Enhancing ASR with Pinyin-to-Character Pre-training in LLMs
par: Yuhang, Yang, et autres
Publié: (2024)
par: Yuhang, Yang, et autres
Publié: (2024)
Exploring Pathological Speech Quality Assessment with ASR-Powered Wav2Vec2 in Data-Scarce Context
par: Nguyen, Tuan, et autres
Publié: (2024)
par: Nguyen, Tuan, et autres
Publié: (2024)
LOTUSDIS: A Thai far-field meeting corpus for robust conversational ASR
par: Tipaksorn, Pattara, et autres
Publié: (2025)
par: Tipaksorn, Pattara, et autres
Publié: (2025)
A Calculus-Based Framework for Determining Vocabulary Size in End-to-End ASR
par: Kopparapu, Sunil Kumar
Publié: (2026)
par: Kopparapu, Sunil Kumar
Publié: (2026)
SVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions
par: Baali, Massa, et autres
Publié: (2025)
par: Baali, Massa, et autres
Publié: (2025)
AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition
par: Awobade, Busayo, et autres
Publié: (2026)
par: Awobade, Busayo, et autres
Publié: (2026)
SloPal: A 60-Million-Word Slovak Parliamentary Corpus with Aligned Speech and Fine-Tuned ASR Models
par: Božík, Erik, et autres
Publié: (2025)
par: Božík, Erik, et autres
Publié: (2025)
Documents similaires
-
ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark
par: Wang, He, et autres
Publié: (2025) -
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
par: Piskala, Deepak Babu
Publié: (2026) -
Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark
par: Turetzky, Arnon, et autres
Publié: (2026) -
Nwāchā Munā: A Devanagari Speech Corpus and Proximal Transfer Benchmark for Nepal Bhasha ASR
par: Sharma, Rishikesh Kumar, et autres
Publié: (2026) -
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
par: Rai, Anand, et autres
Publié: (2025)