ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kumar, Subham, Shivaprakash, Prakrithi, Manoharan, Abhishek, Kurariya, Astut, Mukherjee, Diptadhi, Shukla, Lekhansh, Mukherjee, Animesh, Chand, Prabhat, Murthy, Pratima |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lost without translation -- Can transformer (language models) understand mood states?
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
Benchmarking Motivational Interviewing Competence of Large Language Models
von: Jha, Aishwariya, et al.
Veröffentlicht: (2026)
von: Jha, Aishwariya, et al.
Veröffentlicht: (2026)
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
von: Haritha Gireesh, et al.
Veröffentlicht: (2025)
von: Haritha Gireesh, et al.
Veröffentlicht: (2025)
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
von: Rai, Anand, et al.
Veröffentlicht: (2025)
von: Rai, Anand, et al.
Veröffentlicht: (2025)
Speech Emotion Recognition with ASR Integration
von: Li, Yuanchao
Veröffentlicht: (2026)
von: Li, Yuanchao
Veröffentlicht: (2026)
BR-ASR: Efficient and Scalable Bias Retrieval Framework for Contextual Biasing ASR in Speech LLM
von: Gong, Xun, et al.
Veröffentlicht: (2025)
von: Gong, Xun, et al.
Veröffentlicht: (2025)
Streaming Speech-to-Confusion Network Speech Recognition
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
Contextual Biasing for ASR in Speech LLM with Common Word Cues and Bias Word Position Prediction
von: Novitasari, Sashi, et al.
Veröffentlicht: (2026)
von: Novitasari, Sashi, et al.
Veröffentlicht: (2026)
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition
von: Kim, Jongsuk, et al.
Veröffentlicht: (2025)
von: Kim, Jongsuk, et al.
Veröffentlicht: (2025)
A Unified Framework for Collecting Text-to-Speech Synthesis Datasets for 22 Indian Languages
von: Sathiyamoorthy, Sujitha, et al.
Veröffentlicht: (2024)
von: Sathiyamoorthy, Sujitha, et al.
Veröffentlicht: (2024)
Elevating Robust Multi-Talker ASR by Decoupling Speaker Separation and Speech Recognition
von: Yang, Yufeng, et al.
Veröffentlicht: (2025)
von: Yang, Yufeng, et al.
Veröffentlicht: (2025)
Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses
von: Yang, Yufeng, et al.
Veröffentlicht: (2025)
von: Yang, Yufeng, et al.
Veröffentlicht: (2025)
AS-ASR: A Lightweight Framework for Aphasia-Specific Automatic Speech Recognition
von: Bao, Chen, et al.
Veröffentlicht: (2025)
von: Bao, Chen, et al.
Veröffentlicht: (2025)
Improving ASR Contextual Biasing with Guided Attention
von: Tang, Jiyang, et al.
Veröffentlicht: (2024)
von: Tang, Jiyang, et al.
Veröffentlicht: (2024)
dLLM-ASR: A Faster Diffusion LLM-based Framework for Speech Recognition
von: Tian, Wenjie, et al.
Veröffentlicht: (2026)
von: Tian, Wenjie, et al.
Veröffentlicht: (2026)
SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding
von: Wei, Linye, et al.
Veröffentlicht: (2025)
von: Wei, Linye, et al.
Veröffentlicht: (2025)
Contextual Biasing for Streaming ASR via CTC-based Word Spotting
von: Tsai, Kai-Chen, et al.
Veröffentlicht: (2026)
von: Tsai, Kai-Chen, et al.
Veröffentlicht: (2026)
Contextual Biasing for LLM-Based ASR with Hotword Retrieval and Reinforcement Learning
von: Kong, YuXiang, et al.
Veröffentlicht: (2025)
von: Kong, YuXiang, et al.
Veröffentlicht: (2025)
Speech Recognition on TV Series with Video-guided Post-ASR Correction
von: Yang, Haoyuan, et al.
Veröffentlicht: (2025)
von: Yang, Haoyuan, et al.
Veröffentlicht: (2025)
ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark
von: Wang, He, et al.
Veröffentlicht: (2025)
von: Wang, He, et al.
Veröffentlicht: (2025)
Self-supervised ASR Models and Features For Dysarthric and Elderly Speech Recognition
von: Hu, Shujie, et al.
Veröffentlicht: (2024)
von: Hu, Shujie, et al.
Veröffentlicht: (2024)
LCB-net: Long-Context Biasing for Audio-Visual Speech Recognition
von: Yu, Fan, et al.
Veröffentlicht: (2024)
von: Yu, Fan, et al.
Veröffentlicht: (2024)
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
Selective Invocation for Multilingual ASR: A Cost-effective Approach Adapting to Speech Recognition Difficulty
von: Xue, Hongfei, et al.
Veröffentlicht: (2025)
von: Xue, Hongfei, et al.
Veröffentlicht: (2025)
Towards Robust Dysarthric Speech Recognition: LLM-Agent Post-ASR Correction Beyond WER
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2026)
von: Zheng, Xiuwen, et al.
Veröffentlicht: (2026)
ASR for Affective Speech: Investigating Impact of Emotion and Speech Generative Strategy
von: Wu, Ya-Tse, et al.
Veröffentlicht: (2026)
von: Wu, Ya-Tse, et al.
Veröffentlicht: (2026)
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models
von: Feng, Chen, et al.
Veröffentlicht: (2025)
von: Feng, Chen, et al.
Veröffentlicht: (2025)
Lightweight Prompt Biasing for Contextualized End-to-End ASR Systems
von: Ren, Bo, et al.
Veröffentlicht: (2025)
von: Ren, Bo, et al.
Veröffentlicht: (2025)
Initial Decoding with Minimally Augmented Language Model for Improved Lattice Rescoring in Low Resource ASR
von: Murthy, Savitha, et al.
Veröffentlicht: (2024)
von: Murthy, Savitha, et al.
Veröffentlicht: (2024)
EfficientASR: Speech Recognition Network Compression via Attention Redundancy and Chunk-Level FFN Optimization
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024)
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024)
AdaLTM: Adaptive Layer-wise Task Vector Merging for Categorical Speech Emotion Recognition with ASR Knowledge Integration
von: Lee, Chia-Yu, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Yu, et al.
Veröffentlicht: (2026)
WCTC-Biasing: Retraining-free Contextual Biasing ASR with Wildcard CTC-based Keyword Spotting and Inter-layer Biasing
von: Nakagome, Yu, et al.
Veröffentlicht: (2025)
von: Nakagome, Yu, et al.
Veröffentlicht: (2025)
LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximation
von: Kamahori, Keisuke, et al.
Veröffentlicht: (2025)
von: Kamahori, Keisuke, et al.
Veröffentlicht: (2025)
Bridging ASR and LLMs for Dysarthric Speech Recognition: Benchmarking Self-Supervised and Generative Approaches
von: Aboeitta, Ahmed, et al.
Veröffentlicht: (2025)
von: Aboeitta, Ahmed, et al.
Veröffentlicht: (2025)
A Comprehensive Study on the Effectiveness of ASR Representations for Noise-Robust Speech Emotion Recognition
von: Shi, Xiaohan, et al.
Veröffentlicht: (2023)
von: Shi, Xiaohan, et al.
Veröffentlicht: (2023)
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
von: Wang, He, et al.
Veröffentlicht: (2024)
von: Wang, He, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Lost without translation -- Can transformer (language models) understand mood states?
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025) -
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025) -
Benchmarking Motivational Interviewing Competence of Large Language Models
von: Jha, Aishwariya, et al.
Veröffentlicht: (2026) -
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
von: Haritha Gireesh, et al.
Veröffentlicht: (2025) -
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
von: Kumar, Subham, et al.
Veröffentlicht: (2025)