Causal Structure Discovery for Error Diagnostics of Children's ASR
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Vishwanath Pratap, Sahidullah, Md., Kinnunen, Tomi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ChildAugment: Data Augmentation Methods for Zero-Resource Children's Speaker Verification
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2024)
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2024)
Causal Analysis of ASR Errors for Children: Quantifying the Impact of Physiological, Cognitive, and Extrinsic Factors
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2025)
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2025)
ROAR: Reinforcing Original to Augmented Data Ratio Dynamics for Wav2Vec2.0 Based ASR
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2024)
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2024)
Continuous Learning for Children's ASR: Overcoming Catastrophic Forgetting with Elastic Weight Consolidation and Synaptic Intelligence
di: Ahadzi, Edem, et al.
Pubblicazione: (2025)
di: Ahadzi, Edem, et al.
Pubblicazione: (2025)
Multilingual Source Tracing of Speech Deepfakes: A First Benchmark
di: Xuan, Xi, et al.
Pubblicazione: (2025)
di: Xuan, Xi, et al.
Pubblicazione: (2025)
Generalizing Speaker Verification for Spoof Awareness in the Embedding Space
di: Liu, Xuechen, et al.
Pubblicazione: (2024)
di: Liu, Xuechen, et al.
Pubblicazione: (2024)
Explaining Speaker and Spoof Embeddings via Probing
di: Liu, Xuechen, et al.
Pubblicazione: (2024)
di: Liu, Xuechen, et al.
Pubblicazione: (2024)
Beyond Silence: Bias Analysis through Loss and Asymmetric Approach in Audio Anti-Spoofing
di: Shim, Hye-jin, et al.
Pubblicazione: (2024)
di: Shim, Hye-jin, et al.
Pubblicazione: (2024)
Disentangling Speaker Traits for Deepfake Source Verification via Chebyshev Polynomial and Riemannian Metric Learning
di: Xuan, Xi, et al.
Pubblicazione: (2026)
di: Xuan, Xi, et al.
Pubblicazione: (2026)
Improving Multilingual ASR in the Wild Using Simple N-best Re-ranking
di: Yan, Brian, et al.
Pubblicazione: (2024)
di: Yan, Brian, et al.
Pubblicazione: (2024)
ASR-EC Benchmark: Evaluating Large Language Models on Chinese ASR Error Correction
di: Wei, Victor Junqiu, et al.
Pubblicazione: (2024)
di: Wei, Victor Junqiu, et al.
Pubblicazione: (2024)
Cyclostationarity Analysis as a Complement to Self-Supervised Representations for Speech Deepfake Detection
di: Hanilçi, Cemal, et al.
Pubblicazione: (2026)
di: Hanilçi, Cemal, et al.
Pubblicazione: (2026)
ASR Error Correction using Large Language Models
di: Ma, Rao, et al.
Pubblicazione: (2024)
di: Ma, Rao, et al.
Pubblicazione: (2024)
Crossmodal ASR Error Correction with Discrete Speech Units
di: Li, Yuanchao, et al.
Pubblicazione: (2024)
di: Li, Yuanchao, et al.
Pubblicazione: (2024)
Advocating Character Error Rate for Multilingual ASR Evaluation
di: K, Thennal D, et al.
Pubblicazione: (2024)
di: K, Thennal D, et al.
Pubblicazione: (2024)
Effects of Speaker Count, Duration, and Accent Diversity on Zero-Shot Accent Robustness in Low-Resource ASR
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2025)
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2025)
Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades
di: Jung, Donghyuk, et al.
Pubblicazione: (2026)
di: Jung, Donghyuk, et al.
Pubblicazione: (2026)
Evolutionary Prompt Design for LLM-Based Post-ASR Error Correction
di: Sachdev, Rithik, et al.
Pubblicazione: (2024)
di: Sachdev, Rithik, et al.
Pubblicazione: (2024)
Leveraging AM and FM Rhythm Spectrograms for Dementia Classification and Assessment
di: Gogoi, Parismita, et al.
Pubblicazione: (2025)
di: Gogoi, Parismita, et al.
Pubblicazione: (2025)
Minimising Biasing Word Errors for Contextual ASR with the Tree-Constrained Pointer Generator
di: Sun, Guangzhi, et al.
Pubblicazione: (2022)
di: Sun, Guangzhi, et al.
Pubblicazione: (2022)
Large Language Model Should Understand Pinyin for Chinese ASR Error Correction
di: Li, Yuang, et al.
Pubblicazione: (2024)
di: Li, Yuang, et al.
Pubblicazione: (2024)
Benchmarking Children's ASR with Supervised and Self-supervised Speech Foundation Models
di: Fan, Ruchao, et al.
Pubblicazione: (2024)
di: Fan, Ruchao, et al.
Pubblicazione: (2024)
The Multicultural Medical Assistant: Can LLMs Improve Medical ASR Errors Across Borders?
di: Adedeji, Ayo, et al.
Pubblicazione: (2025)
di: Adedeji, Ayo, et al.
Pubblicazione: (2025)
Failing Forward: Improving Generative Error Correction for ASR with Synthetic Data and Retrieval Augmentation
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
Diagnostic-Driven Layer-Wise Compensation for Post-Training Quantization of Encoder-Decoder ASR Models
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
PromptASR for contextualized ASR with controllable style
di: Yang, Xiaoyu, et al.
Pubblicazione: (2023)
di: Yang, Xiaoyu, et al.
Pubblicazione: (2023)
Benchmarking Japanese Speech Recognition on ASR-LLM Setups with Multi-Pass Augmented Generative Error Correction
di: Ko, Yuka, et al.
Pubblicazione: (2024)
di: Ko, Yuka, et al.
Pubblicazione: (2024)
Mind the Gap: Entity-Preserved Context-Aware ASR Structured Transcriptions
di: Altinok, Duygu
Pubblicazione: (2025)
di: Altinok, Duygu
Pubblicazione: (2025)
MSA-ASR: Efficient Multilingual Speaker Attribution with frozen ASR Models
di: Nguyen, Thai-Binh, et al.
Pubblicazione: (2024)
di: Nguyen, Thai-Binh, et al.
Pubblicazione: (2024)
AutoMode-ASR: Learning to Select ASR Systems for Better Quality and Cost
di: Gündüz, Ahmet, et al.
Pubblicazione: (2024)
di: Gündüz, Ahmet, et al.
Pubblicazione: (2024)
SSCFormer: Push the Limit of Chunk-wise Conformer for Streaming ASR Using Sequentially Sampled Chunks and Chunked Causal Convolution
di: Wang, Fangyuan, et al.
Pubblicazione: (2022)
di: Wang, Fangyuan, et al.
Pubblicazione: (2022)
Locality enhanced dynamic biasing and sampling strategies for contextual ASR
di: Jalal, Md Asif, et al.
Pubblicazione: (2024)
di: Jalal, Md Asif, et al.
Pubblicazione: (2024)
An Explainable Probabilistic Attribute Embedding Approach for Spoofed Speech Characterization
di: Chhibber, Manasi, et al.
Pubblicazione: (2024)
di: Chhibber, Manasi, et al.
Pubblicazione: (2024)
Romanization Encoding For Multilingual ASR
di: Ding, Wen, et al.
Pubblicazione: (2024)
di: Ding, Wen, et al.
Pubblicazione: (2024)
NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR
di: Xie, Yuan, et al.
Pubblicazione: (2026)
di: Xie, Yuan, et al.
Pubblicazione: (2026)
Revisiting ASR Error Correction with Specialized Models
di: Gu, Zijin, et al.
Pubblicazione: (2024)
di: Gu, Zijin, et al.
Pubblicazione: (2024)
IITKGP-ABSP Submission to LRE22: Language Recognition in Low-Resource Settings
di: Dey, Spandan, et al.
Pubblicazione: (2025)
di: Dey, Spandan, et al.
Pubblicazione: (2025)
Promptformer: Prompted Conformer Transducer for ASR
di: Duarte-Torres, Sergio, et al.
Pubblicazione: (2024)
di: Duarte-Torres, Sergio, et al.
Pubblicazione: (2024)
Qwen3-ASR Technical Report
di: Shi, Xian, et al.
Pubblicazione: (2026)
di: Shi, Xian, et al.
Pubblicazione: (2026)
Revisiting Acoustic Features for Robust ASR
di: Shah, Muhammad A., et al.
Pubblicazione: (2024)
di: Shah, Muhammad A., et al.
Pubblicazione: (2024)
Documenti analoghi
-
ChildAugment: Data Augmentation Methods for Zero-Resource Children's Speaker Verification
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2024) -
Causal Analysis of ASR Errors for Children: Quantifying the Impact of Physiological, Cognitive, and Extrinsic Factors
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2025) -
ROAR: Reinforcing Original to Augmented Data Ratio Dynamics for Wav2Vec2.0 Based ASR
di: Singh, Vishwanath Pratap, et al.
Pubblicazione: (2024) -
Continuous Learning for Children's ASR: Overcoming Catastrophic Forgetting with Elastic Weight Consolidation and Synaptic Intelligence
di: Ahadzi, Edem, et al.
Pubblicazione: (2025) -
Multilingual Source Tracing of Speech Deepfakes: A First Benchmark
di: Xuan, Xi, et al.
Pubblicazione: (2025)