Semantically Corrected Amharic Automatic Speech Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Adnew, Samuael, Liang, Paul Pu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025)
von: Storey, Edward, et al.
Veröffentlicht: (2025)
Reading Miscue Detection in Primary School through Automatic Speech Recognition
von: Gao, Lingyun, et al.
Veröffentlicht: (2024)
von: Gao, Lingyun, et al.
Veröffentlicht: (2024)
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020)
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020)
Task Arithmetic can Mitigate Synthetic-to-Real Gap in Automatic Speech Recognition
von: Su, Hsuan, et al.
Veröffentlicht: (2024)
von: Su, Hsuan, et al.
Veröffentlicht: (2024)
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)
TICL+: A Case Study On Speech In-Context Learning for Children's Speech Recognition
von: Zheng, Haolong, et al.
Veröffentlicht: (2025)
von: Zheng, Haolong, et al.
Veröffentlicht: (2025)
Generative Speech Recognition Error Correction with Large Language Models and Task-Activating Prompting
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2024)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2024)
Improved Contextual Recognition In Automatic Speech Recognition Systems By Semantic Lattice Rescoring
von: Sudarshan, Ankitha, et al.
Veröffentlicht: (2023)
von: Sudarshan, Ankitha, et al.
Veröffentlicht: (2023)
OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models
von: Chen, William, et al.
Veröffentlicht: (2025)
von: Chen, William, et al.
Veröffentlicht: (2025)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
von: Min, Do June, et al.
Veröffentlicht: (2024)
von: Min, Do June, et al.
Veröffentlicht: (2024)
Handling Numeric Expressions in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024)
von: Huber, Christian, et al.
Veröffentlicht: (2024)
Benchmarking Rotary Position Embeddings for Automatic Speech Recognition
von: Zhang, Shucong, et al.
Veröffentlicht: (2025)
von: Zhang, Shucong, et al.
Veröffentlicht: (2025)
Combining X-Vectors and Bayesian Batch Active Learning: Two-Stage Active Learning Pipeline for Speech Recognition
von: Kundacina, Ognjen, et al.
Veröffentlicht: (2024)
von: Kundacina, Ognjen, et al.
Veröffentlicht: (2024)
SpeechPrompt: Prompting Speech Language Models for Speech Processing Tasks
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2024)
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2024)
RNN-Transducer-based Losses for Speech Recognition on Noisy Targets
von: Bataev, Vladimir
Veröffentlicht: (2025)
von: Bataev, Vladimir
Veröffentlicht: (2025)
Attention-Based Recurrent Neural Network For Automatic Behavior Laying Hen Recognition
von: Laleye, Fréjus A. A., et al.
Veröffentlicht: (2024)
von: Laleye, Fréjus A. A., et al.
Veröffentlicht: (2024)
TICL: Text-Embedding KNN For Speech In-Context Learning Unlocks Speech Recognition Abilities of Large Multimodal Models
von: Zheng, Haolong, et al.
Veröffentlicht: (2025)
von: Zheng, Haolong, et al.
Veröffentlicht: (2025)
Large Language Models are Efficient Learners of Noise-Robust Speech Recognition
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021)
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
Automatic Pronunciation Error Detection and Correction of the Holy Quran's Learners Using Deep Learning
von: Abdelfattah, Abdullah, et al.
Veröffentlicht: (2025)
von: Abdelfattah, Abdullah, et al.
Veröffentlicht: (2025)
Omni-Router: Sharing Routing Decisions in Sparse Mixture-of-Experts for Speech Recognition
von: Gu, Zijin, et al.
Veröffentlicht: (2025)
von: Gu, Zijin, et al.
Veröffentlicht: (2025)
TokenChain: A Discrete Speech Chain via Semantic Token Modeling
von: Wang, Mingxuan, et al.
Veröffentlicht: (2025)
von: Wang, Mingxuan, et al.
Veröffentlicht: (2025)
Re-evaluating Minimum Bayes Risk Decoding for Automatic Speech Recognition
von: Jinnai, Yuu
Veröffentlicht: (2025)
von: Jinnai, Yuu
Veröffentlicht: (2025)
VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain
von: Le-Duc, Khai
Veröffentlicht: (2024)
von: Le-Duc, Khai
Veröffentlicht: (2024)
Latent Speech-Text Transformer
von: Lu, Yen-Ju, et al.
Veröffentlicht: (2025)
von: Lu, Yen-Ju, et al.
Veröffentlicht: (2025)
Benchmarking Automatic Speech Recognition for Indian Languages in Agricultural Contexts
von: S, Chandrashekar M, et al.
Veröffentlicht: (2026)
von: S, Chandrashekar M, et al.
Veröffentlicht: (2026)
Gated Low-rank Adaptation for personalized Code-Switching Automatic Speech Recognition on the low-spec devices
von: Kim, Gwantae, et al.
Veröffentlicht: (2024)
von: Kim, Gwantae, et al.
Veröffentlicht: (2024)
Instruction Data Generation and Unsupervised Adaptation for Speech Language Models
von: Noroozi, Vahid, et al.
Veröffentlicht: (2024)
von: Noroozi, Vahid, et al.
Veröffentlicht: (2024)
DeepEmoNet: Building Machine Learning Models for Automatic Emotion Recognition in Human Speeches
von: Vu, Tai
Veröffentlicht: (2025)
von: Vu, Tai
Veröffentlicht: (2025)
Transducers with Pronunciation-aware Embeddings for Automatic Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2024)
von: Xu, Hainan, et al.
Veröffentlicht: (2024)
Regularizing Learnable Feature Extraction for Automatic Speech Recognition
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
Leveraging Allophony in Self-Supervised Speech Models for Atypical Pronunciation Assessment
von: Choi, Kwanghee, et al.
Veröffentlicht: (2025)
von: Choi, Kwanghee, et al.
Veröffentlicht: (2025)
Analyzing and Fine-Tuning Whisper Models for Multilingual Pilot Speech Transcription in the Cockpit
von: Nareddy, Kartheek Kumar Reddy, et al.
Veröffentlicht: (2025)
von: Nareddy, Kartheek Kumar Reddy, et al.
Veröffentlicht: (2025)
Mini-Omni-Reasoner: Token-Level Thinking-in-Speaking in Large Speech Models
von: Xie, Zhifei, et al.
Veröffentlicht: (2025)
von: Xie, Zhifei, et al.
Veröffentlicht: (2025)
AG-LSEC: Audio Grounded Lexical Speaker Error Correction
von: Paturi, Rohit, et al.
Veröffentlicht: (2024)
von: Paturi, Rohit, et al.
Veröffentlicht: (2024)
FlashSpeech: Efficient Zero-Shot Speech Synthesis
von: Ye, Zhen, et al.
Veröffentlicht: (2024)
von: Ye, Zhen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025) -
Reading Miscue Detection in Primary School through Automatic Speech Recognition
von: Gao, Lingyun, et al.
Veröffentlicht: (2024) -
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020) -
Task Arithmetic can Mitigate Synthetic-to-Real Gap in Automatic Speech Recognition
von: Su, Hsuan, et al.
Veröffentlicht: (2024) -
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)