Guardado en:
| Autores principales: | Taguchi, Chihiro, Saransig, Jefferson, Velásquez, Dayana, Chiang, David |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2404.15501 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automatic Speech Recognition for Documenting Endangered Languages: Case Study of Ikema Miyakoan
por: Taguchi, Chihiro, et al.
Publicado: (2026)
por: Taguchi, Chihiro, et al.
Publicado: (2026)
Language Complexity and Speech Recognition Accuracy: Orthographic Complexity Hurts, Phonological Complexity Doesn't
por: Taguchi, Chihiro, et al.
Publicado: (2024)
por: Taguchi, Chihiro, et al.
Publicado: (2024)
Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark
por: Taguchi, Chihiro, et al.
Publicado: (2025)
por: Taguchi, Chihiro, et al.
Publicado: (2025)
Efficient Context Selection for Long-Context QA: No Tuning, No Iteration, Just Adaptive-$k$
por: Taguchi, Chihiro, et al.
Publicado: (2025)
por: Taguchi, Chihiro, et al.
Publicado: (2025)
Automatic Speech Recognition for Sanskrit with Transfer Learning
por: Sadhukhan, Bidit, et al.
Publicado: (2025)
por: Sadhukhan, Bidit, et al.
Publicado: (2025)
Automatic Speech Recognition for Greek Medical Dictation
por: Georgilas, Vardis, et al.
Publicado: (2025)
por: Georgilas, Vardis, et al.
Publicado: (2025)
Multi-Stage Multi-Modal Pre-Training for Automatic Speech Recognition
por: Jain, Yash, et al.
Publicado: (2024)
por: Jain, Yash, et al.
Publicado: (2024)
Augmenting Automatic Speech Recognition Models with Disfluency Detection
por: Amann, Robin, et al.
Publicado: (2024)
por: Amann, Robin, et al.
Publicado: (2024)
VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain
por: Le-Duc, Khai
Publicado: (2024)
por: Le-Duc, Khai
Publicado: (2024)
Error-preserving Automatic Speech Recognition of Young English Learners' Language
por: Michot, Janick, et al.
Publicado: (2024)
por: Michot, Janick, et al.
Publicado: (2024)
Fairness of Automatic Speech Recognition: Looking Through a Philosophical Lens
por: Choi, Anna Seo Gyeong, et al.
Publicado: (2025)
por: Choi, Anna Seo Gyeong, et al.
Publicado: (2025)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
por: Min, Do June, et al.
Publicado: (2024)
por: Min, Do June, et al.
Publicado: (2024)
Handling Numeric Expressions in Automatic Speech Recognition
por: Huber, Christian, et al.
Publicado: (2024)
por: Huber, Christian, et al.
Publicado: (2024)
A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain
por: Obaidah, Qusai Abo, et al.
Publicado: (2024)
por: Obaidah, Qusai Abo, et al.
Publicado: (2024)
Doing More with Less: Data Augmentation for Sudanese Dialect Automatic Speech Recognition
por: Mansour, Ayman
Publicado: (2026)
por: Mansour, Ayman
Publicado: (2026)
SimClass: A Classroom Speech Dataset Generated via Game Engine Simulation For Automatic Speech Recognition Research
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
SENS-ASR: Semantic Embedding injection in Neural-transducer for Streaming Automatic Speech Recognition
por: Dkhissi, Youness, et al.
Publicado: (2026)
por: Dkhissi, Youness, et al.
Publicado: (2026)
Benchmarking Rotary Position Embeddings for Automatic Speech Recognition
por: Zhang, Shucong, et al.
Publicado: (2025)
por: Zhang, Shucong, et al.
Publicado: (2025)
LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families
por: Chen, Jianan, et al.
Publicado: (2026)
por: Chen, Jianan, et al.
Publicado: (2026)
Semantically Corrected Amharic Automatic Speech Recognition
por: Adnew, Samuael, et al.
Publicado: (2024)
por: Adnew, Samuael, et al.
Publicado: (2024)
Improved Contextual Recognition In Automatic Speech Recognition Systems By Semantic Lattice Rescoring
por: Sudarshan, Ankitha, et al.
Publicado: (2023)
por: Sudarshan, Ankitha, et al.
Publicado: (2023)
ASKD-Whisper: Adaptive Self-knowledge Distillation for Efficient and Low-Latency Automatic Speech Recognition
por: Lee, Junseok, et al.
Publicado: (2026)
por: Lee, Junseok, et al.
Publicado: (2026)
Benchmarking Automatic Speech Recognition for Indian Languages in Agricultural Contexts
por: S, Chandrashekar M, et al.
Publicado: (2026)
por: S, Chandrashekar M, et al.
Publicado: (2026)
Towards Unsupervised Speech Recognition at the Syllable-Level
por: Wang, Liming, et al.
Publicado: (2025)
por: Wang, Liming, et al.
Publicado: (2025)
It's Never Too Late: Fusing Acoustic Information into Large Language Models for Automatic Speech Recognition
por: Chen, Chen, et al.
Publicado: (2024)
por: Chen, Chen, et al.
Publicado: (2024)
A Comparative Analysis of Bilingual and Trilingual Wav2Vec Models for Automatic Speech Recognition in Multilingual Oral History Archives
por: Lehečka, Jan, et al.
Publicado: (2024)
por: Lehečka, Jan, et al.
Publicado: (2024)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
por: Ghosh, Sreyan, et al.
Publicado: (2024)
por: Ghosh, Sreyan, et al.
Publicado: (2024)
Multistage Fine-tuning Strategies for Automatic Speech Recognition in Low-resource Languages
por: Pillai, Leena G, et al.
Publicado: (2024)
por: Pillai, Leena G, et al.
Publicado: (2024)
Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio
por: He, Xinlu, et al.
Publicado: (2025)
por: He, Xinlu, et al.
Publicado: (2025)
Arabic Little STT: Arabic Children Speech Recognition Dataset
por: Alkadri, Mouhand, et al.
Publicado: (2025)
por: Alkadri, Mouhand, et al.
Publicado: (2025)
Bias Vector: Mitigating Biases in Language Models with Task Arithmetic Approach
por: Shirafuji, Daiki, et al.
Publicado: (2024)
por: Shirafuji, Daiki, et al.
Publicado: (2024)
Improving Speech Recognition Error Prediction for Modern and Off-the-shelf Speech Recognizers
por: Serai, Prashant, et al.
Publicado: (2024)
por: Serai, Prashant, et al.
Publicado: (2024)
Breaking Through the Spike: Spike Window Decoding for Accelerated and Precise Automatic Speech Recognition
por: Zhang, Wei, et al.
Publicado: (2025)
por: Zhang, Wei, et al.
Publicado: (2025)
Gated Low-rank Adaptation for personalized Code-Switching Automatic Speech Recognition on the low-spec devices
por: Kim, Gwantae, et al.
Publicado: (2024)
por: Kim, Gwantae, et al.
Publicado: (2024)
How do Hyenas deal with Human Speech? Speech Recognition and Translation with ConfHyena
por: Gaido, Marco, et al.
Publicado: (2024)
por: Gaido, Marco, et al.
Publicado: (2024)
Semantic Differentiation in Speech Emotion Recognition: Insights from Descriptive and Expressive Speech Roles
por: Guo, Rongchen, et al.
Publicado: (2025)
por: Guo, Rongchen, et al.
Publicado: (2025)
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
por: Rufai, Amina Mardiyyah, et al.
Publicado: (2020)
por: Rufai, Amina Mardiyyah, et al.
Publicado: (2020)
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
por: Storey, Edward, et al.
Publicado: (2025)
por: Storey, Edward, et al.
Publicado: (2025)
Empirical Evaluation of Public HateSpeech Datasets
por: Jaf, Sadar, et al.
Publicado: (2024)
por: Jaf, Sadar, et al.
Publicado: (2024)
In-context Language Learning for Endangered Languages in Speech Recognition
por: Li, Zhaolin, et al.
Publicado: (2025)
por: Li, Zhaolin, et al.
Publicado: (2025)
Ejemplares similares
-
Automatic Speech Recognition for Documenting Endangered Languages: Case Study of Ikema Miyakoan
por: Taguchi, Chihiro, et al.
Publicado: (2026) -
Language Complexity and Speech Recognition Accuracy: Orthographic Complexity Hurts, Phonological Complexity Doesn't
por: Taguchi, Chihiro, et al.
Publicado: (2024) -
Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark
por: Taguchi, Chihiro, et al.
Publicado: (2025) -
Efficient Context Selection for Long-Context QA: No Tuning, No Iteration, Just Adaptive-$k$
por: Taguchi, Chihiro, et al.
Publicado: (2025) -
Automatic Speech Recognition for Sanskrit with Transfer Learning
por: Sadhukhan, Bidit, et al.
Publicado: (2025)