Automatic Speech Recognition System-Independent Word Error Rate Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Park, Chanho, Chen, Mingjie, Hain, Thomas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fast Word Error Rate Estimation Using Self-Supervised Representations for Speech and Text
por: Park, Chanho, et al.
Publicado: (2023)
por: Park, Chanho, et al.
Publicado: (2023)
Improving Acoustic Word Embeddings through Correspondence Training of Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2024)
por: Meghanani, Amit, et al.
Publicado: (2024)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
por: Hu, Ke, et al.
Publicado: (2025)
por: Hu, Ke, et al.
Publicado: (2025)
Position-invariant Fine-tuning of Speech Enhancement Models with Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2026)
por: Meghanani, Amit, et al.
Publicado: (2026)
Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis
por: Do, Cong-Thanh, et al.
Publicado: (2024)
por: Do, Cong-Thanh, et al.
Publicado: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
por: Park, ChaeHun, et al.
Publicado: (2024)
por: Park, ChaeHun, et al.
Publicado: (2024)
LASER: Learning by Aligning Self-supervised Representations of Speech for Improving Content-related Tasks
por: Meghanani, Amit, et al.
Publicado: (2024)
por: Meghanani, Amit, et al.
Publicado: (2024)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
por: Guo, Jiaxin, et al.
Publicado: (2024)
por: Guo, Jiaxin, et al.
Publicado: (2024)
Hallucinations in Neural Automatic Speech Recognition: Identifying Errors and Hallucinatory Models
por: Frieske, Rita, et al.
Publicado: (2024)
por: Frieske, Rita, et al.
Publicado: (2024)
Speech Emotion Recognition with ASR Transcripts: A Comprehensive Study on Word Error Rate and Fusion Techniques
por: Li, Yuanchao, et al.
Publicado: (2024)
por: Li, Yuanchao, et al.
Publicado: (2024)
Error Correction by Paying Attention to Both Acoustic and Confidence References for Automatic Speech Recognition
por: Shu, Yuchun, et al.
Publicado: (2024)
por: Shu, Yuchun, et al.
Publicado: (2024)
Automatic Speech Recognition for Hindi
por: Saha, Anish, et al.
Publicado: (2024)
por: Saha, Anish, et al.
Publicado: (2024)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
por: Wang, Yujin, et al.
Publicado: (2022)
por: Wang, Yujin, et al.
Publicado: (2022)
Automatic Speech Recognition Biases in Newcastle English: an Error Analysis
por: Serditova, Dana, et al.
Publicado: (2025)
por: Serditova, Dana, et al.
Publicado: (2025)
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
por: Abu, Turi, et al.
Publicado: (2025)
por: Abu, Turi, et al.
Publicado: (2025)
Dynamic Data Pruning for Automatic Speech Recognition
por: Xiao, Qiao, et al.
Publicado: (2024)
por: Xiao, Qiao, et al.
Publicado: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary
por: Sudo, Yui, et al.
Publicado: (2024)
por: Sudo, Yui, et al.
Publicado: (2024)
SCORE: Self-supervised Correspondence Fine-tuning for Improved Content Representations
por: Meghanani, Amit, et al.
Publicado: (2024)
por: Meghanani, Amit, et al.
Publicado: (2024)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
por: Hu, Jiliang, et al.
Publicado: (2025)
por: Hu, Jiliang, et al.
Publicado: (2025)
EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
por: Ma, Ziyang, et al.
Publicado: (2024)
por: Ma, Ziyang, et al.
Publicado: (2024)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
por: Shi, Hao, et al.
Publicado: (2024)
por: Shi, Hao, et al.
Publicado: (2024)
Automatic Speech Recognition for Biomedical Data in Bengali Language
por: Kabir, Shariar, et al.
Publicado: (2024)
por: Kabir, Shariar, et al.
Publicado: (2024)
Benchmarking Automatic Speech Recognition Models for African Languages
por: Nahabwe, Alvin, et al.
Publicado: (2025)
por: Nahabwe, Alvin, et al.
Publicado: (2025)
Exploring Gender Disparities in Automatic Speech Recognition Technology
por: ElGhazaly, Hend, et al.
Publicado: (2025)
por: ElGhazaly, Hend, et al.
Publicado: (2025)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
por: Zhang, Hezhao, et al.
Publicado: (2026)
por: Zhang, Hezhao, et al.
Publicado: (2026)
PI-Whisper: Designing an Adaptive and Incremental Automatic Speech Recognition System for Edge Devices
por: Nassereldine, Amir, et al.
Publicado: (2024)
por: Nassereldine, Amir, et al.
Publicado: (2024)
A Comprehensive Study of the Current State-of-the-Art in Nepali Automatic Speech Recognition Systems
por: Ghimire, Rupak Raj, et al.
Publicado: (2024)
por: Ghimire, Rupak Raj, et al.
Publicado: (2024)
Exploring the Integration of Large Language Models into Automatic Speech Recognition Systems: An Empirical Study
por: Min, Zeping, et al.
Publicado: (2023)
por: Min, Zeping, et al.
Publicado: (2023)
Convolutional Variational Autoencoders for Spectrogram Compression in Automatic Speech Recognition
por: Iakovenko, Olga, et al.
Publicado: (2024)
por: Iakovenko, Olga, et al.
Publicado: (2024)
Transliterated Zero-Shot Domain Adaptation for Automatic Speech Recognition
por: Zhu, Han, et al.
Publicado: (2024)
por: Zhu, Han, et al.
Publicado: (2024)
Supporting SENCOTEN Language Documentation Efforts with Automatic Speech Recognition
por: Geng, Mengzhe, et al.
Publicado: (2025)
por: Geng, Mengzhe, et al.
Publicado: (2025)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
por: Sasu, David, et al.
Publicado: (2025)
por: Sasu, David, et al.
Publicado: (2025)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
por: Lin, Zhennan, et al.
Publicado: (2025)
por: Lin, Zhennan, et al.
Publicado: (2025)
Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems
por: Kulkarni, Ajinkya, et al.
Publicado: (2025)
por: Kulkarni, Ajinkya, et al.
Publicado: (2025)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
por: Adila, Aulia, et al.
Publicado: (2024)
por: Adila, Aulia, et al.
Publicado: (2024)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
por: Hsu, Ming-Hao, et al.
Publicado: (2024)
por: Hsu, Ming-Hao, et al.
Publicado: (2024)
SSHR: Leveraging Self-supervised Hierarchical Representations for Multilingual Automatic Speech Recognition
por: Xue, Hongfei, et al.
Publicado: (2023)
por: Xue, Hongfei, et al.
Publicado: (2023)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
por: Shakeel, Muhammad, et al.
Publicado: (2024)
por: Shakeel, Muhammad, et al.
Publicado: (2024)
Automatic Speech Recognition for Non-Native English: Accuracy and Disfluency Handling
por: McGuire, Michael
Publicado: (2025)
por: McGuire, Michael
Publicado: (2025)
A Deep Learning Automatic Speech Recognition Model for Shona Language
por: Sirora, Leslie Wellington, et al.
Publicado: (2025)
por: Sirora, Leslie Wellington, et al.
Publicado: (2025)
Ejemplares similares
-
Fast Word Error Rate Estimation Using Self-Supervised Representations for Speech and Text
por: Park, Chanho, et al.
Publicado: (2023) -
Improving Acoustic Word Embeddings through Correspondence Training of Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2024) -
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
por: Hu, Ke, et al.
Publicado: (2025) -
Position-invariant Fine-tuning of Speech Enhancement Models with Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2026) -
Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis
por: Do, Cong-Thanh, et al.
Publicado: (2024)