Continuously Learning New Words in Automatic Speech Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huber, Christian, Waibel, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2025)
von: Huber, Christian, et al.
Veröffentlicht: (2025)
Handling Numeric Expressions in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024)
von: Huber, Christian, et al.
Veröffentlicht: (2024)
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021)
Adapting Language Balance in Code-Switching Speech
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
Huntington Disease Automatic Speech Recognition with Biomarker Supervision
von: Wang, Charles L., et al.
Veröffentlicht: (2026)
von: Wang, Charles L., et al.
Veröffentlicht: (2026)
Weight Factorization and Centralization for Continual Learning in Speech Recognition
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
Joint Unsupervised and Supervised Training for Automatic Speech Recognition via Bilevel Optimization
von: Saif, A F M, et al.
Veröffentlicht: (2024)
von: Saif, A F M, et al.
Veröffentlicht: (2024)
Supplementary Resources and Analysis for Automatic Speech Recognition Systems Trained on the Loquacious Dataset
von: Rossenbach, Nick, et al.
Veröffentlicht: (2025)
von: Rossenbach, Nick, et al.
Veröffentlicht: (2025)
ArEEG_Words: Dataset for Envisioned Speech Recognition using EEG for Arabic Words
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
von: Mou, Guanyi, et al.
Veröffentlicht: (2024)
Transducers with Pronunciation-aware Embeddings for Automatic Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2024)
von: Xu, Hainan, et al.
Veröffentlicht: (2024)
Regularizing Learnable Feature Extraction for Automatic Speech Recognition
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
Semantically Corrected Amharic Automatic Speech Recognition
von: Adnew, Samuael, et al.
Veröffentlicht: (2024)
von: Adnew, Samuael, et al.
Veröffentlicht: (2024)
Re-evaluating Minimum Bayes Risk Decoding for Automatic Speech Recognition
von: Jinnai, Yuu
Veröffentlicht: (2025)
von: Jinnai, Yuu
Veröffentlicht: (2025)
The Impact of Automatic Speech Transcription on Speaker Attribution
von: Aggazzotti, Cristina, et al.
Veröffentlicht: (2025)
von: Aggazzotti, Cristina, et al.
Veröffentlicht: (2025)
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
von: Rossenbach, Nick, et al.
Veröffentlicht: (2024)
von: Rossenbach, Nick, et al.
Veröffentlicht: (2024)
Robust Long-Form Bangla Speech Processing: Automatic Speech Recognition and Speaker Diarization
von: Chowdhury, MD. Sagor, et al.
Veröffentlicht: (2026)
von: Chowdhury, MD. Sagor, et al.
Veröffentlicht: (2026)
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025)
von: Storey, Edward, et al.
Veröffentlicht: (2025)
Policies and Evaluation for Online Meeting Summarization
von: Schneider, Felix, et al.
Veröffentlicht: (2025)
von: Schneider, Felix, et al.
Veröffentlicht: (2025)
Task Oriented Dialogue as a Catalyst for Self-Supervised Automatic Speech Recognition
von: Chan, David M., et al.
Veröffentlicht: (2024)
von: Chan, David M., et al.
Veröffentlicht: (2024)
On the Effect of Purely Synthetic Training Data for Different Automatic Speech Recognition Architectures
von: Hilmes, Benedikt, et al.
Veröffentlicht: (2024)
von: Hilmes, Benedikt, et al.
Veröffentlicht: (2024)
Clinical BERTScore: An Improved Measure of Automatic Speech Recognition Performance in Clinical Settings
von: Shor, Joel, et al.
Veröffentlicht: (2023)
von: Shor, Joel, et al.
Veröffentlicht: (2023)
Cocktail-Party Audio-Visual Speech Recognition
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
Empirical Study of Named Entity Recognition Performance Using Distribution-aware Word Embedding
von: Chen, Xin, et al.
Veröffentlicht: (2021)
von: Chen, Xin, et al.
Veröffentlicht: (2021)
PIER: A Novel Metric for Evaluating What Matters in Code-Switching
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
Bayesian Low-Rank Factorization for Robust Model Adaptation
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
A Survey on Automatic Online Hate Speech Detection in Low-Resource Languages
von: Das, Susmita, et al.
Veröffentlicht: (2024)
von: Das, Susmita, et al.
Veröffentlicht: (2024)
A Word is Worth 4-bit: Efficient Log Parsing with Binary Coded Decimal Recognition
von: Srivastava, Prerak, et al.
Veröffentlicht: (2025)
von: Srivastava, Prerak, et al.
Veröffentlicht: (2025)
MSNER: A Multilingual Speech Dataset for Named Entity Recognition
von: Meeus, Quentin, et al.
Veröffentlicht: (2024)
von: Meeus, Quentin, et al.
Veröffentlicht: (2024)
Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition
von: Zhu, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhu, Wenjing, et al.
Veröffentlicht: (2024)
WhisperNER: Unified Open Named Entity and Speech Recognition
von: Ayache, Gil, et al.
Veröffentlicht: (2024)
von: Ayache, Gil, et al.
Veröffentlicht: (2024)
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020)
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020)
Verifying the Robustness of Automatic Credibility Assessment
von: Przybyła, Piotr, et al.
Veröffentlicht: (2023)
von: Przybyła, Piotr, et al.
Veröffentlicht: (2023)
Speech Recognition With LLMs Adapted to Disordered Speech Using Reinforcement Learning
von: Nagpal, Chirag, et al.
Veröffentlicht: (2024)
von: Nagpal, Chirag, et al.
Veröffentlicht: (2024)
Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 2025
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
Mai Ho'omāuna i ka 'Ai: Language Models Improve Automatic Speech Recognition in Hawaiian
von: Chaparala, Kaavya, et al.
Veröffentlicht: (2024)
von: Chaparala, Kaavya, et al.
Veröffentlicht: (2024)
Paragraph Segmentation Revisited: Towards a Standard Task for Structuring Speech
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
Task Arithmetic can Mitigate Synthetic-to-Real Gap in Automatic Speech Recognition
von: Su, Hsuan, et al.
Veröffentlicht: (2024)
von: Su, Hsuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2025) -
Handling Numeric Expressions in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024) -
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021) -
Adapting Language Balance in Code-Switching Speech
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025) -
Huntington Disease Automatic Speech Recognition with Biomarker Supervision
von: Wang, Charles L., et al.
Veröffentlicht: (2026)