Algorithms For Automatic Accentuation And Transcription Of Russian Texts In Speech Recognition Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Iakovenko, Olga, Bondarenko, Ivan, Borovikova, Mariya, Vodolazsky, Daniil |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Convolutional Variational Autoencoders for Spectrogram Compression in Automatic Speech Recognition
por: Iakovenko, Olga, et al.
Publicado: (2024)
por: Iakovenko, Olga, et al.
Publicado: (2024)
Methods of Automatic Matrix Language Determination for Code-Switched Speech
por: Iakovenko, Olga, et al.
Publicado: (2024)
por: Iakovenko, Olga, et al.
Publicado: (2024)
Pisets: A Robust Speech Recognition System for Lectures and Interviews
por: Bondarenko, Ivan, et al.
Publicado: (2026)
por: Bondarenko, Ivan, et al.
Publicado: (2026)
Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
por: Nédellec, Claire, et al.
Publicado: (2024)
por: Nédellec, Claire, et al.
Publicado: (2024)
Lost in Transcription: Identifying and Quantifying the Accuracy Biases of Automatic Speech Recognition Systems Against Disfluent Speech
por: Mujtaba, Dena, et al.
Publicado: (2024)
por: Mujtaba, Dena, et al.
Publicado: (2024)
A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems
por: Borgholt, Lasse, et al.
Publicado: (2025)
por: Borgholt, Lasse, et al.
Publicado: (2025)
Automatic Speech Recognition Advancements for Indigenous Languages of the Americas
por: Romero, Monica, et al.
Publicado: (2024)
por: Romero, Monica, et al.
Publicado: (2024)
The Impact of Automatic Speech Transcription on Speaker Attribution
por: Aggazzotti, Cristina, et al.
Publicado: (2025)
por: Aggazzotti, Cristina, et al.
Publicado: (2025)
Team MTS @ AutoMin 2021: An Overview of Existing Summarization Approaches and Comparison to Unsupervised Summarization Techniques
por: Iakovenko, Olga, et al.
Publicado: (2024)
por: Iakovenko, Olga, et al.
Publicado: (2024)
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
por: Lajčinová, Bibiána, et al.
Publicado: (2024)
por: Lajčinová, Bibiána, et al.
Publicado: (2024)
Automatic Speech Recognition for the Ika Language
por: Nzenwata, Uchenna, et al.
Publicado: (2024)
por: Nzenwata, Uchenna, et al.
Publicado: (2024)
On the Semantic and Syntactic Information Encoded in Proto-Tokens for One-Step Text Reconstruction
por: Bondarenko, Ivan, et al.
Publicado: (2026)
por: Bondarenko, Ivan, et al.
Publicado: (2026)
Streaming Translation and Transcription Through Speech-to-Text Causal Alignment
por: Koshkin, Roman, et al.
Publicado: (2026)
por: Koshkin, Roman, et al.
Publicado: (2026)
Responsible Benchmarking of Fairness for Automatic Speech Recognition
por: Herron, Felix, et al.
Publicado: (2026)
por: Herron, Felix, et al.
Publicado: (2026)
Vietnamese Automatic Speech Recognition: A Revisit
por: Vu, Thi, et al.
Publicado: (2026)
por: Vu, Thi, et al.
Publicado: (2026)
Automatic Speech Recognition for Hindi
por: Saha, Anish, et al.
Publicado: (2024)
por: Saha, Anish, et al.
Publicado: (2024)
Quantifying the Role of Textual Predictability in Automatic Speech Recognition
por: Robertson, Sean, et al.
Publicado: (2024)
por: Robertson, Sean, et al.
Publicado: (2024)
WST: Weakly Supervised Transducer for Automatic Speech Recognition
por: Gao, Dongji, et al.
Publicado: (2025)
por: Gao, Dongji, et al.
Publicado: (2025)
Stuttering-Aware Automatic Speech Recognition for Indonesian Language
por: Muhammad, Fadhil, et al.
Publicado: (2026)
por: Muhammad, Fadhil, et al.
Publicado: (2026)
Where Are We At with Automatic Speech Recognition for the Bambara Language?
por: Diallo, Seydou, et al.
Publicado: (2026)
por: Diallo, Seydou, et al.
Publicado: (2026)
Syllabic-Structure Decoder for Automatic Speech Recognition in Vietnamese
por: Nguyen, Nghia Hieu, et al.
Publicado: (2026)
por: Nguyen, Nghia Hieu, et al.
Publicado: (2026)
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
por: Rossenbach, Nick, et al.
Publicado: (2024)
por: Rossenbach, Nick, et al.
Publicado: (2024)
Algorithm for Automatic Legislative Text Consolidation
por: Etcheverry, Matias, et al.
Publicado: (2025)
por: Etcheverry, Matias, et al.
Publicado: (2025)
Speech-Aware Long Context Pruning and Integration for Contextualized Automatic Speech Recognition
por: Rong, Yiming, et al.
Publicado: (2025)
por: Rong, Yiming, et al.
Publicado: (2025)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
por: Nguyen, Khoa Anh, et al.
Publicado: (2026)
por: Nguyen, Khoa Anh, et al.
Publicado: (2026)
EmoAra: Emotion-Preserving English Speech Transcription and Cross-Lingual Translation with Arabic Text-to-Speech
por: Hassan, Besher, et al.
Publicado: (2026)
por: Hassan, Besher, et al.
Publicado: (2026)
Facilitating large language model Russian adaptation with Learned Embedding Propagation
por: Tikhomirov, Mikhail, et al.
Publicado: (2024)
por: Tikhomirov, Mikhail, et al.
Publicado: (2024)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
por: Luu, Nam, et al.
Publicado: (2025)
por: Luu, Nam, et al.
Publicado: (2025)
Speech Recognition Rescoring with Large Speech-Text Foundation Models
por: Shivakumar, Prashanth Gurunath, et al.
Publicado: (2024)
por: Shivakumar, Prashanth Gurunath, et al.
Publicado: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
por: Park, ChaeHun, et al.
Publicado: (2024)
por: Park, ChaeHun, et al.
Publicado: (2024)
Automatic Speech Recognition for Sanskrit with Transfer Learning
por: Sadhukhan, Bidit, et al.
Publicado: (2025)
por: Sadhukhan, Bidit, et al.
Publicado: (2025)
Automatic Speech Recognition for Greek Medical Dictation
por: Georgilas, Vardis, et al.
Publicado: (2025)
por: Georgilas, Vardis, et al.
Publicado: (2025)
Brain-tuning Improves Generalizability and Efficiency of Brain Alignment in Speech Models
por: Moussa, Omer, et al.
Publicado: (2025)
por: Moussa, Omer, et al.
Publicado: (2025)
Bilevel Joint Unsupervised and Supervised Training for Automatic Speech Recognition
por: Cui, Xiaodong, et al.
Publicado: (2024)
por: Cui, Xiaodong, et al.
Publicado: (2024)
Efficient infusion of self-supervised representations in Automatic Speech Recognition
por: Prabhu, Darshan, et al.
Publicado: (2024)
por: Prabhu, Darshan, et al.
Publicado: (2024)
Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
Supplementary Resources and Analysis for Automatic Speech Recognition Systems Trained on the Loquacious Dataset
por: Rossenbach, Nick, et al.
Publicado: (2025)
por: Rossenbach, Nick, et al.
Publicado: (2025)
A Pilot Study of GSLM-based Simulation of Foreign Accentuation Only Using Native Speech Corpora
por: Onda, Kentaro, et al.
Publicado: (2024)
por: Onda, Kentaro, et al.
Publicado: (2024)
Fotheidil: an Automatic Transcription System for the Irish Language
por: Lonergan, Liam, et al.
Publicado: (2024)
por: Lonergan, Liam, et al.
Publicado: (2024)
Mark My Words: A Robust Multilingual Model for Punctuation in Text and Speech Transcripts
por: Pulipaka, Sidharth, et al.
Publicado: (2025)
por: Pulipaka, Sidharth, et al.
Publicado: (2025)
Ejemplares similares
-
Convolutional Variational Autoencoders for Spectrogram Compression in Automatic Speech Recognition
por: Iakovenko, Olga, et al.
Publicado: (2024) -
Methods of Automatic Matrix Language Determination for Code-Switched Speech
por: Iakovenko, Olga, et al.
Publicado: (2024) -
Pisets: A Robust Speech Recognition System for Lectures and Interviews
por: Bondarenko, Ivan, et al.
Publicado: (2026) -
Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
por: Nédellec, Claire, et al.
Publicado: (2024) -
Lost in Transcription: Identifying and Quantifying the Accuracy Biases of Automatic Speech Recognition Systems Against Disfluent Speech
por: Mujtaba, Dena, et al.
Publicado: (2024)