Quality of Automatic Speech Recognition -- Polish Language case study -- from Wav2Vec to Scribe ElevenLabs
Fuente:
arXiv
Saved in:
| Main Authors: | Pietroń, Marcin, Piórkowski, Szymon, Faber, Kamil, Żurek, Dominik, Karwatowski, Michał, Duda, Jerzy, Zieliński, Hubert, Lipnicki, Piotr, Leszczuk, Mikołaj |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AD-NEv++ : The multi-architecture neuroevolution-based multivariate anomaly detection framework
by: Pietroń, Marcin, et al.
Published: (2024)
by: Pietroń, Marcin, et al.
Published: (2024)
Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning
by: Gajewski, Paweł, et al.
Published: (2024)
by: Gajewski, Paweł, et al.
Published: (2024)
xLSTMAD: A Powerful xLSTM-based Method for Anomaly Detection
by: Faber, Kamil, et al.
Published: (2025)
by: Faber, Kamil, et al.
Published: (2025)
TinySubNets: An efficient and low capacity continual learning strategy
by: Pietroń, Marcin, et al.
Published: (2024)
by: Pietroń, Marcin, et al.
Published: (2024)
Towards efficient deep autoencoders for multivariate time series anomaly detection
by: Pietroń, Marcin, et al.
Published: (2024)
by: Pietroń, Marcin, et al.
Published: (2024)
Chatterbox TTS: Open Source vs. ElevenLabs
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
TSN-Affinity: Similarity-Driven Parameter Reuse for Continual Offline Reinforcement Learning
by: Żurek, Dominik, et al.
Published: (2026)
by: Żurek, Dominik, et al.
Published: (2026)
Evolutionary fine tuning of quantized convolution-based deep learning models
by: Pietroń, Marcin
Published: (2026)
by: Pietroń, Marcin
Published: (2026)
Digitale Privatisierung
by: Pietron, Dominik
Published: (2026)
by: Pietron, Dominik
Published: (2026)
Teaching Wav2Vec2 the Language of the Brain
by: Fiedler, Tobias, et al.
Published: (2025)
by: Fiedler, Tobias, et al.
Published: (2025)
Rethinking the Harmonic Loss via Non-Euclidean Distance Layers
by: Miller-Golub, Maxwell, et al.
Published: (2026)
by: Miller-Golub, Maxwell, et al.
Published: (2026)
SpecWav-Attack: Leveraging Spectrogram Resizing and Wav2Vec 2.0 for Attacking Anonymized Speech
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Chapter Quality Assessment in Video Surveillance
by: Leszczuk, Mikoaj, et al.
Published: (2021)
by: Leszczuk, Mikoaj, et al.
Published: (2021)
Wav2Small: Distilling Wav2Vec2 to 72K parameters for Low-Resource Speech emotion recognition
by: Kounadis-Bastian, Dionyssos, et al.
Published: (2024)
by: Kounadis-Bastian, Dionyssos, et al.
Published: (2024)
A Comparative Analysis of Bilingual and Trilingual Wav2Vec Models for Automatic Speech Recognition in Multilingual Oral History Archives
by: Lehečka, Jan, et al.
Published: (2024)
by: Lehečka, Jan, et al.
Published: (2024)
Efficient argument classification with compact language models and ChatGPT-4 refinements
by: Pietron, Marcin, et al.
Published: (2024)
by: Pietron, Marcin, et al.
Published: (2024)
Pulling Back the Curtain on Deep Networks
by: Satkiewicz, Maciej, et al.
Published: (2025)
by: Satkiewicz, Maciej, et al.
Published: (2025)
Over-the-air White-box Attack on the Wav2Vec Speech Recognition Neural Network
by: Alexey, Protopopov
Published: (2026)
by: Alexey, Protopopov
Published: (2026)
The melodic forms of Polish Folk Ballads
by: Zbigniew Jerzy Przerembski
Published: (2000)
by: Zbigniew Jerzy Przerembski
Published: (2000)
Building a Bridge between the Two Schools: Realizing a Practical Path to Include Literacy-based Skills within the STEM Curricula
by: Gómez, Jorge Torres, et al.
Published: (2026)
by: Gómez, Jorge Torres, et al.
Published: (2026)
A Closer Look at Wav2Vec2 Embeddings for On-Device Single-Channel Speech Enhancement
by: Shankar, Ravi, et al.
Published: (2024)
by: Shankar, Ravi, et al.
Published: (2024)
Exploring ASR-Based Wav2Vec2 for Automated Speech Disorder Assessment: Insights and Analysis
by: Nguyen, Tuan, et al.
Published: (2024)
by: Nguyen, Tuan, et al.
Published: (2024)
The Memory of Communist Poland in the Third Polish Republic. A Tentative Systematisation
by: Jerzy Łazor
Published: (2016)
by: Jerzy Łazor
Published: (2016)
Evaluating the Representation of Vowels in Wav2Vec Feature Extractor: A Layer-Wise Analysis Using MFCCs
by: De Cristofaro, Domenico, et al.
Published: (2025)
by: De Cristofaro, Domenico, et al.
Published: (2025)
Exploring Pathological Speech Quality Assessment with ASR-Powered Wav2Vec2 in Data-Scarce Context
by: Nguyen, Tuan, et al.
Published: (2024)
by: Nguyen, Tuan, et al.
Published: (2024)
ROAR: Reinforcing Original to Augmented Data Ratio Dynamics for Wav2Vec2.0 Based ASR
by: Singh, Vishwanath Pratap, et al.
Published: (2024)
by: Singh, Vishwanath Pratap, et al.
Published: (2024)
Whisper Turns Stronger: Augmenting Wav2Vec 2.0 for Superior ASR in Low-Resource Languages
by: Anidjar, Or Haim, et al.
Published: (2024)
by: Anidjar, Or Haim, et al.
Published: (2024)
A comparison of data filtering techniques for English-Polish LLM-based machine translation in the biomedical domain
by: Lérida, Jorge del Pozo, et al.
Published: (2025)
by: Lérida, Jorge del Pozo, et al.
Published: (2025)
Efficient Language Adaptive Pre-training: Extending State-of-the-Art Large Language Models for Polish
by: Ruciński, Szymon
Published: (2024)
by: Ruciński, Szymon
Published: (2024)
Adaptability of ASR Models on Low-Resource Language: A Comparative Study of Whisper and Wav2Vec-BERT on Bangla
by: Ridoy, Md Sazzadul Islam, et al.
Published: (2025)
by: Ridoy, Md Sazzadul Islam, et al.
Published: (2025)
Evaluating the Effectiveness of Transformer Layers in Wav2Vec 2.0, XLS-R, and Whisper for Speaker Identification Tasks
by: Stuhlmann, Linus, et al.
Published: (2025)
by: Stuhlmann, Linus, et al.
Published: (2025)
Which one Performs Better? Wav2Vec or Whisper? Applying both in Badini Kurdish Speech to Text (BKSTT)
by: Adnan, Renas, et al.
Published: (2025)
by: Adnan, Renas, et al.
Published: (2025)
Transcription and translation of videos using fine-tuned XLSR Wav2Vec2 on custom dataset and mBART
by: Tathe, Aniket, et al.
Published: (2024)
by: Tathe, Aniket, et al.
Published: (2024)
Arctic curves of periodic dimer models and generalized discriminants
by: Piorkowski, Mateusz
Published: (2024)
by: Piorkowski, Mateusz
Published: (2024)
Riemann-Hilbert Theory without local Parametrix Problems: Applications to Orthogonal Polynomials
by: Piorkowski, Mateusz
Published: (2019)
by: Piorkowski, Mateusz
Published: (2019)
Scoped MSO, Register Automata, and Expressions: Equivalence over Data Words
by: Piórkowski, Radosław
Published: (2026)
by: Piórkowski, Radosław
Published: (2026)
Hybrid spectral-spatial domain registration for nanometric tracking in digital in-line holographic microscopy
by: Kalinowski, Kamil, et al.
Published: (2026)
by: Kalinowski, Kamil, et al.
Published: (2026)
Speaker Emotion Recognition: Leveraging Self-Supervised Models for Feature Extraction Using Wav2Vec2 and HuBERT
by: Jafarzadeh, Pourya, et al.
Published: (2024)
by: Jafarzadeh, Pourya, et al.
Published: (2024)
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
On arrangements of smooth plane quartics and their bitangents
by: Janasz, Marek, et al.
Published: (2023)
by: Janasz, Marek, et al.
Published: (2023)
Similar Items
-
AD-NEv++ : The multi-architecture neuroevolution-based multivariate anomaly detection framework
by: Pietroń, Marcin, et al.
Published: (2024) -
Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning
by: Gajewski, Paweł, et al.
Published: (2024) -
xLSTMAD: A Powerful xLSTM-based Method for Anomaly Detection
by: Faber, Kamil, et al.
Published: (2025) -
TinySubNets: An efficient and low capacity continual learning strategy
by: Pietroń, Marcin, et al.
Published: (2024) -
Towards efficient deep autoencoders for multivariate time series anomaly detection
by: Pietroń, Marcin, et al.
Published: (2024)