LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximation
Fuente:
arXiv
Salvato in:
| Autori principali: | Kamahori, Keisuke, Kasai, Jungo, Kojima, Noriyuki, Kasikci, Baris |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VoxServe: Streaming-Centric Serving System for Speech Language Models
di: Kamahori, Keisuke, et al.
Pubblicazione: (2026)
di: Kamahori, Keisuke, et al.
Pubblicazione: (2026)
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
di: Wang, He, et al.
Pubblicazione: (2024)
di: Wang, He, et al.
Pubblicazione: (2024)
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models
di: Feng, Chen, et al.
Pubblicazione: (2025)
di: Feng, Chen, et al.
Pubblicazione: (2025)
Speech Recognition on TV Series with Video-guided Post-ASR Correction
di: Yang, Haoyuan, et al.
Pubblicazione: (2025)
di: Yang, Haoyuan, et al.
Pubblicazione: (2025)
Self-supervised ASR Models and Features For Dysarthric and Elderly Speech Recognition
di: Hu, Shujie, et al.
Pubblicazione: (2024)
di: Hu, Shujie, et al.
Pubblicazione: (2024)
TG-ASR: Translation-Guided Learning with Parallel Gated Cross Attention for Low-Resource Automatic Speech Recognition
di: Yang, Cheng-Yeh, et al.
Pubblicazione: (2026)
di: Yang, Cheng-Yeh, et al.
Pubblicazione: (2026)
Bridging ASR and LLMs for Dysarthric Speech Recognition: Benchmarking Self-Supervised and Generative Approaches
di: Aboeitta, Ahmed, et al.
Pubblicazione: (2025)
di: Aboeitta, Ahmed, et al.
Pubblicazione: (2025)
Enhancing Automatic Speech Recognition Through Integrated Noise Detection Architecture
di: Singh, Karamvir
Pubblicazione: (2025)
di: Singh, Karamvir
Pubblicazione: (2025)
Automatic Speech Recognition in the Modern Era: Architectures, Training, and Evaluation
di: Nayeem, Md., et al.
Pubblicazione: (2025)
di: Nayeem, Md., et al.
Pubblicazione: (2025)
Speech Recognition-based Feature Extraction for Enhanced Automatic Severity Classification in Dysarthric Speech
di: Choi, Yerin, et al.
Pubblicazione: (2024)
di: Choi, Yerin, et al.
Pubblicazione: (2024)
Adaptation and Optimization of Automatic Speech Recognition (ASR) for the Maritime Domain in the Field of VHF Communication
di: Nakilcioglu, Emin Cagatay, et al.
Pubblicazione: (2023)
di: Nakilcioglu, Emin Cagatay, et al.
Pubblicazione: (2023)
Serialized Speech Information Guidance with Overlapped Encoding Separation for Multi-Speaker Automatic Speech Recognition
di: Shi, Hao, et al.
Pubblicazione: (2024)
di: Shi, Hao, et al.
Pubblicazione: (2024)
Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech
di: Hajal, Karl El, et al.
Pubblicazione: (2025)
di: Hajal, Karl El, et al.
Pubblicazione: (2025)
Unsupervised Rhythm and Voice Conversion of Dysarthric to Healthy Speech for ASR
di: Hajal, Karl El, et al.
Pubblicazione: (2025)
di: Hajal, Karl El, et al.
Pubblicazione: (2025)
LoRP-TTS: Low-Rank Personalized Text-To-Speech
di: Bondaruk, Łukasz, et al.
Pubblicazione: (2025)
di: Bondaruk, Łukasz, et al.
Pubblicazione: (2025)
Do we really need Self-Attention for Streaming Automatic Speech Recognition?
di: Dkhissi, Youness, et al.
Pubblicazione: (2026)
di: Dkhissi, Youness, et al.
Pubblicazione: (2026)
Tiny-Align: Bridging Automatic Speech Recognition and Large Language Model on the Edge
di: Qin, Ruiyang, et al.
Pubblicazione: (2024)
di: Qin, Ruiyang, et al.
Pubblicazione: (2024)
ACES: Accent Subspaces for Coupling, Explanations, and Stress-Testing in Automatic Speech Recognition
di: Parekh, Swapnil
Pubblicazione: (2026)
di: Parekh, Swapnil
Pubblicazione: (2026)
LLMs-Integrated Automatic Hate Speech Recognition Using Controllable Text Generation Models
di: Oshima, Ryutaro, et al.
Pubblicazione: (2026)
di: Oshima, Ryutaro, et al.
Pubblicazione: (2026)
Probing the Information Encoded in Neural-based Acoustic Models of Automatic Speech Recognition Systems
di: Raymondaud, Quentin, et al.
Pubblicazione: (2024)
di: Raymondaud, Quentin, et al.
Pubblicazione: (2024)
Investigation of Whisper ASR Hallucinations Induced by Non-Speech Audio
di: Barański, Mateusz, et al.
Pubblicazione: (2025)
di: Barański, Mateusz, et al.
Pubblicazione: (2025)
HuBERT-VIC: Improving Noise-Robust Automatic Speech Recognition of Speech Foundation Model via Variance-Invariance-Covariance Regularization
di: Ahn, Hyebin, et al.
Pubblicazione: (2025)
di: Ahn, Hyebin, et al.
Pubblicazione: (2025)
Data-Efficient ASR Personalization for Non-Normative Speech Using an Uncertainty-Based Phoneme Difficulty Score for Guided Sampling
di: Pokel, Niclas, et al.
Pubblicazione: (2025)
di: Pokel, Niclas, et al.
Pubblicazione: (2025)
Efficient Finetuning for Dimensional Speech Emotion Recognition in the Age of Transformers
di: Sampath, Aneesha, et al.
Pubblicazione: (2025)
di: Sampath, Aneesha, et al.
Pubblicazione: (2025)
Multistage Fine-tuning Strategies for Automatic Speech Recognition in Low-resource Languages
di: Pillai, Leena G, et al.
Pubblicazione: (2024)
di: Pillai, Leena G, et al.
Pubblicazione: (2024)
Boosting Code-Switching ASR with Mixture of Experts Enhanced Speech-Conditioned LLM
di: Zhang, Fengrun, et al.
Pubblicazione: (2024)
di: Zhang, Fengrun, et al.
Pubblicazione: (2024)
CleanMel: Mel-Spectrogram Enhancement for Improving Both Speech Quality and ASR
di: Shao, Nian, et al.
Pubblicazione: (2025)
di: Shao, Nian, et al.
Pubblicazione: (2025)
Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics
di: Goswami, Mandip
Pubblicazione: (2026)
di: Goswami, Mandip
Pubblicazione: (2026)
GEC-RAG: Improving Generative Error Correction via Retrieval-Augmented Generation for Automatic Speech Recognition Systems
di: Robatian, Amin, et al.
Pubblicazione: (2025)
di: Robatian, Amin, et al.
Pubblicazione: (2025)
Low-Rank and Sparse Model Merging for Multi-Lingual Speech Recognition and Translation
di: Zhao, Qiuming, et al.
Pubblicazione: (2025)
di: Zhao, Qiuming, et al.
Pubblicazione: (2025)
Toward Efficient Speech Emotion Recognition via Spectral Learning and Attention
di: Lee, HyeYoung, et al.
Pubblicazione: (2025)
di: Lee, HyeYoung, et al.
Pubblicazione: (2025)
When De-noising Hurts: A Systematic Study of Speech Enhancement Effects on Modern Medical ASR Systems
di: Chondhekar, Sujal, et al.
Pubblicazione: (2025)
di: Chondhekar, Sujal, et al.
Pubblicazione: (2025)
Whisper in Medusa's Ear: Multi-head Efficient Decoding for Transformer-based ASR
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2024)
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2024)
Interpreting Pretrained Speech Models for Automatic Speech Assessment of Voice Disorders
di: Lau, Hok-Shing, et al.
Pubblicazione: (2024)
di: Lau, Hok-Shing, et al.
Pubblicazione: (2024)
Enhanced Speech Emotion Recognition with Efficient Channel Attention Guided Deep CNN-BiLSTM Framework
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024)
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024)
Enhancing Synthetic Training Data for Speech Commands: From ASR-Based Filtering to Domain Adaptation in SSL Latent Space
di: Quintas, Sebastião, et al.
Pubblicazione: (2024)
di: Quintas, Sebastião, et al.
Pubblicazione: (2024)
TinyML for Speech Recognition
di: Barovic, Andrew, et al.
Pubblicazione: (2025)
di: Barovic, Andrew, et al.
Pubblicazione: (2025)
SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding
di: Wei, Linye, et al.
Pubblicazione: (2025)
di: Wei, Linye, et al.
Pubblicazione: (2025)
Automatic Speech Recognition using Advanced Deep Learning Approaches: A survey
di: Kheddar, Hamza, et al.
Pubblicazione: (2024)
di: Kheddar, Hamza, et al.
Pubblicazione: (2024)
Clustering and Mining Accented Speech for Inclusive and Fair Speech Recognition
di: Kim, Jaeyoung, et al.
Pubblicazione: (2024)
di: Kim, Jaeyoung, et al.
Pubblicazione: (2024)
Documenti analoghi
-
VoxServe: Streaming-Centric Serving System for Speech Language Models
di: Kamahori, Keisuke, et al.
Pubblicazione: (2026) -
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
di: Wang, He, et al.
Pubblicazione: (2024) -
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models
di: Feng, Chen, et al.
Pubblicazione: (2025) -
Speech Recognition on TV Series with Video-guided Post-ASR Correction
di: Yang, Haoyuan, et al.
Pubblicazione: (2025) -
Self-supervised ASR Models and Features For Dysarthric and Elderly Speech Recognition
di: Hu, Shujie, et al.
Pubblicazione: (2024)