Large Vocabulary Spontaneous Speech Recognition for Tigrigna
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kahsu, Ataklti, Teferra, Solomon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction
von: Cheripally, Sowmya
Veröffentlicht: (2024)
von: Cheripally, Sowmya
Veröffentlicht: (2024)
Training for Speech Recognition on Coprocessors
von: Baunsgaard, Sebastian, et al.
Veröffentlicht: (2020)
von: Baunsgaard, Sebastian, et al.
Veröffentlicht: (2020)
Real-time CARFAC Cochlea Model Acceleration on FPGA for Underwater Acoustic Sensing Systems
von: Bremer, Bram, et al.
Veröffentlicht: (2025)
von: Bremer, Bram, et al.
Veröffentlicht: (2025)
Quantization for OpenAI's Whisper Models: A Comparative Analysis
von: Andreyev, Allison
Veröffentlicht: (2025)
von: Andreyev, Allison
Veröffentlicht: (2025)
Spectral oversubtraction? An approach for speech enhancement after robot ego speech filtering in semi-real-time
von: Li, Yue, et al.
Veröffentlicht: (2024)
von: Li, Yue, et al.
Veröffentlicht: (2024)
SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning
von: Chopra, Anuradha, et al.
Veröffentlicht: (2025)
von: Chopra, Anuradha, et al.
Veröffentlicht: (2025)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
von: Bhadra, Dipayan, et al.
Veröffentlicht: (2025)
von: Bhadra, Dipayan, et al.
Veröffentlicht: (2025)
Toward Low-Latency End-to-End Voice Agents for Telecommunications Using Streaming ASR, Quantized LLMs, and Real-Time TTS
von: Ethiraj, Vignesh, et al.
Veröffentlicht: (2025)
von: Ethiraj, Vignesh, et al.
Veröffentlicht: (2025)
Benchmarking Foundation Speech and Language Models for Alzheimer's Disease and Related Dementia Detection from Spontaneous Speech
von: Li, Jingyu, et al.
Veröffentlicht: (2025)
von: Li, Jingyu, et al.
Veröffentlicht: (2025)
SARA: Stress Test Reasoning in Audio Deepfake Detection
von: Nguyen, Binh, et al.
Veröffentlicht: (2026)
von: Nguyen, Binh, et al.
Veröffentlicht: (2026)
Improving Speech Recognition Accuracy Using Custom Language Models with the Vosk Toolkit
von: Soni, Aniket Abhishek
Veröffentlicht: (2025)
von: Soni, Aniket Abhishek
Veröffentlicht: (2025)
Monaural Multi-Speaker Speech Separation Using Efficient Transformer Model
von: Rijal, S., et al.
Veröffentlicht: (2023)
von: Rijal, S., et al.
Veröffentlicht: (2023)
Splitformer: An improved early-exit architecture for automatic speech recognition on edge devices
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2025)
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2025)
Are Music Foundation Models Better at Singing Voice Deepfake Detection? Far-Better Fuse them with Speech Foundation Models
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Multi-blank Transducers for Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
Speech Slytherin: Examining the Performance and Efficiency of Mamba for Speech Separation, Recognition, and Synthesis
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
Test-Time Adaptation for Speech Emotion Recognition
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026)
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026)
Drax: Speech Recognition with Discrete Flow Matching
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
Adapting WavLM for Speech Emotion Recognition
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)
Keyword-Guided Adaptation of Automatic Speech Recognition
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
CogniVoice: Multimodal and Multilingual Fusion Networks for Mild Cognitive Impairment Assessment from Spontaneous Speech
von: Cheng, Jiali, et al.
Veröffentlicht: (2024)
von: Cheng, Jiali, et al.
Veröffentlicht: (2024)
Developing Acoustic Models for Automatic Speech Recognition in Swedish
von: Salvi, Giampiero
Veröffentlicht: (2024)
von: Salvi, Giampiero
Veröffentlicht: (2024)
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023)
TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition
von: Chen, Chengxin, et al.
Veröffentlicht: (2024)
von: Chen, Chengxin, et al.
Veröffentlicht: (2024)
Parameter Efficient Finetuning for Speech Emotion Recognition and Domain Adaptation
von: Lashkarashvili, Nineli, et al.
Veröffentlicht: (2024)
von: Lashkarashvili, Nineli, et al.
Veröffentlicht: (2024)
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2025)
SeQuiFi: Mitigating Catastrophic Forgetting in Speech Emotion Recognition with Sequential Class-Finetuning
von: Jain, Sarthak, et al.
Veröffentlicht: (2024)
von: Jain, Sarthak, et al.
Veröffentlicht: (2024)
Transcription-Free Fine-Tuning of Speech Separation Models for Noisy and Reverberant Multi-Speaker Automatic Speech Recognition
von: Ravenscroft, William, et al.
Veröffentlicht: (2024)
von: Ravenscroft, William, et al.
Veröffentlicht: (2024)
Multimodal Attention Merging for Improved Speech Recognition and Audio Event Classification
von: Sundar, Anirudh S., et al.
Veröffentlicht: (2023)
von: Sundar, Anirudh S., et al.
Veröffentlicht: (2023)
Chunked Attention-based Encoder-Decoder Model for Streaming Speech Recognition
von: Zeineldeen, Mohammad, et al.
Veröffentlicht: (2023)
von: Zeineldeen, Mohammad, et al.
Veröffentlicht: (2023)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
Adapting Automatic Speech Recognition for Accented Air Traffic Control Communications
von: Wee, Marcus Yu Zhe, et al.
Veröffentlicht: (2025)
von: Wee, Marcus Yu Zhe, et al.
Veröffentlicht: (2025)
Sequence-Level Unsupervised Training in Speech Recognition: A Theoretical Study
von: Yang, Zijian, et al.
Veröffentlicht: (2026)
von: Yang, Zijian, et al.
Veröffentlicht: (2026)
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
von: Du, Zongyang, et al.
Veröffentlicht: (2025)
von: Du, Zongyang, et al.
Veröffentlicht: (2025)
STAR: Speech-to-Audio Generation via Representation Learning
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
Detecting Check-Worthy Claims in Political Debates, Speeches, and Interviews Using Audio Data
von: Ivanov, Petar, et al.
Veröffentlicht: (2023)
von: Ivanov, Petar, et al.
Veröffentlicht: (2023)
Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing
von: Hizabri, Marouane El, et al.
Veröffentlicht: (2026)
von: Hizabri, Marouane El, et al.
Veröffentlicht: (2026)
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models
von: Feng, Chen, et al.
Veröffentlicht: (2025)
von: Feng, Chen, et al.
Veröffentlicht: (2025)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
A Layer-Anchoring Strategy for Enhancing Cross-Lingual Speech Emotion Recognition
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction
von: Cheripally, Sowmya
Veröffentlicht: (2024) -
Training for Speech Recognition on Coprocessors
von: Baunsgaard, Sebastian, et al.
Veröffentlicht: (2020) -
Real-time CARFAC Cochlea Model Acceleration on FPGA for Underwater Acoustic Sensing Systems
von: Bremer, Bram, et al.
Veröffentlicht: (2025) -
Quantization for OpenAI's Whisper Models: A Comparative Analysis
von: Andreyev, Allison
Veröffentlicht: (2025) -
Spectral oversubtraction? An approach for speech enhancement after robot ego speech filtering in semi-real-time
von: Li, Yue, et al.
Veröffentlicht: (2024)