UME: Upcycling Mixture-of-Experts for Scalable and Efficient Automatic Speech Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fu, Li, Yu, Shanyong, Li, Siqi, Fan, Lu, Wu, Youzheng, He, Xiaodong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PAC: Pronunciation-Aware Contextualized Large Language Model-based Automatic Speech Recognition
von: Fu, Li, et al.
Veröffentlicht: (2025)
von: Fu, Li, et al.
Veröffentlicht: (2025)
MOPSA: Mixture of Prompt-Experts Based Speaker Adaptation for Elderly Speech Recognition
von: Deng, Chengxi, et al.
Veröffentlicht: (2025)
von: Deng, Chengxi, et al.
Veröffentlicht: (2025)
CAMEL: Cross-Attention Enhanced Mixture-of-Experts and Language Bias for Code-Switching Speech Recognition
von: Wang, He, et al.
Veröffentlicht: (2024)
von: Wang, He, et al.
Veröffentlicht: (2024)
FNH-TTS: Mixture-of-Experts Duration Modeling for Robust Neural Speech Synthesis
von: Meng, Qingliang, et al.
Veröffentlicht: (2025)
von: Meng, Qingliang, et al.
Veröffentlicht: (2025)
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)
Diarization-Aware Multi-Speaker Automatic Speech Recognition via Large Language Models
von: Lin, Yuke, et al.
Veröffentlicht: (2025)
von: Lin, Yuke, et al.
Veröffentlicht: (2025)
FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition
von: Kim, Jongsuk, et al.
Veröffentlicht: (2025)
von: Kim, Jongsuk, et al.
Veröffentlicht: (2025)
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS
von: Aronowitz, Hagai, et al.
Veröffentlicht: (2026)
von: Aronowitz, Hagai, et al.
Veröffentlicht: (2026)
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
von: Huang, Hukai, et al.
Veröffentlicht: (2024)
von: Huang, Hukai, et al.
Veröffentlicht: (2024)
Unsupervised Online Continual Learning for Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2024)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2024)
Using Songs to Improve Kazakh Automatic Speech Recognition
von: Yeshpanov, Rustem
Veröffentlicht: (2026)
von: Yeshpanov, Rustem
Veröffentlicht: (2026)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
The DKU System for Multi-Speaker Automatic Speech Recognition in MLC-SLM Challenge
von: Lin, Yuke, et al.
Veröffentlicht: (2025)
von: Lin, Yuke, et al.
Veröffentlicht: (2025)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
Non-Intrusive Automatic Speech Recognition Refinement: A Survey
von: Peyghan, Mohammad Reza, et al.
Veröffentlicht: (2025)
von: Peyghan, Mohammad Reza, et al.
Veröffentlicht: (2025)
Too Good to Be True: A Study on Modern Automatic Speech Recognition for the Evaluation of Speech Enhancement
von: de Oliveira, Danilo, et al.
Veröffentlicht: (2026)
von: de Oliveira, Danilo, et al.
Veröffentlicht: (2026)
Group-Aware Partial Model Merging for Children's Automatic Speech Recognition
von: Rolland, Thomas, et al.
Veröffentlicht: (2025)
von: Rolland, Thomas, et al.
Veröffentlicht: (2025)
Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition
von: Mu, Bingshen, et al.
Veröffentlicht: (2025)
von: Mu, Bingshen, et al.
Veröffentlicht: (2025)
Joint Learning using Mixture-of-Expert-Based Representation for Speech Enhancement and Robust Emotion Recognition
von: Tzeng, Jing-Tong, et al.
Veröffentlicht: (2025)
von: Tzeng, Jing-Tong, et al.
Veröffentlicht: (2025)
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
von: Wang, Pu, et al.
Veröffentlicht: (2024)
von: Wang, Pu, et al.
Veröffentlicht: (2024)
Multi-Scale Temporal Transformer For Speech Emotion Recognition
von: Li, Zhipeng, et al.
Veröffentlicht: (2024)
von: Li, Zhipeng, et al.
Veröffentlicht: (2024)
Robust Audiovisual Speech Recognition Models with Mixture-of-Experts
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
SpeechColab Leaderboard: An Open-Source Platform for Automatic Speech Recognition Evaluation
von: Du, Jiayu, et al.
Veröffentlicht: (2024)
von: Du, Jiayu, et al.
Veröffentlicht: (2024)
Improving Automatic Speech Recognition with Decoder-Centric Regularisation in Encoder-Decoder Models
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2022)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2022)
Adaptive Mixture of Low-Rank Experts for Robust Audio Spoofing Detection
von: Chen, Qixian, et al.
Veröffentlicht: (2025)
von: Chen, Qixian, et al.
Veröffentlicht: (2025)
AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition
von: Dai, Yuhang, et al.
Veröffentlicht: (2025)
von: Dai, Yuhang, et al.
Veröffentlicht: (2025)
Unifying Speech Recognition, Synthesis and Conversion with Autoregressive Transformers
von: Cai, Runyuan, et al.
Veröffentlicht: (2026)
von: Cai, Runyuan, et al.
Veröffentlicht: (2026)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
Detecting and Defending Against Adversarial Attacks on Automatic Speech Recognition via Diffusion Models
von: Kühne, Nikolai L., et al.
Veröffentlicht: (2024)
von: Kühne, Nikolai L., et al.
Veröffentlicht: (2024)
SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding
von: Wei, Linye, et al.
Veröffentlicht: (2025)
von: Wei, Linye, et al.
Veröffentlicht: (2025)
Dynamic Data Pruning for Automatic Speech Recognition
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
Dolphin: A Large-Scale Automatic Speech Recognition Model for Eastern Languages
von: Meng, Yangyang, et al.
Veröffentlicht: (2025)
von: Meng, Yangyang, et al.
Veröffentlicht: (2025)
Zero-Shot Recognition of Dysarthric Speech Using Commercial Automatic Speech Recognition and Multimodal Large Language Models
von: Alsayegh, Ali, et al.
Veröffentlicht: (2025)
von: Alsayegh, Ali, et al.
Veröffentlicht: (2025)
Enhancing Automatic Chord Recognition through LLM Chain-of-Thought Reasoning
von: Chang, Chih-Cheng, et al.
Veröffentlicht: (2025)
von: Chang, Chih-Cheng, et al.
Veröffentlicht: (2025)
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2023)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2023)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PAC: Pronunciation-Aware Contextualized Large Language Model-based Automatic Speech Recognition
von: Fu, Li, et al.
Veröffentlicht: (2025) -
MOPSA: Mixture of Prompt-Experts Based Speaker Adaptation for Elderly Speech Recognition
von: Deng, Chengxi, et al.
Veröffentlicht: (2025) -
CAMEL: Cross-Attention Enhanced Mixture-of-Experts and Language Bias for Code-Switching Speech Recognition
von: Wang, He, et al.
Veröffentlicht: (2024) -
FNH-TTS: Mixture-of-Experts Duration Modeling for Robust Neural Speech Synthesis
von: Meng, Qingliang, et al.
Veröffentlicht: (2025) -
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)