Dynamic Multi-Expert Projectors with Stabilized Routing for Multilingual Speech Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Pandey, Isha, Mittal, Ashish, Bahuguna, Vartul, Ramakrishnan, Ganesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Language translation, and change of accent for speech-to-speech task using diffusion model
por: Mishra, Abhishek, et al.
Publicado: (2025)
por: Mishra, Abhishek, et al.
Publicado: (2025)
A2TTS: TTS for Low Resource Indian Languages
por: Bhadoriya, Ayush Singh, et al.
Publicado: (2025)
por: Bhadoriya, Ayush Singh, et al.
Publicado: (2025)
A Three-Pronged Approach to Cross-Lingual Adaptation with Multilingual LLMs
por: Singh, Vaibhav, et al.
Publicado: (2024)
por: Singh, Vaibhav, et al.
Publicado: (2024)
SpeechMapper: Speech-to-text Embedding Projector for LLMs
por: Mohapatra, Biswesh, et al.
Publicado: (2026)
por: Mohapatra, Biswesh, et al.
Publicado: (2026)
Multilingual Routing in Mixture-of-Experts
por: Bandarkar, Lucas, et al.
Publicado: (2025)
por: Bandarkar, Lucas, et al.
Publicado: (2025)
Automatic Speech Recognition for Hindi
por: Saha, Anish, et al.
Publicado: (2024)
por: Saha, Anish, et al.
Publicado: (2024)
Understanding Multilingualism in Mixture-of-Experts LLMs: Routing Mechanism, Expert Specialization, and Layerwise Steering
por: Chen, Yuxin, et al.
Publicado: (2026)
por: Chen, Yuxin, et al.
Publicado: (2026)
Multilingual Tokenization through the Lens of Indian Languages: Challenges and Insights
por: Karthika, N J, et al.
Publicado: (2025)
por: Karthika, N J, et al.
Publicado: (2025)
LexGen: Domain-aware Multilingual Lexicon Generation
por: Maheshwari, Ayush, et al.
Publicado: (2024)
por: Maheshwari, Ayush, et al.
Publicado: (2024)
Zipper-LoRA: Dynamic Parameter Decoupling for Speech-LLM based Multilingual Speech Recognition
por: Mei, Yuxiang, et al.
Publicado: (2026)
por: Mei, Yuxiang, et al.
Publicado: (2026)
Efficient infusion of self-supervised representations in Automatic Speech Recognition
por: Prabhu, Darshan, et al.
Publicado: (2024)
por: Prabhu, Darshan, et al.
Publicado: (2024)
Mixture of LoRA Experts for Low-Resourced Multi-Accent Automatic Speech Recognition
por: Bagat, Raphaël, et al.
Publicado: (2025)
por: Bagat, Raphaël, et al.
Publicado: (2025)
Multilingual Phonological Feature Recognition with Self-Supervised Speech Models
por: Hernandez, Abner, et al.
Publicado: (2026)
por: Hernandez, Abner, et al.
Publicado: (2026)
Multilingual Extraction and Recognition of Implicit Discourse Relations in Speech and Text
por: Ruby, Ahmed, et al.
Publicado: (2026)
por: Ruby, Ahmed, et al.
Publicado: (2026)
Omni-Router: Sharing Routing Decisions in Sparse Mixture-of-Experts for Speech Recognition
por: Gu, Zijin, et al.
Publicado: (2025)
por: Gu, Zijin, et al.
Publicado: (2025)
Routing-Aligned Fine-Tuning for Multilingual Downstream Tasks in Mixture-of-Experts Models
por: Deng, Guanzhi, et al.
Publicado: (2026)
por: Deng, Guanzhi, et al.
Publicado: (2026)
Omnilingual ASR: Open-Source Multilingual Speech Recognition for 1600+ Languages
por: Omnilingual ASR team, et al.
Publicado: (2025)
por: Omnilingual ASR team, et al.
Publicado: (2025)
Lost in Transcription: How Speech-to-Text Errors Derail Code Understanding
por: Havare, Jayant, et al.
Publicado: (2026)
por: Havare, Jayant, et al.
Publicado: (2026)
CodeVaani: A Multilingual, Voice-Based Code Learning Assistant
por: Havare, Jayant, et al.
Publicado: (2025)
por: Havare, Jayant, et al.
Publicado: (2025)
MultiMed: Multilingual Medical Speech Recognition via Attention Encoder Decoder
por: Le-Duc, Khai, et al.
Publicado: (2024)
por: Le-Duc, Khai, et al.
Publicado: (2024)
Multi-Teacher Language-Aware Knowledge Distillation for Multilingual Speech Emotion Recognition
por: Bijoy, Mehedi Hasan, et al.
Publicado: (2025)
por: Bijoy, Mehedi Hasan, et al.
Publicado: (2025)
ViSpeR: Multilingual Audio-Visual Speech Recognition
por: Narayan, Sanath, et al.
Publicado: (2024)
por: Narayan, Sanath, et al.
Publicado: (2024)
MSNER: A Multilingual Speech Dataset for Named Entity Recognition
por: Meeus, Quentin, et al.
Publicado: (2024)
por: Meeus, Quentin, et al.
Publicado: (2024)
Ethio-ASR: Joint Multilingual Speech Recognition and Language Identification for Ethiopian Languages
por: Abdullah, Badr M., et al.
Publicado: (2026)
por: Abdullah, Badr M., et al.
Publicado: (2026)
Streaming Speech-to-Confusion Network Speech Recognition
por: Filimonov, Denis, et al.
Publicado: (2023)
por: Filimonov, Denis, et al.
Publicado: (2023)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
por: Zhou, Hao, et al.
Publicado: (2024)
por: Zhou, Hao, et al.
Publicado: (2024)
DICTDIS: Dictionary Constrained Disambiguation for Improved NMT
por: Maheshwari, Ayush, et al.
Publicado: (2022)
por: Maheshwari, Ayush, et al.
Publicado: (2022)
Twists, Humps, and Pebbles: Multilingual Speech Recognition Models Exhibit Gender Performance Gaps
por: Attanasio, Giuseppe, et al.
Publicado: (2024)
por: Attanasio, Giuseppe, et al.
Publicado: (2024)
Dynamic Language Group-Based MoE: Enhancing Code-Switching Speech Recognition with Hierarchical Routing
por: Huang, Hukai, et al.
Publicado: (2024)
por: Huang, Hukai, et al.
Publicado: (2024)
Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition
por: Lee, Wonjun, et al.
Publicado: (2026)
por: Lee, Wonjun, et al.
Publicado: (2026)
Multilingual DistilWhisper: Efficient Distillation of Multi-task Speech Models via Language-Specific Experts
por: Ferraz, Thomas Palmeira, et al.
Publicado: (2023)
por: Ferraz, Thomas Palmeira, et al.
Publicado: (2023)
A Unified Speech LLM for Diarization and Speech Recognition in Multilingual Conversations
por: Saengthong, Phurich, et al.
Publicado: (2025)
por: Saengthong, Phurich, et al.
Publicado: (2025)
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training
por: Ponkshe, Kaustubh, et al.
Publicado: (2024)
por: Ponkshe, Kaustubh, et al.
Publicado: (2024)
EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
por: Ma, Ziyang, et al.
Publicado: (2024)
por: Ma, Ziyang, et al.
Publicado: (2024)
ILT-Iterative LoRA Training through Focus-Feedback-Fix for Multilingual Speech Recognition
por: Meng, Qingliang, et al.
Publicado: (2025)
por: Meng, Qingliang, et al.
Publicado: (2025)
Group then Scale: Dynamic Mixture-of-Experts Multilingual Language Model
por: Li, Chong, et al.
Publicado: (2025)
por: Li, Chong, et al.
Publicado: (2025)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
por: Adila, Aulia, et al.
Publicado: (2024)
por: Adila, Aulia, et al.
Publicado: (2024)
The Art of Breaking Words: Rethinking Multilingual Tokenizer Design
por: Thakur, Aamod, et al.
Publicado: (2025)
por: Thakur, Aamod, et al.
Publicado: (2025)
Learning to Route Languages for Multilingual Policy Optimization
por: Guo, Geyang, et al.
Publicado: (2026)
por: Guo, Geyang, et al.
Publicado: (2026)
Robust Audiovisual Speech Recognition Models with Mixture-of-Experts
por: Wu, Yihan, et al.
Publicado: (2024)
por: Wu, Yihan, et al.
Publicado: (2024)
Ejemplares similares
-
Language translation, and change of accent for speech-to-speech task using diffusion model
por: Mishra, Abhishek, et al.
Publicado: (2025) -
A2TTS: TTS for Low Resource Indian Languages
por: Bhadoriya, Ayush Singh, et al.
Publicado: (2025) -
A Three-Pronged Approach to Cross-Lingual Adaptation with Multilingual LLMs
por: Singh, Vaibhav, et al.
Publicado: (2024) -
SpeechMapper: Speech-to-text Embedding Projector for LLMs
por: Mohapatra, Biswesh, et al.
Publicado: (2026) -
Multilingual Routing in Mixture-of-Experts
por: Bandarkar, Lucas, et al.
Publicado: (2025)