ELP-Adapters: Parameter Efficient Adapter Tuning for Various Speech Processing Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Inoue, Nakamasa, Otake, Shinta, Hirose, Takumi, Ohi, Masanari, Kawakami, Rei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech Synthesis
von: Nishimura, Yuto, et al.
Veröffentlicht: (2024)
von: Nishimura, Yuto, et al.
Veröffentlicht: (2024)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2024)
von: Sang, Mufan, et al.
Veröffentlicht: (2024)
Lightweight Zero-shot Text-to-Speech with Mixture of Adapters
von: Fujita, Kenichi, et al.
Veröffentlicht: (2024)
von: Fujita, Kenichi, et al.
Veröffentlicht: (2024)
Improving Code Switching with Supervised Fine Tuning and GELU Adapters
von: Pham, Linh
Veröffentlicht: (2025)
von: Pham, Linh
Veröffentlicht: (2025)
Unseen Speaker and Language Adaptation for Lightweight Text-To-Speech with Adapters
von: Falai, Alessio, et al.
Veröffentlicht: (2025)
von: Falai, Alessio, et al.
Veröffentlicht: (2025)
VoiceTailor: Lightweight Plug-In Adapter for Diffusion-Based Personalized Text-to-Speech
von: Kim, Heeseung, et al.
Veröffentlicht: (2024)
von: Kim, Heeseung, et al.
Veröffentlicht: (2024)
Adapter Incremental Continual Learning of Efficient Audio Spectrogram Transformers
von: Selvaraj, Nithish Muthuchamy, et al.
Veröffentlicht: (2023)
von: Selvaraj, Nithish Muthuchamy, et al.
Veröffentlicht: (2023)
Efficient Adapter Finetuning for Tail Languages in Streaming Multilingual ASR
von: Bai, Junwen, et al.
Veröffentlicht: (2024)
von: Bai, Junwen, et al.
Veröffentlicht: (2024)
SE/BN Adapter: Parametric Efficient Domain Adaptation for Speaker Recognition
von: Wang, Tianhao, et al.
Veröffentlicht: (2024)
von: Wang, Tianhao, et al.
Veröffentlicht: (2024)
Adapting Whisper for Parameter-efficient Code-Switching Speech Recognition via Soft Prompt Tuning
von: Yang, Hongli, et al.
Veröffentlicht: (2025)
von: Yang, Hongli, et al.
Veröffentlicht: (2025)
Efficient Adapter Tuning for Joint Singing Voice Beat and Downbeat Tracking with Self-supervised Learning Features
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
Language-Aware Prompt Tuning for Parameter-Efficient Seamless Language Expansion in Multilingual ASR
von: Yang, Hongli, et al.
Veröffentlicht: (2025)
von: Yang, Hongli, et al.
Veröffentlicht: (2025)
Audio Prompt Adapter: Unleashing Music Editing Abilities for Text-to-Music with Lightweight Finetuning
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2024)
H-PRM: A Pluggable Hotword Pre-Retrieval Module for Various Speech Recognition Systems
von: Dai, Huangyu, et al.
Veröffentlicht: (2025)
von: Dai, Huangyu, et al.
Veröffentlicht: (2025)
SpeechComposer: Unifying Multiple Speech Tasks with Prompt Composition
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts
von: Jin, Hojun, et al.
Veröffentlicht: (2025)
von: Jin, Hojun, et al.
Veröffentlicht: (2025)
MoE Adapter for Large Audio Language Models: Sparsity, Disentanglement, and Gradient-Conflict-Free
von: Lei, Yishu, et al.
Veröffentlicht: (2026)
von: Lei, Yishu, et al.
Veröffentlicht: (2026)
Text Prompt is Not Enough: Sound Event Enhanced Prompt Adapter for Target Style Audio Generation
von: Xiong, Chenxu, et al.
Veröffentlicht: (2024)
von: Xiong, Chenxu, et al.
Veröffentlicht: (2024)
Efficient Dialect-Aware Modeling and Conditioning for Low-Resource Taiwanese Hakka Speech Processing
von: Peng, An-Ci, et al.
Veröffentlicht: (2026)
von: Peng, An-Ci, et al.
Veröffentlicht: (2026)
Communication-Efficient Personalized Federated Learning for Speech-to-Text Tasks
von: Du, Yichao, et al.
Veröffentlicht: (2024)
von: Du, Yichao, et al.
Veröffentlicht: (2024)
Speech-Copilot: Leveraging Large Language Models for Speech Processing via Task Decomposition, Modularization, and Program Generation
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2024)
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2024)
Gammatonegram Representation for End-to-End Dysarthric Speech Processing Tasks: Speech Recognition, Speaker Identification, and Intelligibility Assessment
von: Farhadipour, Aref, et al.
Veröffentlicht: (2023)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2023)
Leveraging Parameter-Efficient Transfer Learning for Multi-Lingual Text-to-Speech Adaptation
von: Li, Yingting, et al.
Veröffentlicht: (2024)
von: Li, Yingting, et al.
Veröffentlicht: (2024)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
Exploring Adapter Design Tradeoffs for Low Resource Music Generation
von: Mehta, Atharva, et al.
Veröffentlicht: (2025)
von: Mehta, Atharva, et al.
Veröffentlicht: (2025)
Speech Prefix-Tuning with RNNT Loss for Improving LLM Predictions
von: Baskar, Murali Karthick, et al.
Veröffentlicht: (2024)
von: Baskar, Murali Karthick, et al.
Veröffentlicht: (2024)
SSVD-O: Parameter-Efficient Fine-Tuning with Structured SVD for Speech Recognition
von: Wang, Pu, et al.
Veröffentlicht: (2026)
von: Wang, Pu, et al.
Veröffentlicht: (2026)
Task-Lens: Cross-Task Utility Based Speech Dataset Profiling for Low-Resource Indian Languages
von: Sharma, Swati, et al.
Veröffentlicht: (2026)
von: Sharma, Swati, et al.
Veröffentlicht: (2026)
Efficient Streaming LLM for Speech Recognition
von: Jia, Junteng, et al.
Veröffentlicht: (2024)
von: Jia, Junteng, et al.
Veröffentlicht: (2024)
FLEURS-R: A Restored Multilingual Speech Corpus for Generation Tasks
von: Ma, Min, et al.
Veröffentlicht: (2024)
von: Ma, Min, et al.
Veröffentlicht: (2024)
Efficient Compression of Multitask Multilingual Speech Models
von: Ferraz, Thomas Palmeira
Veröffentlicht: (2024)
von: Ferraz, Thomas Palmeira
Veröffentlicht: (2024)
Unified Pathological Speech Analysis with Prompt Tuning
von: Yang, Fei, et al.
Veröffentlicht: (2024)
von: Yang, Fei, et al.
Veröffentlicht: (2024)
AAT: Adapting Audio Transformer for Various Acoustics Recognition Tasks
von: Liang, Yun, et al.
Veröffentlicht: (2024)
von: Liang, Yun, et al.
Veröffentlicht: (2024)
ECTSpeech: Enhancing Efficient Speech Synthesis via Easy Consistency Tuning
von: Zhu, Tao, et al.
Veröffentlicht: (2025)
von: Zhu, Tao, et al.
Veröffentlicht: (2025)
StyleSpeech: Parameter-efficient Fine Tuning for Pre-trained Controllable Text-to-Speech
von: Lou, Haowei, et al.
Veröffentlicht: (2024)
von: Lou, Haowei, et al.
Veröffentlicht: (2024)
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
von: Li, Hanzhao, et al.
Veröffentlicht: (2025)
von: Li, Hanzhao, et al.
Veröffentlicht: (2025)
PersonaTAB: Predicting Personality Traits using Textual, Acoustic, and Behavioral Cues in Fully-Duplex Speech Dialogs
von: Inoue, Sho, et al.
Veröffentlicht: (2025)
von: Inoue, Sho, et al.
Veröffentlicht: (2025)
DQLoRA: A Lightweight Domain-Aware Denoising ASR via Adapter-guided Distillation
von: Yang, Yiru
Veröffentlicht: (2025)
von: Yang, Yiru
Veröffentlicht: (2025)
Jointly Fine-Tuning "BERT-like" Self Supervised Models to Improve Multimodal Speech Emotion Recognition
von: Siriwardhana, Shamane, et al.
Veröffentlicht: (2020)
von: Siriwardhana, Shamane, et al.
Veröffentlicht: (2020)
Ähnliche Einträge
-
HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech Synthesis
von: Nishimura, Yuto, et al.
Veröffentlicht: (2024) -
Exploration of Adapter for Noise Robust Automatic Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2024) -
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2024) -
Lightweight Zero-shot Text-to-Speech with Mixture of Adapters
von: Fujita, Kenichi, et al.
Veröffentlicht: (2024) -
Improving Code Switching with Supervised Fine Tuning and GELU Adapters
von: Pham, Linh
Veröffentlicht: (2025)