Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sang, Mufan, Hansen, John H. L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UniPET-SPK: A Unified Framework for Parameter-Efficient Tuning of Pre-trained Speech Models for Robust Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2025)
von: Sang, Mufan, et al.
Veröffentlicht: (2025)
Adversarial Speaker Distillation for Countermeasure Model on Automatic Speaker Verification
von: Liao, Yen-Lun, et al.
Veröffentlicht: (2022)
von: Liao, Yen-Lun, et al.
Veröffentlicht: (2022)
Transcription-Free Fine-Tuning of Speech Separation Models for Noisy and Reverberant Multi-Speaker Automatic Speech Recognition
von: Ravenscroft, William, et al.
Veröffentlicht: (2024)
von: Ravenscroft, William, et al.
Veröffentlicht: (2024)
Speaker Disentanglement of Speech Pre-trained Model Based on Interpretability
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
Unseen Speaker and Language Adaptation for Lightweight Text-To-Speech with Adapters
von: Falai, Alessio, et al.
Veröffentlicht: (2025)
von: Falai, Alessio, et al.
Veröffentlicht: (2025)
Multiple Choice Learning for Efficient Speech Separation with Many Speakers
von: Perera, David, et al.
Veröffentlicht: (2024)
von: Perera, David, et al.
Veröffentlicht: (2024)
HiddenSpeaker: Generate Imperceptible Unlearnable Audios for Speaker Verification System
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2024)
Label-Efficient Self-Supervised Speaker Verification With Information Maximization and Contrastive Learning
von: Lepage, Théo, et al.
Veröffentlicht: (2022)
von: Lepage, Théo, et al.
Veröffentlicht: (2022)
Adversarial Data Augmentation for Robust Speaker Verification
von: Zhou, Zhenyu, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenyu, et al.
Veröffentlicht: (2024)
Gradient Norm-based Fine-Tuning for Backdoor Defense in Automatic Speech Recognition
von: Zhou, Nanjun, et al.
Veröffentlicht: (2025)
von: Zhou, Nanjun, et al.
Veröffentlicht: (2025)
Adapter-Based Multi-Agent AVSR Extension for Pre-Trained ASR Models
von: Simic, Christopher, et al.
Veröffentlicht: (2025)
von: Simic, Christopher, et al.
Veröffentlicht: (2025)
Zero-Shot Multi-Lingual Speaker Verification in Clinical Trials
von: Akram, Ali, et al.
Veröffentlicht: (2024)
von: Akram, Ali, et al.
Veröffentlicht: (2024)
TinySV: Speaker Verification in TinyML with On-device Learning
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)
von: Pavan, Massimo, et al.
Veröffentlicht: (2024)
Robust Channel Learning for Large-Scale Radio Speaker Verification
von: Yang, Wenhao, et al.
Veröffentlicht: (2024)
von: Yang, Wenhao, et al.
Veröffentlicht: (2024)
Improving the Adversarial Robustness for Speaker Verification by Self-Supervised Learning
von: Wu, Haibin, et al.
Veröffentlicht: (2021)
von: Wu, Haibin, et al.
Veröffentlicht: (2021)
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
von: Chen, Zhengyang, et al.
Veröffentlicht: (2024)
von: Chen, Zhengyang, et al.
Veröffentlicht: (2024)
Self-Supervised Frameworks for Speaker Verification via Bootstrapped Positive Sampling
von: Lepage, Theo, et al.
Veröffentlicht: (2025)
von: Lepage, Theo, et al.
Veröffentlicht: (2025)
SVSNet+: Enhancing Speaker Voice Similarity Assessment Models with Representations from Speech Foundation Models
von: Yin, Chun, et al.
Veröffentlicht: (2024)
von: Yin, Chun, et al.
Veröffentlicht: (2024)
Towards Supervised Performance on Speaker Verification with Self-Supervised Learning by Leveraging Large-Scale ASR Models
von: Miara, Victor, et al.
Veröffentlicht: (2024)
von: Miara, Victor, et al.
Veröffentlicht: (2024)
ELP-Adapters: Parameter Efficient Adapter Tuning for Various Speech Processing Tasks
von: Inoue, Nakamasa, et al.
Veröffentlicht: (2024)
von: Inoue, Nakamasa, et al.
Veröffentlicht: (2024)
AMT-APC: Automatic Piano Cover by Fine-Tuning an Automatic Music Transcription Model
von: Komiya, Kazuma, et al.
Veröffentlicht: (2024)
von: Komiya, Kazuma, et al.
Veröffentlicht: (2024)
Generative Pre-training for Speech with Flow Matching
von: Liu, Alexander H., et al.
Veröffentlicht: (2023)
von: Liu, Alexander H., et al.
Veröffentlicht: (2023)
Multi-stream Convolutional Neural Network with Frequency Selection for Robust Speaker Verification
von: Yao, Wei, et al.
Veröffentlicht: (2020)
von: Yao, Wei, et al.
Veröffentlicht: (2020)
LMD: A Learnable Mask Network to Detect Adversarial Examples for Speaker Verification
von: Chen, Xing, et al.
Veröffentlicht: (2022)
von: Chen, Xing, et al.
Veröffentlicht: (2022)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Phonetic Richness for Improved Automatic Speaker Verification
von: Klein, Nicholas, et al.
Veröffentlicht: (2024)
von: Klein, Nicholas, et al.
Veröffentlicht: (2024)
Source -Free Domain Adaptation for Speaker Verification in Data-Scarce Languages and Noisy Channels
von: Elia, Shlomo Salo, et al.
Veröffentlicht: (2024)
von: Elia, Shlomo Salo, et al.
Veröffentlicht: (2024)
VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis
von: Lin, Weiwei, et al.
Veröffentlicht: (2024)
von: Lin, Weiwei, et al.
Veröffentlicht: (2024)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
Pre-training Feature Guided Diffusion Model for Speech Enhancement
von: Yang, Yiyuan, et al.
Veröffentlicht: (2024)
von: Yang, Yiyuan, et al.
Veröffentlicht: (2024)
Windowed SummaryMixing: An Efficient Fine-Tuning of Self-Supervised Learning Models for Low-resource Speech Recognition
von: Menon, Aditya Srinivas, et al.
Veröffentlicht: (2026)
von: Menon, Aditya Srinivas, et al.
Veröffentlicht: (2026)
Keyword-Guided Adaptation of Automatic Speech Recognition
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models
von: Feng, Chen, et al.
Veröffentlicht: (2025)
von: Feng, Chen, et al.
Veröffentlicht: (2025)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
von: Lu, Cheng, et al.
Veröffentlicht: (2024)
von: Lu, Cheng, et al.
Veröffentlicht: (2024)
Self-Tuning Spectral Clustering for Speaker Diarization
von: Raghav, Nikhil, et al.
Veröffentlicht: (2024)
von: Raghav, Nikhil, et al.
Veröffentlicht: (2024)
Meta Audiobox Aesthetics: Unified Automatic Quality Assessment for Speech, Music, and Sound
von: Tjandra, Andros, et al.
Veröffentlicht: (2025)
von: Tjandra, Andros, et al.
Veröffentlicht: (2025)
Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2024)
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2024)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
von: Kaloga, Yacouba, et al.
Veröffentlicht: (2024)
Cross-Domain Knowledge Transfer for Underwater Acoustic Classification Using Pre-trained Models
von: Mohammadi, Amirmohammad, et al.
Veröffentlicht: (2024)
von: Mohammadi, Amirmohammad, et al.
Veröffentlicht: (2024)
An Enhanced Audio Feature Tailored for Anomalous Sound Detection Based on Pre-trained Models
von: Zhong, Guirui, et al.
Veröffentlicht: (2025)
von: Zhong, Guirui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UniPET-SPK: A Unified Framework for Parameter-Efficient Tuning of Pre-trained Speech Models for Robust Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2025) -
Adversarial Speaker Distillation for Countermeasure Model on Automatic Speaker Verification
von: Liao, Yen-Lun, et al.
Veröffentlicht: (2022) -
Transcription-Free Fine-Tuning of Speech Separation Models for Noisy and Reverberant Multi-Speaker Automatic Speech Recognition
von: Ravenscroft, William, et al.
Veröffentlicht: (2024) -
Speaker Disentanglement of Speech Pre-trained Model Based on Interpretability
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025) -
Unseen Speaker and Language Adaptation for Lightweight Text-To-Speech with Adapters
von: Falai, Alessio, et al.
Veröffentlicht: (2025)