Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Junyi, Zhang, Lin, Han, Jiangyu, Plchot, Oldřich, Rohdin, Johan, Stafylakis, Themos, Wang, Shuai, Černocký, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024)
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024)
State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data
von: Barahona, Sara, et al.
Veröffentlicht: (2024)
von: Barahona, Sara, et al.
Veröffentlicht: (2024)
Efficient and Generalizable Speaker Diarization via Structured Pruning of Self-Supervised Models
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
BUT Systems for WildSpoof Challenge: SASV in the Wild
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
Fine-tune Before Structured Pruning: Towards Compact and Accurate Self-Supervised Models for Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
Target Speech Extraction with Pre-trained Self-supervised Learning Models
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
BUT Systems and Analyses for the ASVspoof 5 Challenge
von: Rohdin, Johan, et al.
Veröffentlicht: (2024)
von: Rohdin, Johan, et al.
Veröffentlicht: (2024)
Bayesian Learning for Domain-Invariant Speaker Verification and Anti-Spoofing
von: Li, Jin, et al.
Veröffentlicht: (2025)
von: Li, Jin, et al.
Veröffentlicht: (2025)
BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
Probing Self-supervised Learning Models with Target Speech Extraction
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
Leveraging Self-Supervised Learning for Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2024)
von: Han, Jiangyu, et al.
Veröffentlicht: (2024)
Spatially Aware Self-Supervised Models for Multi-Channel Neural Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
Analysis of ABC Frontend Audio Systems for the NIST-SRE24
von: Barahona, Sara, et al.
Veröffentlicht: (2025)
von: Barahona, Sara, et al.
Veröffentlicht: (2025)
Approaching Dialogue State Tracking via Aligning Speech Encoders and LLMs
von: Sedláček, Šimon, et al.
Veröffentlicht: (2025)
von: Sedláček, Šimon, et al.
Veröffentlicht: (2025)
Joint Speech and Text Training for LLM-Based End-to-End Spoken Dialogue State Tracking
von: Vendrame, Katia, et al.
Veröffentlicht: (2025)
von: Vendrame, Katia, et al.
Veröffentlicht: (2025)
Integrated Spoofing-Robust Automatic Speaker Verification via a Three-Class Formulation and LLR
von: Tan, Kai, et al.
Veröffentlicht: (2026)
von: Tan, Kai, et al.
Veröffentlicht: (2026)
Joint Optimization of Speaker and Spoof Detectors for Spoofing-Robust Automatic Speaker Verification
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2025)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2025)
Do End-to-End Neural Diarization Attractors Need to Encode Speaker Characteristic Information?
von: Zhang, Lin, et al.
Veröffentlicht: (2024)
von: Zhang, Lin, et al.
Veröffentlicht: (2024)
DiCoW: Diarization-Conditioned Whisper for Target Speaker Automatic Speech Recognition
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
DiaPer: End-to-End Neural Diarization with Perceiver-Based Attractors
von: Landini, Federico, et al.
Veröffentlicht: (2023)
von: Landini, Federico, et al.
Veröffentlicht: (2023)
Optimizing a-DCF for Spoofing-Robust Speaker Verification
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
SA-WavLM: Speaker-Aware Self-Supervised Pre-training for Mixture Speech
von: Lin, Jingru, et al.
Veröffentlicht: (2024)
von: Lin, Jingru, et al.
Veröffentlicht: (2024)
SV-Mixer: Replacing the Transformer Encoder with Lightweight MLPs for Self-Supervised Model Compression in Speaker Verification
von: Heo, Jungwoo, et al.
Veröffentlicht: (2025)
von: Heo, Jungwoo, et al.
Veröffentlicht: (2025)
A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification
von: Zhang, You, et al.
Veröffentlicht: (2022)
von: Zhang, You, et al.
Veröffentlicht: (2022)
Investigating the Potential of Multi-Stage Score Fusion in Spoofing-Aware Speaker Verification
von: Kurnaz, Oguzhan, et al.
Veröffentlicht: (2025)
von: Kurnaz, Oguzhan, et al.
Veröffentlicht: (2025)
BUT System for the MLC-SLM Challenge
von: Polok, Alexander, et al.
Veröffentlicht: (2025)
von: Polok, Alexander, et al.
Veröffentlicht: (2025)
Spoofing-Aware Speaker Verification Robust Against Domain and Channel Mismatches
von: Zeng, Chang, et al.
Veröffentlicht: (2024)
von: Zeng, Chang, et al.
Veröffentlicht: (2024)
Spoofing-Robust Speaker Verification Using Parallel Embedding Fusion: BTU Speech Group's Approach for ASVspoof5 Challenge
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
Asymmetric Clean Segments-Guided Self-Supervised Learning for Robust Speaker Verification
von: Gan, Chong-Xin, et al.
Veröffentlicht: (2023)
von: Gan, Chong-Xin, et al.
Veröffentlicht: (2023)
ELEAT-SAGA: Early & Late Integration with Evading Alternating Training for Spoof-Robust Speaker Verification
von: Asali, Amro, et al.
Veröffentlicht: (2026)
von: Asali, Amro, et al.
Veröffentlicht: (2026)
Target Speaker ASR with Whisper
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Spoofing-Aware Speaker Verification via Wavelet Prompt Tuning and Multi-Model Ensembles
von: Farhadipour, Aref, et al.
Veröffentlicht: (2026)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2026)
Mind the Gap: Impact of Synthetic Conversational Data on Multi-Talker ASR and Speaker Diarization
von: Polok, Alexander, et al.
Veröffentlicht: (2026)
von: Polok, Alexander, et al.
Veröffentlicht: (2026)
Experimenting with Additive Margins for Contrastive Self-Supervised Speaker Verification
von: Lepage, Theo, et al.
Veröffentlicht: (2023)
von: Lepage, Theo, et al.
Veröffentlicht: (2023)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models
von: C, Shiva Kumar, et al.
Veröffentlicht: (2025)
von: C, Shiva Kumar, et al.
Veröffentlicht: (2025)
Improving Automatic Speech Recognition with Decoder-Centric Regularisation in Encoder-Decoder Models
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features
von: van Rensburg, Kyle Janse, et al.
Veröffentlicht: (2026)
von: van Rensburg, Kyle Janse, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
von: Peng, Junyi, et al.
Veröffentlicht: (2024) -
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024) -
State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data
von: Barahona, Sara, et al.
Veröffentlicht: (2024) -
Efficient and Generalizable Speaker Diarization via Structured Pruning of Self-Supervised Models
von: Han, Jiangyu, et al.
Veröffentlicht: (2025) -
BUT Systems for WildSpoof Challenge: SASV in the Wild
von: Peng, Junyi, et al.
Veröffentlicht: (2025)