Gespeichert in:
| Hauptverfasser: | Myoung, Jisoo, Han, Sangwook, Kim, Kihyuk, Shin, Jong Won |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2509.19721 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Speech Enhancement based on cascaded two flows
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
SV-Mixer: Replacing the Transformer Encoder with Lightweight MLPs for Self-Supervised Model Compression in Speaker Verification
von: Heo, Jungwoo, et al.
Veröffentlicht: (2025)
von: Heo, Jungwoo, et al.
Veröffentlicht: (2025)
FlowSE: Flow Matching-based Speech Enhancement
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
Layer-aware TDNN: Speaker Recognition Using Multi-Layer Features from Pre-Trained Models
von: Kim, Jin Sob, et al.
Veröffentlicht: (2024)
von: Kim, Jin Sob, et al.
Veröffentlicht: (2024)
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2024)
von: Sang, Mufan, et al.
Veröffentlicht: (2024)
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
von: Chen, Zhengyang, et al.
Veröffentlicht: (2024)
von: Chen, Zhengyang, et al.
Veröffentlicht: (2024)
Speaker Disentanglement of Speech Pre-trained Model Based on Interpretability
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaoxu, et al.
Veröffentlicht: (2025)
Multi-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment
von: Shao, Yiwen, et al.
Veröffentlicht: (2024)
von: Shao, Yiwen, et al.
Veröffentlicht: (2024)
Asymmetric Clean Segments-Guided Self-Supervised Learning for Robust Speaker Verification
von: Gan, Chong-Xin, et al.
Veröffentlicht: (2023)
von: Gan, Chong-Xin, et al.
Veröffentlicht: (2023)
NeXt-TDNN: Modernizing Multi-Scale Temporal Convolution Backbone for Speaker Verification
von: Heo, Hyun-Jun, et al.
Veröffentlicht: (2023)
von: Heo, Hyun-Jun, et al.
Veröffentlicht: (2023)
Speaker-Conditioned Phrase Break Prediction for Text-to-Speech with Phoneme-Level Pre-trained Language Model
von: Yang, Dong, et al.
Veröffentlicht: (2025)
von: Yang, Dong, et al.
Veröffentlicht: (2025)
SA-WavLM: Speaker-Aware Self-Supervised Pre-training for Mixture Speech
von: Lin, Jingru, et al.
Veröffentlicht: (2024)
von: Lin, Jingru, et al.
Veröffentlicht: (2024)
FUN-SSL: Full-band Layer Followed by U-Net with Narrow-band Layers for Multiple Moving Sound Source Localization
von: Choi, Yuseon, et al.
Veröffentlicht: (2025)
von: Choi, Yuseon, et al.
Veröffentlicht: (2025)
Joint Optimization of Speaker and Spoof Detectors for Spoofing-Robust Automatic Speaker Verification
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2025)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2025)
Learning Emotion-Invariant Speaker Representations for Speaker Verification
von: Tian, Jingguang, et al.
Veröffentlicht: (2025)
von: Tian, Jingguang, et al.
Veröffentlicht: (2025)
Investigating the Potential of Multi-Stage Score Fusion in Spoofing-Aware Speaker Verification
von: Kurnaz, Oguzhan, et al.
Veröffentlicht: (2025)
von: Kurnaz, Oguzhan, et al.
Veröffentlicht: (2025)
An Age-Agnostic System for Robust Speaker Verification
von: Zheng, Jiusi, et al.
Veröffentlicht: (2025)
von: Zheng, Jiusi, et al.
Veröffentlicht: (2025)
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
ERes2NetV2: Boosting Short-Duration Speaker Verification Performance with Computational Efficiency
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
UniPET-SPK: A Unified Framework for Parameter-Efficient Tuning of Pre-trained Speech Models for Robust Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2025)
von: Sang, Mufan, et al.
Veröffentlicht: (2025)
Integrated Multi-Level Knowledge Distillation for Enhanced Speaker Verification
von: Yang, Wenhao, et al.
Veröffentlicht: (2024)
von: Yang, Wenhao, et al.
Veröffentlicht: (2024)
Effective Modeling of Critical Contextual Information for TDNN-based Speaker Verification
von: Weng, Shilong, et al.
Veröffentlicht: (2025)
von: Weng, Shilong, et al.
Veröffentlicht: (2025)
Towards Unsupervised Speaker Diarization System for Multilingual Telephone Calls Using Pre-trained Whisper Model and Mixture of Sparse Autoencoders
von: Lam, Phat, et al.
Veröffentlicht: (2024)
von: Lam, Phat, et al.
Veröffentlicht: (2024)
Spoofing-Aware Speaker Verification via Wavelet Prompt Tuning and Multi-Model Ensembles
von: Farhadipour, Aref, et al.
Veröffentlicht: (2026)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2026)
An Adaptive X-vector Model for Text-independent Speaker Verification
von: Gu, Bin, et al.
Veröffentlicht: (2020)
von: Gu, Bin, et al.
Veröffentlicht: (2020)
Trainable Adaptive Score Normalization for Automatic Speaker Verification
von: Choi, Jeong-Hwan, et al.
Veröffentlicht: (2025)
von: Choi, Jeong-Hwan, et al.
Veröffentlicht: (2025)
Optimizing a-DCF for Spoofing-Robust Speaker Verification
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
LG Uplus System with Multi-Speaker IDs and Discriminator-based Sub-Judges for the WildSpoof Challenge
von: Park, Jinyoung, et al.
Veröffentlicht: (2025)
von: Park, Jinyoung, et al.
Veröffentlicht: (2025)
Adversarial Reweighting for Speaker Verification Fairness
von: Jin, Minho, et al.
Veröffentlicht: (2022)
von: Jin, Minho, et al.
Veröffentlicht: (2022)
Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
von: Huang, Wen, et al.
Veröffentlicht: (2024)
von: Huang, Wen, et al.
Veröffentlicht: (2024)
Separate and Reconstruct: Asymmetric Encoder-Decoder for Speech Separation
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2024)
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2024)
Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2025)
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2025)
Whisper-PMFA: Partial Multi-Scale Feature Aggregation for Speaker Verification using Whisper Models
von: Zhao, Yiyang, et al.
Veröffentlicht: (2024)
von: Zhao, Yiyang, et al.
Veröffentlicht: (2024)
Bayesian Learning for Domain-Invariant Speaker Verification and Anti-Spoofing
von: Li, Jin, et al.
Veröffentlicht: (2025)
von: Li, Jin, et al.
Veröffentlicht: (2025)
DAME: Duration-Aware Matryoshka Embedding for Duration-Robust Speaker Verification
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
3D-Speaker-Toolkit: An Open-Source Toolkit for Multimodal Speaker Verification and Diarization
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
Towards Fine-Grained and Multi-Granular Contrastive Language-Speech Pre-training
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
Diffusion-Based Adversarial Purification for Speaker Verification
von: Bai, Yibo, et al.
Veröffentlicht: (2023)
von: Bai, Yibo, et al.
Veröffentlicht: (2023)
PhiNet: Speaker Verification with Phonetic Interpretability
von: Ma, Yi, et al.
Veröffentlicht: (2026)
von: Ma, Yi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Speech Enhancement based on cascaded two flows
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025) -
SV-Mixer: Replacing the Transformer Encoder with Lightweight MLPs for Self-Supervised Model Compression in Speaker Verification
von: Heo, Jungwoo, et al.
Veröffentlicht: (2025) -
FlowSE: Flow Matching-based Speech Enhancement
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025) -
Layer-aware TDNN: Speaker Recognition Using Multi-Layer Features from Pre-Trained Models
von: Kim, Jin Sob, et al.
Veröffentlicht: (2024) -
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
von: Sang, Mufan, et al.
Veröffentlicht: (2024)