Continual Speaker Identity Unlearning with Minimal Interference
Fuente:
arXiv
Guardado en:
| Autores principales: | Kim, Jinju, Kang, Yunsung, Park, Gyeong-Moon, Ko, Jong Hwan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do Not Mimic My Voice: Speaker Identity Unlearning for Zero-Shot Text-to-Speech
por: Kim, Taesoo, et al.
Publicado: (2025)
por: Kim, Taesoo, et al.
Publicado: (2025)
Mask2Flow-TSE: Two-Stage Target Speaker Extraction with Masking and Flow Matching
por: Moon, Junwon, et al.
Publicado: (2026)
por: Moon, Junwon, et al.
Publicado: (2026)
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment
por: Kim, Taesoo, et al.
Publicado: (2025)
por: Kim, Taesoo, et al.
Publicado: (2025)
Evaluating Identity Leakage in Speaker De-Identification Systems
por: Seo, Seungmin, et al.
Publicado: (2025)
por: Seo, Seungmin, et al.
Publicado: (2025)
SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection
por: Jung, Kyudan, et al.
Publicado: (2026)
por: Jung, Kyudan, et al.
Publicado: (2026)
Disentangling Age and Identity with a Mutual Information Minimization Approach for Cross-Age Speaker Verification
por: Zhang, Fengrun, et al.
Publicado: (2024)
por: Zhang, Fengrun, et al.
Publicado: (2024)
Rethinking Leveraging Pre-Trained Multi-Layer Representations for Speaker Verification
por: Kim, Jin Sob, et al.
Publicado: (2025)
por: Kim, Jin Sob, et al.
Publicado: (2025)
Evaluating Speaker Identity Coding in Self-supervised Models and Humans
por: Elbanna, Gasser
Publicado: (2024)
por: Elbanna, Gasser
Publicado: (2024)
Robust Target Speaker Diarization and Separation via Augmented Speaker Embedding Sampling
por: Jalal, Md Asif, et al.
Publicado: (2025)
por: Jalal, Md Asif, et al.
Publicado: (2025)
Generative Unlearning for Any Identity
por: Seo, Juwon, et al.
Publicado: (2024)
por: Seo, Juwon, et al.
Publicado: (2024)
SpeakerLM: End-to-End Versatile Speaker Diarization and Recognition with Multimodal Large Language Models
por: Yin, Han, et al.
Publicado: (2025)
por: Yin, Han, et al.
Publicado: (2025)
A Novel Automatic Framework for Speaker Drift Detection in Synthesized Speech
por: Huang, Jia-Hong, et al.
Publicado: (2026)
por: Huang, Jia-Hong, et al.
Publicado: (2026)
Exploring Speaker Diarization with Mixture of Experts
por: Yang, Gaobin, et al.
Publicado: (2025)
por: Yang, Gaobin, et al.
Publicado: (2025)
Stable-TTS: Stable Speaker-Adaptive Text-to-Speech Synthesis via Prosody Prompting
por: Han, Wooseok, et al.
Publicado: (2024)
por: Han, Wooseok, et al.
Publicado: (2024)
Disentangling Speakers in Multi-Talker Speech Recognition with Speaker-Aware CTC
por: Kang, Jiawen, et al.
Publicado: (2024)
por: Kang, Jiawen, et al.
Publicado: (2024)
Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model
por: Peng, Shuhai, et al.
Publicado: (2026)
por: Peng, Shuhai, et al.
Publicado: (2026)
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation
por: Kim, Ji-Hoon, et al.
Publicado: (2024)
por: Kim, Ji-Hoon, et al.
Publicado: (2024)
Speaker Verification with Speech-Aware LLMs: Evaluation and Augmentation
por: Thebaud, Thomas, et al.
Publicado: (2026)
por: Thebaud, Thomas, et al.
Publicado: (2026)
Probabilistic Fusion and Calibration of Neural Speaker Diarization Models
por: Alvarez-Trejos, Juan Ignacio, et al.
Publicado: (2025)
por: Alvarez-Trejos, Juan Ignacio, et al.
Publicado: (2025)
Hyperbolic Additive Margin Softmax with Hierarchical Information for Speaker Verification
por: Fang, Zhihua, et al.
Publicado: (2026)
por: Fang, Zhihua, et al.
Publicado: (2026)
Neural Multi-Speaker Voice Cloning for Nepali in Low-Resource Settings
por: Shrestha, Aayush M., et al.
Publicado: (2026)
por: Shrestha, Aayush M., et al.
Publicado: (2026)
End-to-End Multi-Microphone Speaker Extraction Using Relative Transfer Functions
por: Eisenberg, Aviad, et al.
Publicado: (2025)
por: Eisenberg, Aviad, et al.
Publicado: (2025)
Perturb a Model, Not an Image: Towards Robust Privacy Protection via Anti-Personalized Diffusion Models
por: Lee, Tae-Young, et al.
Publicado: (2025)
por: Lee, Tae-Young, et al.
Publicado: (2025)
An Investigation Into Various Approaches For Bengali Long-Form Speech Transcription and Bengali Speaker Diarization
por: Jahan, Epshita, et al.
Publicado: (2026)
por: Jahan, Epshita, et al.
Publicado: (2026)
TFGA-Net: Temporal-Frequency Graph Attention Network for Brain-Controlled Speaker Extraction
por: Si, Youhao, et al.
Publicado: (2025)
por: Si, Youhao, et al.
Publicado: (2025)
MultiActor-Audiobook: Zero-Shot Audiobook Generation with Faces and Voices of Multiple Speakers
por: Park, Kyeongman, et al.
Publicado: (2025)
por: Park, Kyeongman, et al.
Publicado: (2025)
A Holistic Framework for Robust Bangla ASR and Speaker Diarization with Optimized VAD and CTC Alignment
por: Ishmam, Zarif, et al.
Publicado: (2026)
por: Ishmam, Zarif, et al.
Publicado: (2026)
Speaker Embeddings to Improve Tracking of Intermittent and Moving Speakers
por: Iatariene, Taous, et al.
Publicado: (2025)
por: Iatariene, Taous, et al.
Publicado: (2025)
Eta-WavLM: Efficient Speaker Identity Removal in Self-Supervised Speech Representations Using a Simple Linear Equation
por: Ruggiero, Giuseppe, et al.
Publicado: (2025)
por: Ruggiero, Giuseppe, et al.
Publicado: (2025)
ERIS: Evolutionary Real-world Interference Scheme for Jailbreaking Audio Large Models
por: Zhang, Yibo, et al.
Publicado: (2025)
por: Zhang, Yibo, et al.
Publicado: (2025)
AlphaFlowTSE: One-Step Generative Target Speaker Extraction via Conditional AlphaFlow
por: Li, Duojia, et al.
Publicado: (2026)
por: Li, Duojia, et al.
Publicado: (2026)
Memory-Efficient Training for Deep Speaker Embedding Learning in Speaker Verification
por: Liu, Bei, et al.
Publicado: (2024)
por: Liu, Bei, et al.
Publicado: (2024)
Stage-Adaptive Reliability Modeling for Continuous Valence-Arousal Estimation
por: Lee, Yubeen, et al.
Publicado: (2026)
por: Lee, Yubeen, et al.
Publicado: (2026)
Leveraging Speaker Embeddings in End-to-End Neural Diarization for Two-Speaker Scenarios
por: Alvarez-Trejos, Juan Ignacio, et al.
Publicado: (2024)
por: Alvarez-Trejos, Juan Ignacio, et al.
Publicado: (2024)
Subject-Independent Imagined Speech Detection via Cross-Subject Generalization and Calibration
por: Ko, Byung-Kwan, et al.
Publicado: (2025)
por: Ko, Byung-Kwan, et al.
Publicado: (2025)
Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought
por: Li, Xuanchen, et al.
Publicado: (2026)
por: Li, Xuanchen, et al.
Publicado: (2026)
Raon-Speech Technical Report
por: Kim, Beomsoo, et al.
Publicado: (2026)
por: Kim, Beomsoo, et al.
Publicado: (2026)
Towards Speaker Identification with Minimal Dataset and Constrained Resources using 1D-Convolution Neural Network
por: Shahan, Irfan Nafiz, et al.
Publicado: (2024)
por: Shahan, Irfan Nafiz, et al.
Publicado: (2024)
Explainable Attribute-Based Speaker Verification
por: Wu, Xiaoliang, et al.
Publicado: (2024)
por: Wu, Xiaoliang, et al.
Publicado: (2024)
SALF-MOS: Speaker Agnostic Latent Features Downsampled for MOS Prediction
por: Agrawal, Saurabh, et al.
Publicado: (2025)
por: Agrawal, Saurabh, et al.
Publicado: (2025)
Ejemplares similares
-
Do Not Mimic My Voice: Speaker Identity Unlearning for Zero-Shot Text-to-Speech
por: Kim, Taesoo, et al.
Publicado: (2025) -
Mask2Flow-TSE: Two-Stage Target Speaker Extraction with Masking and Flow Matching
por: Moon, Junwon, et al.
Publicado: (2026) -
TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment
por: Kim, Taesoo, et al.
Publicado: (2025) -
Evaluating Identity Leakage in Speaker De-Identification Systems
por: Seo, Seungmin, et al.
Publicado: (2025) -
SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection
por: Jung, Kyudan, et al.
Publicado: (2026)