Dimensionality-Aware Anomaly Detection in Learned Representations of Self-Supervised Speech Models
Fuente:
arXiv
Saved in:
| Main Authors: | Arcos-Holzinger, Sandra, Erfani, Sarah M., Bailey, James, Khudanpur, Sanjeev |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adversarial Attacks and Defenses for Speech Recognition Systems
by: Żelasko, Piotr, et al.
Published: (2021)
by: Żelasko, Piotr, et al.
Published: (2021)
Clean Label Attacks against SLU Systems
by: Xinyuan, Henry Li, et al.
Published: (2024)
by: Xinyuan, Henry Li, et al.
Published: (2024)
Two Views, One Truth: Spectral and Self-Supervised Features Fusion for Robust Speech Deepfake Detection
by: Kheir, Yassine El, et al.
Published: (2025)
by: Kheir, Yassine El, et al.
Published: (2025)
Lightweight Model Attribution and Detection of Synthetic Speech via Audio Residual Fingerprints
by: Pizarro, Matías, et al.
Published: (2024)
by: Pizarro, Matías, et al.
Published: (2024)
Parameter-Efficient Transfer Learning under Federated Learning for Automatic Speech Recognition
by: Kan, Xuan, et al.
Published: (2024)
by: Kan, Xuan, et al.
Published: (2024)
Frame-level Temporal Difference Learning for Partial Deepfake Speech Detection
by: Li, Menglu, et al.
Published: (2025)
by: Li, Menglu, et al.
Published: (2025)
Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation
by: Zhang, Kuiyuan, et al.
Published: (2024)
by: Zhang, Kuiyuan, et al.
Published: (2024)
Cross-Technology Generalization in Synthesized Speech Detection: Evaluating AST Models with Modern Voice Generators
by: Ustinov, Andrew, et al.
Published: (2025)
by: Ustinov, Andrew, et al.
Published: (2025)
Adversarial Representation Learning for Robust Privacy Preservation in Audio
by: Gharib, Shayan, et al.
Published: (2023)
by: Gharib, Shayan, et al.
Published: (2023)
A Systematic Evaluation of Adversarial Attacks against Speech Emotion Recognition Models
by: Facchinetti, Nicolas, et al.
Published: (2024)
by: Facchinetti, Nicolas, et al.
Published: (2024)
Representation Learning for Audio Privacy Preservation using Source Separation and Robust Adversarial Learning
by: Luong, Diep, et al.
Published: (2023)
by: Luong, Diep, et al.
Published: (2023)
LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech
by: Liu, Xuechen, et al.
Published: (2025)
by: Liu, Xuechen, et al.
Published: (2025)
Why Speech Deepfake Detectors Won't Generalize: The Limits of Detection in an Open World
by: Berisha, Visar, et al.
Published: (2025)
by: Berisha, Visar, et al.
Published: (2025)
Exploratory Evaluation of Speech Content Masking
by: Williams, Jennifer, et al.
Published: (2024)
by: Williams, Jennifer, et al.
Published: (2024)
Interpretable Temporal Class Activation Representation for Audio Spoofing Detection
by: Li, Menglu, et al.
Published: (2024)
by: Li, Menglu, et al.
Published: (2024)
A Survey on Speech Deepfake Detection
by: Li, Menglu, et al.
Published: (2024)
by: Li, Menglu, et al.
Published: (2024)
Towards Generalized Source Tracing for Codec-Based Deepfake Speech
by: Chen, Xuanjun, et al.
Published: (2025)
by: Chen, Xuanjun, et al.
Published: (2025)
DIN-CTS: Low-Complexity Depthwise-Inception Neural Network with Contrastive Training Strategy for Deepfake Speech Detection
by: Pham, Lam, et al.
Published: (2025)
by: Pham, Lam, et al.
Published: (2025)
Privacy in Speech Technology
by: Bäckström, Tom
Published: (2023)
by: Bäckström, Tom
Published: (2023)
CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning
by: Wu, Haolin, et al.
Published: (2024)
by: Wu, Haolin, et al.
Published: (2024)
Towards Evaluating the Robustness of Automatic Speech Recognition Systems via Audio Style Transfer
by: Jin, Weifei, et al.
Published: (2024)
by: Jin, Weifei, et al.
Published: (2024)
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
by: Huang, Kunyang, et al.
Published: (2025)
by: Huang, Kunyang, et al.
Published: (2025)
MerkleSpeech: Public-Key Verifiable, Chunk-Localised Speech Provenance via Perceptual Fingerprints and Merkle Commitments
by: Ono, Tatsunori
Published: (2026)
by: Ono, Tatsunori
Published: (2026)
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
by: Jin, Weifei, et al.
Published: (2025)
by: Jin, Weifei, et al.
Published: (2025)
Unraveling Adversarial Examples against Speaker Identification -- Techniques for Attack Detection and Victim Model Classification
by: Joshi, Sonal, et al.
Published: (2024)
by: Joshi, Sonal, et al.
Published: (2024)
Lightweight Protection for Privacy in Offloaded Speech Understanding
by: Cai, Dongqi
Published: (2024)
by: Cai, Dongqi
Published: (2024)
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
by: Neri, Michael, et al.
Published: (2025)
by: Neri, Michael, et al.
Published: (2025)
One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection
by: Kim, Hyun Myung, et al.
Published: (2024)
by: Kim, Hyun Myung, et al.
Published: (2024)
Every Breath You Don't Take: Deepfake Speech Detection Using Breath
by: Layton, Seth, et al.
Published: (2024)
by: Layton, Seth, et al.
Published: (2024)
WaLi: Can Pressure Sensors in HVAC Systems Capture Human Speech?
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
Zero-Query Adversarial Attack on Black-box Automatic Speech Recognition Systems
by: Fang, Zheng, et al.
Published: (2024)
by: Fang, Zheng, et al.
Published: (2024)
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
Efficient Streaming Voice Steganalysis in Challenging Detection Scenarios
by: Zhou, Pengcheng, et al.
Published: (2024)
by: Zhou, Pengcheng, et al.
Published: (2024)
The Impact of Speech Anonymization on Pathology and Its Limits
by: Arasteh, Soroosh Tayebi, et al.
Published: (2024)
by: Arasteh, Soroosh Tayebi, et al.
Published: (2024)
Speech privacy-preserving methods using secret key for convolutional neural network models and their robustness evaluation
by: Niwa, Shoko, et al.
Published: (2024)
by: Niwa, Shoko, et al.
Published: (2024)
Privacy-Preserving End-to-End Spoken Language Understanding
by: Wang, Yinggui, et al.
Published: (2024)
by: Wang, Yinggui, et al.
Published: (2024)
Precision-Varying Prediction (PVP): Robustifying ASR systems against adversarial attacks
by: Pizarro, Matías, et al.
Published: (2026)
by: Pizarro, Matías, et al.
Published: (2026)
Towards the Development of a Real-Time Deepfake Audio Detection System in Communication Platforms
by: Mathew, Jonat John, et al.
Published: (2024)
by: Mathew, Jonat John, et al.
Published: (2024)
ClaritySpeech: Dementia Obfuscation in Speech
by: Woszczyk, Dominika, et al.
Published: (2025)
by: Woszczyk, Dominika, et al.
Published: (2025)
Improving Membership Inference in ASR Model Auditing with Perturbed Loss Features
by: Teixeira, Francisco, et al.
Published: (2024)
by: Teixeira, Francisco, et al.
Published: (2024)
Similar Items
-
Adversarial Attacks and Defenses for Speech Recognition Systems
by: Żelasko, Piotr, et al.
Published: (2021) -
Clean Label Attacks against SLU Systems
by: Xinyuan, Henry Li, et al.
Published: (2024) -
Two Views, One Truth: Spectral and Self-Supervised Features Fusion for Robust Speech Deepfake Detection
by: Kheir, Yassine El, et al.
Published: (2025) -
Lightweight Model Attribution and Detection of Synthetic Speech via Audio Residual Fingerprints
by: Pizarro, Matías, et al.
Published: (2024) -
Parameter-Efficient Transfer Learning under Federated Learning for Automatic Speech Recognition
by: Kan, Xuan, et al.
Published: (2024)