Towards Evaluating the Robustness of Automatic Speech Recognition Systems via Audio Style Transfer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Weifei, Cao, Yuxin, Su, Junjie, Shen, Qi, Ye, Kai, Wang, Derui, Hao, Jie, Liu, Ziyao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
von: Jin, Weifei, et al.
Veröffentlicht: (2025)
von: Jin, Weifei, et al.
Veröffentlicht: (2025)
Zero-Query Adversarial Attack on Black-box Automatic Speech Recognition Systems
von: Fang, Zheng, et al.
Veröffentlicht: (2024)
von: Fang, Zheng, et al.
Veröffentlicht: (2024)
Adversarial Attacks and Defenses for Speech Recognition Systems
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021)
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021)
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
von: Huang, Kunyang, et al.
Veröffentlicht: (2025)
von: Huang, Kunyang, et al.
Veröffentlicht: (2025)
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)
Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation
von: Zhang, Kuiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Kuiyuan, et al.
Veröffentlicht: (2024)
Two Views, One Truth: Spectral and Self-Supervised Features Fusion for Robust Speech Deepfake Detection
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
Mirage Fools the Ear, Mute Hides the Truth: Precise Targeted Adversarial Attacks on Polyphonic Sound Event Detection Systems
von: Su, Junjie, et al.
Veröffentlicht: (2025)
von: Su, Junjie, et al.
Veröffentlicht: (2025)
SilentCipher: Deep Audio Watermarking
von: Singh, Mayank Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Mayank Kumar, et al.
Veröffentlicht: (2024)
Privacy in Speech Technology
von: Bäckström, Tom
Veröffentlicht: (2023)
von: Bäckström, Tom
Veröffentlicht: (2023)
AudioMarkBench: Benchmarking Robustness of Audio Watermarking
von: Liu, Hongbin, et al.
Veröffentlicht: (2024)
von: Liu, Hongbin, et al.
Veröffentlicht: (2024)
Yours or Mine? Overwriting Attacks Against Neural Audio Watermarking
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)
Interpretable Temporal Class Activation Representation for Audio Spoofing Detection
von: Li, Menglu, et al.
Veröffentlicht: (2024)
von: Li, Menglu, et al.
Veröffentlicht: (2024)
MerkleSpeech: Public-Key Verifiable, Chunk-Localised Speech Provenance via Perceptual Fingerprints and Merkle Commitments
von: Ono, Tatsunori
Veröffentlicht: (2026)
von: Ono, Tatsunori
Veröffentlicht: (2026)
Breaking Speaker Recognition with PaddingBack
von: Ye, Zhe, et al.
Veröffentlicht: (2023)
von: Ye, Zhe, et al.
Veröffentlicht: (2023)
Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis
von: Warren, Kevin, et al.
Veröffentlicht: (2025)
von: Warren, Kevin, et al.
Veröffentlicht: (2025)
One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection
von: Kim, Hyun Myung, et al.
Veröffentlicht: (2024)
von: Kim, Hyun Myung, et al.
Veröffentlicht: (2024)
Lightweight Protection for Privacy in Offloaded Speech Understanding
von: Cai, Dongqi
Veröffentlicht: (2024)
von: Cai, Dongqi
Veröffentlicht: (2024)
A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection
von: Liu, Xuechen, et al.
Veröffentlicht: (2024)
von: Liu, Xuechen, et al.
Veröffentlicht: (2024)
PITCH: AI-assisted Tagging of Deepfake Audio Calls using Challenge-Response
von: Mittal, Govind, et al.
Veröffentlicht: (2024)
von: Mittal, Govind, et al.
Veröffentlicht: (2024)
A Universal Identity Backdoor Attack against Speaker Verification based on Siamese Network
von: Zhao, Haodong, et al.
Veröffentlicht: (2023)
von: Zhao, Haodong, et al.
Veröffentlicht: (2023)
TriniMark: A Robust Generative Speech Watermarking Method for Trinity-Level Traceability
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
Frame-level Temporal Difference Learning for Partial Deepfake Speech Detection
von: Li, Menglu, et al.
Veröffentlicht: (2025)
von: Li, Menglu, et al.
Veröffentlicht: (2025)
Gumbel Rao Monte Carlo based Bi-Modal Neural Architecture Search for Audio-Visual Deepfake Detection
von: PN, Aravinda Reddy, et al.
Veröffentlicht: (2024)
von: PN, Aravinda Reddy, et al.
Veröffentlicht: (2024)
The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents
von: Wang, Lixu, et al.
Veröffentlicht: (2025)
von: Wang, Lixu, et al.
Veröffentlicht: (2025)
WaLi: Can Pressure Sensors in HVAC Systems Capture Human Speech?
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
Why Speech Deepfake Detectors Won't Generalize: The Limits of Detection in an Open World
von: Berisha, Visar, et al.
Veröffentlicht: (2025)
von: Berisha, Visar, et al.
Veröffentlicht: (2025)
Adversarial Representation Learning for Robust Privacy Preservation in Audio
von: Gharib, Shayan, et al.
Veröffentlicht: (2023)
von: Gharib, Shayan, et al.
Veröffentlicht: (2023)
Cross-Technology Generalization in Synthesized Speech Detection: Evaluating AST Models with Modern Voice Generators
von: Ustinov, Andrew, et al.
Veröffentlicht: (2025)
von: Ustinov, Andrew, et al.
Veröffentlicht: (2025)
Speech privacy-preserving methods using secret key for convolutional neural network models and their robustness evaluation
von: Niwa, Shoko, et al.
Veröffentlicht: (2024)
von: Niwa, Shoko, et al.
Veröffentlicht: (2024)
An Effective Energy Mask-based Adversarial Evasion Attacks against Misclassification in Speaker Recognition Systems
von: Park, Chanwoo, et al.
Veröffentlicht: (2026)
von: Park, Chanwoo, et al.
Veröffentlicht: (2026)
DIN-CTS: Low-Complexity Depthwise-Inception Neural Network with Contrastive Training Strategy for Deepfake Speech Detection
von: Pham, Lam, et al.
Veröffentlicht: (2025)
von: Pham, Lam, et al.
Veröffentlicht: (2025)
GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis
von: Liu, Weizhi, et al.
Veröffentlicht: (2024)
von: Liu, Weizhi, et al.
Veröffentlicht: (2024)
ALIF: Low-Cost Adversarial Audio Attacks on Black-Box Speech Platforms using Linguistic Features
von: Cheng, Peng, et al.
Veröffentlicht: (2024)
von: Cheng, Peng, et al.
Veröffentlicht: (2024)
A Systematic Evaluation of Adversarial Attacks against Speech Emotion Recognition Models
von: Facchinetti, Nicolas, et al.
Veröffentlicht: (2024)
von: Facchinetti, Nicolas, et al.
Veröffentlicht: (2024)
Quantized Approximate Signal Processing (QASP): Towards Homomorphic Encryption for audio
von: Nguyen, Tu Duyen, et al.
Veröffentlicht: (2025)
von: Nguyen, Tu Duyen, et al.
Veröffentlicht: (2025)
CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning
von: Wu, Haolin, et al.
Veröffentlicht: (2024)
von: Wu, Haolin, et al.
Veröffentlicht: (2024)
Your Microphone Array Retains Your Identity: A Robust Voice Liveness Detection System for Smart Speakers
von: Meng, Yan, et al.
Veröffentlicht: (2025)
von: Meng, Yan, et al.
Veröffentlicht: (2025)
LCANets++: Robust Audio Classification using Multi-layer Neural Networks with Lateral Competition
von: Dibbo, Sayanton V., et al.
Veröffentlicht: (2023)
von: Dibbo, Sayanton V., et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
von: Jin, Weifei, et al.
Veröffentlicht: (2025) -
Zero-Query Adversarial Attack on Black-box Automatic Speech Recognition Systems
von: Fang, Zheng, et al.
Veröffentlicht: (2024) -
Adversarial Attacks and Defenses for Speech Recognition Systems
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021) -
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
von: Huang, Kunyang, et al.
Veröffentlicht: (2025) -
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)