AudioJailbreak: Jailbreak Attacks against End-to-End Large Audio-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Guangke, Song, Fu, Zhao, Zhe, Jia, Xiaojun, Liu, Yang, Qiao, Yanchen, Zhang, Weizhe, Tu, Weiping, Yang, Yuhong, Du, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AdvWave: Stealthy Adversarial Jailbreak Attack against Large Audio-Language Models
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
by: Peng, Zifan, et al.
Published: (2025)
by: Peng, Zifan, et al.
Published: (2025)
Multilingual and Multi-Accent Jailbreaking of Audio LLMs
by: Roh, Jaechul, et al.
Published: (2025)
by: Roh, Jaechul, et al.
Published: (2025)
When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs
by: Dingeto, Hiskias, et al.
Published: (2025)
by: Dingeto, Hiskias, et al.
Published: (2025)
Yours or Mine? Overwriting Attacks Against Neural Audio Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
Synthetic Voices, Real Threats: Evaluating Large Text-to-Speech Models in Generating Harmful Audio
by: Chen, Guangke, et al.
Published: (2025)
by: Chen, Guangke, et al.
Published: (2025)
SilentCipher: Deep Audio Watermarking
by: Singh, Mayank Kumar, et al.
Published: (2024)
by: Singh, Mayank Kumar, et al.
Published: (2024)
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
by: Huang, Kunyang, et al.
Published: (2025)
by: Huang, Kunyang, et al.
Published: (2025)
A Universal Identity Backdoor Attack against Speaker Verification based on Siamese Network
by: Zhao, Haodong, et al.
Published: (2023)
by: Zhao, Haodong, et al.
Published: (2023)
CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning
by: Wu, Haolin, et al.
Published: (2024)
by: Wu, Haolin, et al.
Published: (2024)
The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents
by: Wang, Lixu, et al.
Published: (2025)
by: Wang, Lixu, et al.
Published: (2025)
Interpretable Temporal Class Activation Representation for Audio Spoofing Detection
by: Li, Menglu, et al.
Published: (2024)
by: Li, Menglu, et al.
Published: (2024)
SongBsAb: A Dual Prevention Approach against Singing Voice Conversion based Illegal Song Covers
by: Chen, Guangke, et al.
Published: (2024)
by: Chen, Guangke, et al.
Published: (2024)
Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis
by: Warren, Kevin, et al.
Published: (2025)
by: Warren, Kevin, et al.
Published: (2025)
One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection
by: Kim, Hyun Myung, et al.
Published: (2024)
by: Kim, Hyun Myung, et al.
Published: (2024)
A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection
by: Liu, Xuechen, et al.
Published: (2024)
by: Liu, Xuechen, et al.
Published: (2024)
PITCH: AI-assisted Tagging of Deepfake Audio Calls using Challenge-Response
by: Mittal, Govind, et al.
Published: (2024)
by: Mittal, Govind, et al.
Published: (2024)
IO-RAE: Information-Obfuscation Reversible Adversarial Example for Audio Privacy Protection
by: Zhu, Jiajie, et al.
Published: (2026)
by: Zhu, Jiajie, et al.
Published: (2026)
AudioMarkBench: Benchmarking Robustness of Audio Watermarking
by: Liu, Hongbin, et al.
Published: (2024)
by: Liu, Hongbin, et al.
Published: (2024)
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
Gumbel Rao Monte Carlo based Bi-Modal Neural Architecture Search for Audio-Visual Deepfake Detection
by: PN, Aravinda Reddy, et al.
Published: (2024)
by: PN, Aravinda Reddy, et al.
Published: (2024)
An Effective Energy Mask-based Adversarial Evasion Attacks against Misclassification in Speaker Recognition Systems
by: Park, Chanwoo, et al.
Published: (2026)
by: Park, Chanwoo, et al.
Published: (2026)
ALIF: Low-Cost Adversarial Audio Attacks on Black-Box Speech Platforms using Linguistic Features
by: Cheng, Peng, et al.
Published: (2024)
by: Cheng, Peng, et al.
Published: (2024)
EveGuard: Defeating Vibration-based Side-Channel Eavesdropping with Audio Adversarial Perturbations
by: Chang, Jung-Woo, et al.
Published: (2024)
by: Chang, Jung-Woo, et al.
Published: (2024)
DeePen: Penetration Testing for Audio Deepfake Detection
by: Müller, Nicolas, et al.
Published: (2025)
by: Müller, Nicolas, et al.
Published: (2025)
Adversarial Representation Learning for Robust Privacy Preservation in Audio
by: Gharib, Shayan, et al.
Published: (2023)
by: Gharib, Shayan, et al.
Published: (2023)
PosCUDA: Position based Convolution for Unlearnable Audio Datasets
by: Gokul, Vignesh, et al.
Published: (2024)
by: Gokul, Vignesh, et al.
Published: (2024)
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
by: Jin, Weifei, et al.
Published: (2025)
by: Jin, Weifei, et al.
Published: (2025)
GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis
by: Liu, Weizhi, et al.
Published: (2024)
by: Liu, Weizhi, et al.
Published: (2024)
Inference Attacks for X-Vector Speaker Anonymization
by: Bauer, Luke, et al.
Published: (2025)
by: Bauer, Luke, et al.
Published: (2025)
Adversarial Attacks and Defenses for Speech Recognition Systems
by: Żelasko, Piotr, et al.
Published: (2021)
by: Żelasko, Piotr, et al.
Published: (2021)
Evaluating Synthetic Command Attacks on Smart Voice Assistants
by: He, Zhengxian, et al.
Published: (2024)
by: He, Zhengxian, et al.
Published: (2024)
Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World
by: Sadasivan, Vinu Sankar, et al.
Published: (2025)
by: Sadasivan, Vinu Sankar, et al.
Published: (2025)
Audio Jailbreak: An Open Comprehensive Benchmark for Jailbreaking Large Audio-Language Models
by: Song, Zirui, et al.
Published: (2025)
by: Song, Zirui, et al.
Published: (2025)
Hidden in Plain Sound: Environmental Backdoor Poisoning Attacks on Whisper, and Mitigations
by: Bartolini, Jonatan, et al.
Published: (2024)
by: Bartolini, Jonatan, et al.
Published: (2024)
LCANets++: Robust Audio Classification using Multi-layer Neural Networks with Lateral Competition
by: Dibbo, Sayanton V., et al.
Published: (2023)
by: Dibbo, Sayanton V., et al.
Published: (2023)
Representation Learning for Audio Privacy Preservation using Source Separation and Robust Adversarial Learning
by: Luong, Diep, et al.
Published: (2023)
by: Luong, Diep, et al.
Published: (2023)
Towards Evaluating the Robustness of Automatic Speech Recognition Systems via Audio Style Transfer
by: Jin, Weifei, et al.
Published: (2024)
by: Jin, Weifei, et al.
Published: (2024)
Towards the Development of a Real-Time Deepfake Audio Detection System in Communication Platforms
by: Mathew, Jonat John, et al.
Published: (2024)
by: Mathew, Jonat John, et al.
Published: (2024)
Zero-Query Adversarial Attack on Black-box Automatic Speech Recognition Systems
by: Fang, Zheng, et al.
Published: (2024)
by: Fang, Zheng, et al.
Published: (2024)
Similar Items
-
AdvWave: Stealthy Adversarial Jailbreak Attack against Large Audio-Language Models
by: Kang, Mintong, et al.
Published: (2024) -
JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
by: Peng, Zifan, et al.
Published: (2025) -
Multilingual and Multi-Accent Jailbreaking of Audio LLMs
by: Roh, Jaechul, et al.
Published: (2025) -
When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs
by: Dingeto, Hiskias, et al.
Published: (2025) -
Yours or Mine? Overwriting Attacks Against Neural Audio Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)