Sirens' Whisper: Inaudible Near-Ultrasonic Jailbreaks of Speech-Driven LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Ling, Zijian, Hu, Pingyi, Gao, Xiuyong, Ma, Xiaojing, Zhou, Man, Feng, Jun, Lu, Songfeng, Zhang, Dongmei, Zhu, Bin Benjamin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Whispering Under the Eaves: Protecting User Privacy Against Commercial and LLM-powered Automatic Speech Recognition Systems
by: Jin, Weifei, et al.
Published: (2025)
by: Jin, Weifei, et al.
Published: (2025)
Acoustic Interference: A New Paradigm Weaponizing Acoustic Latent Semantic for Universal Jailbreak against Large Audio Language Models
by: Wang, Yanyun, et al.
Published: (2026)
by: Wang, Yanyun, et al.
Published: (2026)
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
by: Huang, Kunyang, et al.
Published: (2025)
by: Huang, Kunyang, et al.
Published: (2025)
Selective Masking Adversarial Attack on Automatic Speech Recognition Systems
by: Fang, Zheng, et al.
Published: (2025)
by: Fang, Zheng, et al.
Published: (2025)
HVAC-EAR: Eavesdropping Human Speech Using HVAC Systems
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
Vulnerabilities of Audio-Based Biometric Authentication Systems Against Deepfake Speech Synthesis
by: Hong, Mengze, et al.
Published: (2026)
by: Hong, Mengze, et al.
Published: (2026)
Hidden in Plain Sound: Environmental Backdoor Poisoning Attacks on Whisper, and Mitigations
by: Bartolini, Jonatan, et al.
Published: (2024)
by: Bartolini, Jonatan, et al.
Published: (2024)
Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs
by: Roh, Jaechul, et al.
Published: (2026)
by: Roh, Jaechul, et al.
Published: (2026)
MelShield: Robust Mel-Domain Audio Watermarking for Provenance Attribution of AI Generated Synthesized Speech
by: Jin, Yutong, et al.
Published: (2026)
by: Jin, Yutong, et al.
Published: (2026)
When Fine-Tuning is Not Enough: Lessons from HSAD on Hybrid and Adversarial Audio Spoof Detection
by: Hu, Bin, et al.
Published: (2025)
by: Hu, Bin, et al.
Published: (2025)
Multilingual and Multi-Accent Jailbreaking of Audio LLMs
by: Roh, Jaechul, et al.
Published: (2025)
by: Roh, Jaechul, et al.
Published: (2025)
Whisper Smarter, not Harder: Adversarial Attack on Partial Suppression
by: Wong, Zheng Jie, et al.
Published: (2025)
by: Wong, Zheng Jie, et al.
Published: (2025)
Can DeepFake Speech be Reliably Detected?
by: Liu, Hongbin, et al.
Published: (2024)
by: Liu, Hongbin, et al.
Published: (2024)
SCDF: A Speaker Characteristics DeepFake Speech Dataset for Bias Analysis
by: Staněk, Vojtěch, et al.
Published: (2025)
by: Staněk, Vojtěch, et al.
Published: (2025)
Decoding Deception: Understanding Automatic Speech Recognition Vulnerabilities in Evasion and Poisoning Attacks
by: G, Aravindhan, et al.
Published: (2025)
by: G, Aravindhan, et al.
Published: (2025)
Privacy in Speech Technology
by: Bäckström, Tom
Published: (2023)
by: Bäckström, Tom
Published: (2023)
Smark: A Watermark for Text-to-Speech Diffusion Models via Discrete Wavelet Transform
by: Zhang, Yichuan, et al.
Published: (2025)
by: Zhang, Yichuan, et al.
Published: (2025)
SOLIDO: A Robust Watermarking Method for Speech Synthesis via Low-Rank Adaptation
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
Beyond Content: A Comprehensive Speech Toxicity Dataset and Detection Framework Incorporating Paralinguistic Cues
by: Ba, Zhongjie, et al.
Published: (2026)
by: Ba, Zhongjie, et al.
Published: (2026)
Where Do Backdoors Live? A Component-Level Analysis of Backdoor Propagation in Speech Language Models
by: Fortier, Alexandrine, et al.
Published: (2025)
by: Fortier, Alexandrine, et al.
Published: (2025)
MerkleSpeech: Public-Key Verifiable, Chunk-Localised Speech Provenance via Perceptual Fingerprints and Merkle Commitments
by: Ono, Tatsunori
Published: (2026)
by: Ono, Tatsunori
Published: (2026)
Invisible Ears at Your Fingertips: Acoustic Eavesdropping via Mouse Sensors
by: Fakih, Mohamad, et al.
Published: (2025)
by: Fakih, Mohamad, et al.
Published: (2025)
DECKER: Domain-invariant Embedding for Cross-Keyboard Extraction and Recognition
by: Maurya, Bikrant Bikram Pratap, et al.
Published: (2026)
by: Maurya, Bikrant Bikram Pratap, et al.
Published: (2026)
Mirage Fools the Ear, Mute Hides the Truth: Precise Targeted Adversarial Attacks on Polyphonic Sound Event Detection Systems
by: Su, Junjie, et al.
Published: (2025)
by: Su, Junjie, et al.
Published: (2025)
Audio Pirates: Black-box Audio Watermark Removal via Diffusion Priors
by: Yao, Lingfeng, et al.
Published: (2026)
by: Yao, Lingfeng, et al.
Published: (2026)
ClearMask: Noise-Free and Naturalness-Preserving Protection Against Voice Deepfake Attacks
by: Wang, Yuanda, et al.
Published: (2025)
by: Wang, Yuanda, et al.
Published: (2025)
DASM: Domain-Aware Sharpness Minimization for Multi-Domain Voice Stream Steganalysis
by: Zhou, Pengcheng, et al.
Published: (2026)
by: Zhou, Pengcheng, et al.
Published: (2026)
SARSteer: Safeguarding Large Audio-Language Models via Safe-Ablated Refusal Steering
by: Lin, Weilin, et al.
Published: (2025)
by: Lin, Weilin, et al.
Published: (2025)
Lightweight Protection for Privacy in Offloaded Speech Understanding
by: Cai, Dongqi
Published: (2024)
by: Cai, Dongqi
Published: (2024)
Adversarial Attacks and Defenses for Speech Recognition Systems
by: Żelasko, Piotr, et al.
Published: (2021)
by: Żelasko, Piotr, et al.
Published: (2021)
JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
by: Peng, Zifan, et al.
Published: (2025)
by: Peng, Zifan, et al.
Published: (2025)
Frame-level Temporal Difference Learning for Partial Deepfake Speech Detection
by: Li, Menglu, et al.
Published: (2025)
by: Li, Menglu, et al.
Published: (2025)
SafeSpeech: Robust and Universal Voice Protection Against Malicious Speech Synthesis
by: Zhang, Zhisheng, et al.
Published: (2025)
by: Zhang, Zhisheng, et al.
Published: (2025)
WaLi: Can Pressure Sensors in HVAC Systems Capture Human Speech?
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech
by: Liu, Xuechen, et al.
Published: (2025)
by: Liu, Xuechen, et al.
Published: (2025)
Why Speech Deepfake Detectors Won't Generalize: The Limits of Detection in an Open World
by: Berisha, Visar, et al.
Published: (2025)
by: Berisha, Visar, et al.
Published: (2025)
Zero-Query Adversarial Attack on Black-box Automatic Speech Recognition Systems
by: Fang, Zheng, et al.
Published: (2024)
by: Fang, Zheng, et al.
Published: (2024)
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
A Portable and Stealthy Inaudible Voice Attack Based on Acoustic Metamaterials
by: Ning, Zhiyuan, et al.
Published: (2025)
by: Ning, Zhiyuan, et al.
Published: (2025)
Cross-Technology Generalization in Synthesized Speech Detection: Evaluating AST Models with Modern Voice Generators
by: Ustinov, Andrew, et al.
Published: (2025)
by: Ustinov, Andrew, et al.
Published: (2025)
Similar Items
-
Whispering Under the Eaves: Protecting User Privacy Against Commercial and LLM-powered Automatic Speech Recognition Systems
by: Jin, Weifei, et al.
Published: (2025) -
Acoustic Interference: A New Paradigm Weaponizing Acoustic Latent Semantic for Universal Jailbreak against Large Audio Language Models
by: Wang, Yanyun, et al.
Published: (2026) -
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
by: Huang, Kunyang, et al.
Published: (2025) -
Selective Masking Adversarial Attack on Automatic Speech Recognition Systems
by: Fang, Zheng, et al.
Published: (2025) -
HVAC-EAR: Eavesdropping Human Speech Using HVAC Systems
by: Tamiti, Tarikul Islam, et al.
Published: (2025)