ALMGuard: Safety Shortcuts and Where to Find Them as Guardrails for Audio-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Weifei, Cao, Yuxin, Su, Junjie, Xue, Minhui, Hao, Jie, Xu, Ke, Dong, Jin Song, Wang, Derui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mirage Fools the Ear, Mute Hides the Truth: Precise Targeted Adversarial Attacks on Polyphonic Sound Event Detection Systems
von: Su, Junjie, et al.
Veröffentlicht: (2025)
von: Su, Junjie, et al.
Veröffentlicht: (2025)
Whispering Under the Eaves: Protecting User Privacy Against Commercial and LLM-powered Automatic Speech Recognition Systems
von: Jin, Weifei, et al.
Veröffentlicht: (2025)
von: Jin, Weifei, et al.
Veröffentlicht: (2025)
Towards Evaluating the Robustness of Automatic Speech Recognition Systems via Audio Style Transfer
von: Jin, Weifei, et al.
Veröffentlicht: (2024)
von: Jin, Weifei, et al.
Veröffentlicht: (2024)
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
von: Jin, Weifei, et al.
Veröffentlicht: (2025)
von: Jin, Weifei, et al.
Veröffentlicht: (2025)
E2E-VGuard: Adversarial Prevention for Production LLM-based End-To-End Speech Synthesis
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)
Bones of Contention: Exploring Query-Efficient Attacks against Skeleton Recognition Systems
von: Cao, Yuxin, et al.
Veröffentlicht: (2025)
von: Cao, Yuxin, et al.
Veröffentlicht: (2025)
DUAP: Dual-task Universal Adversarial Perturbations Against Voice Control Systems
von: Sun, Suyang, et al.
Veröffentlicht: (2026)
von: Sun, Suyang, et al.
Veröffentlicht: (2026)
Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs
von: Roh, Jaechul, et al.
Veröffentlicht: (2026)
von: Roh, Jaechul, et al.
Veröffentlicht: (2026)
MelShield: Robust Mel-Domain Audio Watermarking for Provenance Attribution of AI Generated Synthesized Speech
von: Jin, Yutong, et al.
Veröffentlicht: (2026)
von: Jin, Yutong, et al.
Veröffentlicht: (2026)
SafeSpeech: Robust and Universal Voice Protection Against Malicious Speech Synthesis
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)
Audio Pirates: Black-box Audio Watermark Removal via Diffusion Priors
von: Yao, Lingfeng, et al.
Veröffentlicht: (2026)
von: Yao, Lingfeng, et al.
Veröffentlicht: (2026)
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
Vulnerabilities of Audio-Based Biometric Authentication Systems Against Deepfake Speech Synthesis
von: Hong, Mengze, et al.
Veröffentlicht: (2026)
von: Hong, Mengze, et al.
Veröffentlicht: (2026)
SARSteer: Safeguarding Large Audio-Language Models via Safe-Ablated Refusal Steering
von: Lin, Weilin, et al.
Veröffentlicht: (2025)
von: Lin, Weilin, et al.
Veröffentlicht: (2025)
When Fine-Tuning is Not Enough: Lessons from HSAD on Hybrid and Adversarial Audio Spoof Detection
von: Hu, Bin, et al.
Veröffentlicht: (2025)
von: Hu, Bin, et al.
Veröffentlicht: (2025)
Acoustic Interference: A New Paradigm Weaponizing Acoustic Latent Semantic for Universal Jailbreak against Large Audio Language Models
von: Wang, Yanyun, et al.
Veröffentlicht: (2026)
von: Wang, Yanyun, et al.
Veröffentlicht: (2026)
SyncGuard: Robust Audio Watermarking Capable of Countering Desynchronization Attacks
von: Gan, Zhenliang, et al.
Veröffentlicht: (2025)
von: Gan, Zhenliang, et al.
Veröffentlicht: (2025)
PRoADS: Provably Secure and Robust Audio Diffusion Steganography with latent optimization and backward Euler Inversion
von: Yan, YongPeng, et al.
Veröffentlicht: (2026)
von: Yan, YongPeng, et al.
Veröffentlicht: (2026)
Measuring the Robustness of Audio Deepfake Detectors
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
von: Ziv, Roee, et al.
Veröffentlicht: (2025)
von: Ziv, Roee, et al.
Veröffentlicht: (2025)
Yours or Mine? Overwriting Attacks Against Neural Audio Watermarking
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)
von: Yao, Lingfeng, et al.
Veröffentlicht: (2025)
Where Do Backdoors Live? A Component-Level Analysis of Backdoor Propagation in Speech Language Models
von: Fortier, Alexandrine, et al.
Veröffentlicht: (2025)
von: Fortier, Alexandrine, et al.
Veröffentlicht: (2025)
The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents
von: Wang, Lixu, et al.
Veröffentlicht: (2025)
von: Wang, Lixu, et al.
Veröffentlicht: (2025)
StyleFool: Fooling Video Classification Systems via Style Transfer
von: Cao, Yuxin, et al.
Veröffentlicht: (2022)
von: Cao, Yuxin, et al.
Veröffentlicht: (2022)
STEP: Detecting Audio Backdoor Attacks via Stability-based Trigger Exposure Profiling
von: Wang, Kun, et al.
Veröffentlicht: (2026)
von: Wang, Kun, et al.
Veröffentlicht: (2026)
EveGuard: Defeating Vibration-based Side-Channel Eavesdropping with Audio Adversarial Perturbations
von: Chang, Jung-Woo, et al.
Veröffentlicht: (2024)
von: Chang, Jung-Woo, et al.
Veröffentlicht: (2024)
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
von: Chen, Meng, et al.
Veröffentlicht: (2026)
von: Chen, Meng, et al.
Veröffentlicht: (2026)
SilentCipher: Deep Audio Watermarking
von: Singh, Mayank Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Mayank Kumar, et al.
Veröffentlicht: (2024)
JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
von: Peng, Zifan, et al.
Veröffentlicht: (2025)
von: Peng, Zifan, et al.
Veröffentlicht: (2025)
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
von: Huang, Kunyang, et al.
Veröffentlicht: (2025)
von: Huang, Kunyang, et al.
Veröffentlicht: (2025)
Interpretable Temporal Class Activation Representation for Audio Spoofing Detection
von: Li, Menglu, et al.
Veröffentlicht: (2024)
von: Li, Menglu, et al.
Veröffentlicht: (2024)
Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis
von: Warren, Kevin, et al.
Veröffentlicht: (2025)
von: Warren, Kevin, et al.
Veröffentlicht: (2025)
One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection
von: Kim, Hyun Myung, et al.
Veröffentlicht: (2024)
von: Kim, Hyun Myung, et al.
Veröffentlicht: (2024)
Invisible Ears at Your Fingertips: Acoustic Eavesdropping via Mouse Sensors
von: Fakih, Mohamad, et al.
Veröffentlicht: (2025)
von: Fakih, Mohamad, et al.
Veröffentlicht: (2025)
DECKER: Domain-invariant Embedding for Cross-Keyboard Extraction and Recognition
von: Maurya, Bikrant Bikram Pratap, et al.
Veröffentlicht: (2026)
von: Maurya, Bikrant Bikram Pratap, et al.
Veröffentlicht: (2026)
Selective Masking Adversarial Attack on Automatic Speech Recognition Systems
von: Fang, Zheng, et al.
Veröffentlicht: (2025)
von: Fang, Zheng, et al.
Veröffentlicht: (2025)
ClearMask: Noise-Free and Naturalness-Preserving Protection Against Voice Deepfake Attacks
von: Wang, Yuanda, et al.
Veröffentlicht: (2025)
von: Wang, Yuanda, et al.
Veröffentlicht: (2025)
HVAC-EAR: Eavesdropping Human Speech Using HVAC Systems
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
DASM: Domain-Aware Sharpness Minimization for Multi-Domain Voice Stream Steganalysis
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2026)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2026)
Fingerprinting SDKs for Mobile Apps and Where to Find Them: Understanding the Market for Device Fingerprinting
von: Specter, Michael A., et al.
Veröffentlicht: (2025)
von: Specter, Michael A., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mirage Fools the Ear, Mute Hides the Truth: Precise Targeted Adversarial Attacks on Polyphonic Sound Event Detection Systems
von: Su, Junjie, et al.
Veröffentlicht: (2025) -
Whispering Under the Eaves: Protecting User Privacy Against Commercial and LLM-powered Automatic Speech Recognition Systems
von: Jin, Weifei, et al.
Veröffentlicht: (2025) -
Towards Evaluating the Robustness of Automatic Speech Recognition Systems via Audio Style Transfer
von: Jin, Weifei, et al.
Veröffentlicht: (2024) -
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
von: Jin, Weifei, et al.
Veröffentlicht: (2025) -
E2E-VGuard: Adversarial Prevention for Production LLM-based End-To-End Speech Synthesis
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)