Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Meng, Wang, Kun, Lu, Li, Zhang, Jiaheng, Zhang, Tianwei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
di: Ziv, Roee, et al.
Pubblicazione: (2025)
di: Ziv, Roee, et al.
Pubblicazione: (2025)
Measuring the Robustness of Audio Deepfake Detectors
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
Smark: A Watermark for Text-to-Speech Diffusion Models via Discrete Wavelet Transform
di: Zhang, Yichuan, et al.
Pubblicazione: (2025)
di: Zhang, Yichuan, et al.
Pubblicazione: (2025)
STEP: Detecting Audio Backdoor Attacks via Stability-based Trigger Exposure Profiling
di: Wang, Kun, et al.
Pubblicazione: (2026)
di: Wang, Kun, et al.
Pubblicazione: (2026)
AdvWave: Stealthy Adversarial Jailbreak Attack against Large Audio-Language Models
di: Kang, Mintong, et al.
Pubblicazione: (2024)
di: Kang, Mintong, et al.
Pubblicazione: (2024)
AudioJailbreak: Jailbreak Attacks against End-to-End Large Audio-Language Models
di: Chen, Guangke, et al.
Pubblicazione: (2025)
di: Chen, Guangke, et al.
Pubblicazione: (2025)
JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
di: Peng, Zifan, et al.
Pubblicazione: (2025)
di: Peng, Zifan, et al.
Pubblicazione: (2025)
Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization
di: Fang, Zheng, et al.
Pubblicazione: (2026)
di: Fang, Zheng, et al.
Pubblicazione: (2026)
An Engorgio Prompt Makes Large Language Model Babble on
di: Dong, Jianshuo, et al.
Pubblicazione: (2024)
di: Dong, Jianshuo, et al.
Pubblicazione: (2024)
SARSteer: Safeguarding Large Audio-Language Models via Safe-Ablated Refusal Steering
di: Lin, Weilin, et al.
Pubblicazione: (2025)
di: Lin, Weilin, et al.
Pubblicazione: (2025)
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models
di: Zhang, Yucheng, et al.
Pubblicazione: (2024)
di: Zhang, Yucheng, et al.
Pubblicazione: (2024)
Beyond Content: A Comprehensive Speech Toxicity Dataset and Detection Framework Incorporating Paralinguistic Cues
di: Ba, Zhongjie, et al.
Pubblicazione: (2026)
di: Ba, Zhongjie, et al.
Pubblicazione: (2026)
SOLIDO: A Robust Watermarking Method for Speech Synthesis via Low-Rank Adaptation
di: Li, Yue, et al.
Pubblicazione: (2025)
di: Li, Yue, et al.
Pubblicazione: (2025)
Sirens' Whisper: Inaudible Near-Ultrasonic Jailbreaks of Speech-Driven LLMs
di: Ling, Zijian, et al.
Pubblicazione: (2026)
di: Ling, Zijian, et al.
Pubblicazione: (2026)
When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs
di: Dingeto, Hiskias, et al.
Pubblicazione: (2025)
di: Dingeto, Hiskias, et al.
Pubblicazione: (2025)
GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis
di: Liu, Weizhi, et al.
Pubblicazione: (2024)
di: Liu, Weizhi, et al.
Pubblicazione: (2024)
Synthetic Voices, Real Threats: Evaluating Large Text-to-Speech Models in Generating Harmful Audio
di: Chen, Guangke, et al.
Pubblicazione: (2025)
di: Chen, Guangke, et al.
Pubblicazione: (2025)
Invisible Prompts, Visible Threats: Malicious Font Injection in External Resources for Large Language Models
di: Xiong, Junjie, et al.
Pubblicazione: (2025)
di: Xiong, Junjie, et al.
Pubblicazione: (2025)
Evaluation of Prompt Injection Defenses in Large Language Models
di: Deep, Priyal, et al.
Pubblicazione: (2026)
di: Deep, Priyal, et al.
Pubblicazione: (2026)
Imperceptible Jailbreaking against Large Language Models
di: Gao, Kuofeng, et al.
Pubblicazione: (2025)
di: Gao, Kuofeng, et al.
Pubblicazione: (2025)
Exploiting Vulnerabilities in Speech Translation Systems through Targeted Adversarial Attacks
di: Liu, Chang, et al.
Pubblicazione: (2025)
di: Liu, Chang, et al.
Pubblicazione: (2025)
DMark: Order-Agnostic Watermarking for Diffusion Large Language Models
di: Wu, Linyu, et al.
Pubblicazione: (2025)
di: Wu, Linyu, et al.
Pubblicazione: (2025)
Protecting Your Voice: Temporal-aware Robust Watermarking
di: Li, Yue, et al.
Pubblicazione: (2025)
di: Li, Yue, et al.
Pubblicazione: (2025)
ALIF: Low-Cost Adversarial Audio Attacks on Black-Box Speech Platforms using Linguistic Features
di: Cheng, Peng, et al.
Pubblicazione: (2024)
di: Cheng, Peng, et al.
Pubblicazione: (2024)
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
Prompt Injection as an Emerging Threat: Evaluating the Resilience of Large Language Models
di: Ganiuly, Daniyal, et al.
Pubblicazione: (2025)
di: Ganiuly, Daniyal, et al.
Pubblicazione: (2025)
System Prompt Poisoning: Persistent Attacks on Large Language Models Beyond User Injection
di: Li, Zongze, et al.
Pubblicazione: (2025)
di: Li, Zongze, et al.
Pubblicazione: (2025)
SafeEar: Content Privacy-Preserving Audio Deepfake Detection
di: Li, Xinfeng, et al.
Pubblicazione: (2024)
di: Li, Xinfeng, et al.
Pubblicazione: (2024)
Proactive Detection of Voice Cloning with Localized Watermarking
di: Roman, Robin San, et al.
Pubblicazione: (2024)
di: Roman, Robin San, et al.
Pubblicazione: (2024)
SCDF: A Speaker Characteristics DeepFake Speech Dataset for Bias Analysis
di: Staněk, Vojtěch, et al.
Pubblicazione: (2025)
di: Staněk, Vojtěch, et al.
Pubblicazione: (2025)
Decoding Deception: Understanding Automatic Speech Recognition Vulnerabilities in Evasion and Poisoning Attacks
di: G, Aravindhan, et al.
Pubblicazione: (2025)
di: G, Aravindhan, et al.
Pubblicazione: (2025)
DeePen: Penetration Testing for Audio Deepfake Detection
di: Müller, Nicolas, et al.
Pubblicazione: (2025)
di: Müller, Nicolas, et al.
Pubblicazione: (2025)
Exploring Audio Editing Features as User-Centric Privacy Defenses Against Large Language Model(LLM) Based Emotion Inference Attacks
di: Soumik, Mohd. Farhan Israk, et al.
Pubblicazione: (2025)
di: Soumik, Mohd. Farhan Israk, et al.
Pubblicazione: (2025)
Hallucinating AI Hijacking Attack: Large Language Models and Malicious Code Recommenders
di: Noever, David, et al.
Pubblicazione: (2024)
di: Noever, David, et al.
Pubblicazione: (2024)
Acoustic Interference: A New Paradigm Weaponizing Acoustic Latent Semantic for Universal Jailbreak against Large Audio Language Models
di: Wang, Yanyun, et al.
Pubblicazione: (2026)
di: Wang, Yanyun, et al.
Pubblicazione: (2026)
Image-based Prompt Injection: Hijacking Multimodal LLMs through Visually Embedded Adversarial Instructions
di: Nagaraja, Neha, et al.
Pubblicazione: (2026)
di: Nagaraja, Neha, et al.
Pubblicazione: (2026)
Goal-guided Generative Prompt Injection Attack on Large Language Models
di: Zhang, Chong, et al.
Pubblicazione: (2024)
di: Zhang, Chong, et al.
Pubblicazione: (2024)
Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World
di: Sadasivan, Vinu Sankar, et al.
Pubblicazione: (2025)
di: Sadasivan, Vinu Sankar, et al.
Pubblicazione: (2025)
Review-Incorporated Model-Agnostic Profile Injection Attacks on Recommender Systems
di: Yang, Shiyi, et al.
Pubblicazione: (2024)
di: Yang, Shiyi, et al.
Pubblicazione: (2024)
Prompt Tuning for Audio Deepfake Detection: Computationally Efficient Test-time Domain Adaptation with Limited Target Dataset
di: Oiso, Hideyuki, et al.
Pubblicazione: (2024)
di: Oiso, Hideyuki, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
di: Ziv, Roee, et al.
Pubblicazione: (2025) -
Measuring the Robustness of Audio Deepfake Detectors
di: Li, Xiang, et al.
Pubblicazione: (2025) -
Smark: A Watermark for Text-to-Speech Diffusion Models via Discrete Wavelet Transform
di: Zhang, Yichuan, et al.
Pubblicazione: (2025) -
STEP: Detecting Audio Backdoor Attacks via Stability-based Trigger Exposure Profiling
di: Wang, Kun, et al.
Pubblicazione: (2026) -
AdvWave: Stealthy Adversarial Jailbreak Attack against Large Audio-Language Models
di: Kang, Mintong, et al.
Pubblicazione: (2024)