DemonAgent: Dynamically Encrypted Multi-Backdoor Implantation Attack on LLM-based Agent
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhu, Pengyu, Zhou, Zhenhong, Zhang, Yuanhe, Yan, Shilinlu, Wang, Kun, Su, Sen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Crabs: Consuming Resource via Auto-generation for LLM-DoS Attack under Black-box Settings
di: Zhang, Yuanhe, et al.
Pubblicazione: (2024)
di: Zhang, Yuanhe, et al.
Pubblicazione: (2024)
AutoBackdoor: Automating Backdoor Attacks via LLM Agents
di: Li, Yige, et al.
Pubblicazione: (2025)
di: Li, Yige, et al.
Pubblicazione: (2025)
Collaborative Shadows: Distributed Backdoor Attacks in LLM-Based Multi-Agent Systems
di: Zhu, Pengyu, et al.
Pubblicazione: (2025)
di: Zhu, Pengyu, et al.
Pubblicazione: (2025)
SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems
di: Feng, Yunhao, et al.
Pubblicazione: (2026)
di: Feng, Yunhao, et al.
Pubblicazione: (2026)
A Spatiotemporal Stealthy Backdoor Attack against Cooperative Multi-Agent Deep Reinforcement Learning
di: Yu, Yinbo, et al.
Pubblicazione: (2024)
di: Yu, Yinbo, et al.
Pubblicazione: (2024)
Can We Trust Embodied Agents? Exploring Backdoor Attacks against Embodied LLM-based Decision-Making Systems
di: Jiao, Ruochen, et al.
Pubblicazione: (2024)
di: Jiao, Ruochen, et al.
Pubblicazione: (2024)
Backdoor Attribution: Elucidating and Controlling Backdoor in Language Models
di: Yu, Miao, et al.
Pubblicazione: (2025)
di: Yu, Miao, et al.
Pubblicazione: (2025)
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
di: Tie, Guiyao, et al.
Pubblicazione: (2026)
di: Tie, Guiyao, et al.
Pubblicazione: (2026)
Compromising Embodied Agents with Contextual Backdoor Attacks
di: Liu, Aishan, et al.
Pubblicazione: (2024)
di: Liu, Aishan, et al.
Pubblicazione: (2024)
BadAgent: Inserting and Activating Backdoor Attacks in LLM Agents
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
ToolTweak: An Attack on Tool Selection in LLM-based Agents
di: Sneh, Jonathan, et al.
Pubblicazione: (2025)
di: Sneh, Jonathan, et al.
Pubblicazione: (2025)
AgentRAE: Remote Action Execution through Notification-based Visual Backdoors against Screenshots-based Mobile GUI Agents
di: Luo, Yutao, et al.
Pubblicazione: (2026)
di: Luo, Yutao, et al.
Pubblicazione: (2026)
Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate
di: Qi, Senmao, et al.
Pubblicazione: (2025)
di: Qi, Senmao, et al.
Pubblicazione: (2025)
Security of Internet of Agents: Attacks and Countermeasures
di: Wang, Yuntao, et al.
Pubblicazione: (2025)
di: Wang, Yuntao, et al.
Pubblicazione: (2025)
RECUR: Resource Exhaustion Attack via Recursive-Entropy Guided Counterfactual Utilization and Reflection
di: Wang, Ziwei, et al.
Pubblicazione: (2026)
di: Wang, Ziwei, et al.
Pubblicazione: (2026)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
Invisible to Humans, Triggered by Agents: Stealthy Jailbreak Attacks on Mobile Vision-Language Agents
di: Ding, Renhua, et al.
Pubblicazione: (2025)
di: Ding, Renhua, et al.
Pubblicazione: (2025)
LoopTrap: Termination Poisoning Attacks on LLM Agents
di: Xu, Huiyu, et al.
Pubblicazione: (2026)
di: Xu, Huiyu, et al.
Pubblicazione: (2026)
BLAST: A Stealthy Backdoor Leverage Attack against Cooperative Multi-Agent Deep Reinforcement Learning based Systems
di: Fang, Jing, et al.
Pubblicazione: (2025)
di: Fang, Jing, et al.
Pubblicazione: (2025)
CheatAgent: Attacking LLM-Empowered Recommender Systems via LLM Agent
di: Ning, Liang-bo, et al.
Pubblicazione: (2025)
di: Ning, Liang-bo, et al.
Pubblicazione: (2025)
Watch Out for Your Agents! Investigating Backdoor Threats to LLM-Based Agents
di: Yang, Wenkai, et al.
Pubblicazione: (2024)
di: Yang, Wenkai, et al.
Pubblicazione: (2024)
A Vision for Access Control in LLM-based Agent Systems
di: Li, Xinfeng, et al.
Pubblicazione: (2025)
di: Li, Xinfeng, et al.
Pubblicazione: (2025)
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks
di: Song, Ruoyu, et al.
Pubblicazione: (2024)
di: Song, Ruoyu, et al.
Pubblicazione: (2024)
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations
di: Ge, Huaizhi, et al.
Pubblicazione: (2024)
di: Ge, Huaizhi, et al.
Pubblicazione: (2024)
Exploring Backdoor Attack and Defense for LLM-empowered Recommendations
di: Ning, Liangbo, et al.
Pubblicazione: (2025)
di: Ning, Liangbo, et al.
Pubblicazione: (2025)
Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement Learning
di: Ma, Oubo, et al.
Pubblicazione: (2026)
di: Ma, Oubo, et al.
Pubblicazione: (2026)
Your Agent Can Defend Itself against Backdoor Attacks
di: Changjiang, Li, et al.
Pubblicazione: (2025)
di: Changjiang, Li, et al.
Pubblicazione: (2025)
Targeted Bit-Flip Attacks on LLM-Based Agents
di: Wang, Jialai, et al.
Pubblicazione: (2026)
di: Wang, Jialai, et al.
Pubblicazione: (2026)
Is Monitoring Enough? Strategic Agent Selection For Stealthy Attack in Multi-Agent Discussions
di: Xiang, Qiuchi, et al.
Pubblicazione: (2026)
di: Xiang, Qiuchi, et al.
Pubblicazione: (2026)
Resource Consumption Threats in Large Language Models
di: Zhang, Yuanhe, et al.
Pubblicazione: (2026)
di: Zhang, Yuanhe, et al.
Pubblicazione: (2026)
Invisible Textual Backdoor Attacks based on Dual-Trigger
di: Hou, Yang, et al.
Pubblicazione: (2024)
di: Hou, Yang, et al.
Pubblicazione: (2024)
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
di: Yan, Bingyu, et al.
Pubblicazione: (2025)
di: Yan, Bingyu, et al.
Pubblicazione: (2025)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
di: Wu, Yixin, et al.
Pubblicazione: (2025)
di: Wu, Yixin, et al.
Pubblicazione: (2025)
Your LLM Agent Can Leak Your Data: Data Exfiltration via Backdoored Tool Use
di: Zhang, Wuyang, et al.
Pubblicazione: (2026)
di: Zhang, Wuyang, et al.
Pubblicazione: (2026)
Secret Stealing Attacks on Local LLM Fine-Tuning through Supply-Chain Model Code Backdoors
di: Li, Zi, et al.
Pubblicazione: (2026)
di: Li, Zi, et al.
Pubblicazione: (2026)
CompressionAttack: Exploiting Prompt Compression as a New Attack Surface in LLM-Powered Agents
di: Liu, Zesen, et al.
Pubblicazione: (2025)
di: Liu, Zesen, et al.
Pubblicazione: (2025)
Enhancing Linux Privilege Escalation Attack Capabilities of Local LLM Agents
di: Probst, Benjamin, et al.
Pubblicazione: (2026)
di: Probst, Benjamin, et al.
Pubblicazione: (2026)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
di: Wang, Zhilong, et al.
Pubblicazione: (2025)
di: Wang, Zhilong, et al.
Pubblicazione: (2025)
When Alignment Isn't Enough: Response-Path Attacks on LLM Agents
di: Luo, Mingyu, et al.
Pubblicazione: (2026)
di: Luo, Mingyu, et al.
Pubblicazione: (2026)
SFIBA: Spatial-based Full-target Invisible Backdoor Attacks
di: Yin, Yangxu, et al.
Pubblicazione: (2025)
di: Yin, Yangxu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Crabs: Consuming Resource via Auto-generation for LLM-DoS Attack under Black-box Settings
di: Zhang, Yuanhe, et al.
Pubblicazione: (2024) -
AutoBackdoor: Automating Backdoor Attacks via LLM Agents
di: Li, Yige, et al.
Pubblicazione: (2025) -
Collaborative Shadows: Distributed Backdoor Attacks in LLM-Based Multi-Agent Systems
di: Zhu, Pengyu, et al.
Pubblicazione: (2025) -
SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems
di: Feng, Yunhao, et al.
Pubblicazione: (2026) -
A Spatiotemporal Stealthy Backdoor Attack against Cooperative Multi-Agent Deep Reinforcement Learning
di: Yu, Yinbo, et al.
Pubblicazione: (2024)