AgentDyn: Are Your Agent Security Defenses Deployable in Real-World Dynamic Environments?
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Hao, Wen, Ruoyao, Shi, Shanghao, Zhang, Ning, Vorobeychik, Yevgeniy, Xiao, Chaowei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AgentSys: Secure and Dynamic LLM Agents Through Explicit Hierarchical Memory Management
di: Wen, Ruoyao, et al.
Pubblicazione: (2026)
di: Wen, Ruoyao, et al.
Pubblicazione: (2026)
Low Rank Adaptation for Adversarial Perturbation
di: Liu, Han, et al.
Pubblicazione: (2026)
di: Liu, Han, et al.
Pubblicazione: (2026)
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
A Scalable Approach to Solving Simulation-Based Network Security Games
di: Lanier, Michael, et al.
Pubblicazione: (2026)
di: Lanier, Michael, et al.
Pubblicazione: (2026)
Sliced Rényi Pufferfish Privacy: Directional Additive Noise Mechanism and Private Learning with Gradient Clipping
di: Zhang, Tao, et al.
Pubblicazione: (2025)
di: Zhang, Tao, et al.
Pubblicazione: (2025)
Residual-PAC Privacy: Automatic Privacy Control Beyond the Gaussian Barrier
di: Zhang, Tao, et al.
Pubblicazione: (2025)
di: Zhang, Tao, et al.
Pubblicazione: (2025)
Multi-Agent Reinforcement Learning for Assessing False-Data Injection Attacks on Transportation Networks
di: Eghtesad, Taha, et al.
Pubblicazione: (2023)
di: Eghtesad, Taha, et al.
Pubblicazione: (2023)
A New Era in LLM Security: Exploring Security Concerns in Real-World LLM-based Systems
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
CyGym: A Simulation-Based Game-Theoretic Analysis Framework for Cybersecurity
di: Lanier, Michael, et al.
Pubblicazione: (2025)
di: Lanier, Michael, et al.
Pubblicazione: (2025)
RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models
di: Wang, Jiongxiao, et al.
Pubblicazione: (2023)
di: Wang, Jiongxiao, et al.
Pubblicazione: (2023)
AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs
di: Liu, Xiaogeng, et al.
Pubblicazione: (2024)
di: Liu, Xiaogeng, et al.
Pubblicazione: (2024)
Adversarial Reinforcement Learning for Detecting False Data Injection Attacks in Vehicular Routing
di: Eghtesad, Taha, et al.
Pubblicazione: (2026)
di: Eghtesad, Taha, et al.
Pubblicazione: (2026)
Differential Confounding Privacy and Inverse Composition
di: Zhang, Tao, et al.
Pubblicazione: (2024)
di: Zhang, Tao, et al.
Pubblicazione: (2024)
AgentSentinel: An End-to-End and Real-Time Security Defense Framework for Computer-Use Agents
di: Hu, Haitao, et al.
Pubblicazione: (2025)
di: Hu, Haitao, et al.
Pubblicazione: (2025)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
di: Xiang, Chong, et al.
Pubblicazione: (2026)
di: Xiang, Chong, et al.
Pubblicazione: (2026)
Defensible Design for OpenClaw: Securing Autonomous Tool-Invoking Agents
di: Li, Zongwei, et al.
Pubblicazione: (2026)
di: Li, Zongwei, et al.
Pubblicazione: (2026)
Adversarial Machine Unlearning
di: Di, Zonglin, et al.
Pubblicazione: (2024)
di: Di, Zonglin, et al.
Pubblicazione: (2024)
TraceGuard: Process-Guided Firewall against Reasoning Backdoors in Large Language Models
di: Guo, Zhen, et al.
Pubblicazione: (2026)
di: Guo, Zhen, et al.
Pubblicazione: (2026)
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
di: Debenedetti, Edoardo, et al.
Pubblicazione: (2024)
di: Debenedetti, Edoardo, et al.
Pubblicazione: (2024)
Bayes-Nash Generative Privacy Against Membership Inference Attacks
di: Zhang, Tao, et al.
Pubblicazione: (2024)
di: Zhang, Tao, et al.
Pubblicazione: (2024)
A Game-Theoretic Approach to Privacy-Utility Tradeoff in Sharing Genomic Summary Statistics
di: Zhang, Tao, et al.
Pubblicazione: (2024)
di: Zhang, Tao, et al.
Pubblicazione: (2024)
Code Agent can be an End-to-end System Hacker: Benchmarking Real-world Threats of Computer-use Agent
di: Luo, Weidi, et al.
Pubblicazione: (2025)
di: Luo, Weidi, et al.
Pubblicazione: (2025)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
di: Kong, Dezhang, et al.
Pubblicazione: (2025)
di: Kong, Dezhang, et al.
Pubblicazione: (2025)
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
di: Ying, Zonghao, et al.
Pubblicazione: (2026)
di: Ying, Zonghao, et al.
Pubblicazione: (2026)
RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments
di: Fu, Yuchuan, et al.
Pubblicazione: (2025)
di: Fu, Yuchuan, et al.
Pubblicazione: (2025)
System-Level Defense against Indirect Prompt Injection Attacks: An Information Flow Control Perspective
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment
di: Lin, Xixun, et al.
Pubblicazione: (2026)
di: Lin, Xixun, et al.
Pubblicazione: (2026)
System Password Security: Attack and Defense Mechanisms
di: Shi, Chaofang, et al.
Pubblicazione: (2025)
di: Shi, Chaofang, et al.
Pubblicazione: (2025)
SEC-bench: Automated Benchmarking of LLM Agents on Real-World Software Security Tasks
di: Lee, Hwiwon, et al.
Pubblicazione: (2025)
di: Lee, Hwiwon, et al.
Pubblicazione: (2025)
Attacks on Node Attributes in Graph Neural Networks
di: Xu, Ying, et al.
Pubblicazione: (2024)
di: Xu, Ying, et al.
Pubblicazione: (2024)
Systems-Level Attack Surface of Edge Agent Deployments on IoT
di: Zhan, Zhonghao, et al.
Pubblicazione: (2026)
di: Zhan, Zhonghao, et al.
Pubblicazione: (2026)
Defenses Against Prompt Attacks Learn Surface Heuristics
di: Li, Shawn, et al.
Pubblicazione: (2026)
di: Li, Shawn, et al.
Pubblicazione: (2026)
OpenClaw PRISM: A Zero-Fork, Defense-in-Depth Runtime Security Layer for Tool-Augmented LLM Agents
di: Li, Frank
Pubblicazione: (2026)
di: Li, Frank
Pubblicazione: (2026)
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
di: Cadet, Xavier, et al.
Pubblicazione: (2025)
di: Cadet, Xavier, et al.
Pubblicazione: (2025)
Don't Let the Claw Grip Your Hand: A Security Analysis and Defense Framework for OpenClaw
di: Shan, Zhengyang, et al.
Pubblicazione: (2026)
di: Shan, Zhengyang, et al.
Pubblicazione: (2026)
Your Agent, Their Asset: A Real-World Safety Analysis of OpenClaw
di: Wang, Zijun, et al.
Pubblicazione: (2026)
di: Wang, Zijun, et al.
Pubblicazione: (2026)
WIPI: A New Web Threat for LLM-Driven Web Agents
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
MindGuard: Intrinsic Decision Inspection for Securing LLM Agents Against Metadata Poisoning
di: Wang, Zhiqiang, et al.
Pubblicazione: (2025)
di: Wang, Zhiqiang, et al.
Pubblicazione: (2025)
Reframing LLM Agent Security as an Agent-Human Interaction Problem
di: Wang, Peiran, et al.
Pubblicazione: (2026)
di: Wang, Peiran, et al.
Pubblicazione: (2026)
Documenti analoghi
-
AgentSys: Secure and Dynamic LLM Agents Through Explicit Hierarchical Memory Management
di: Wen, Ruoyao, et al.
Pubblicazione: (2026) -
Low Rank Adaptation for Adversarial Perturbation
di: Liu, Han, et al.
Pubblicazione: (2026) -
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
di: Li, Hao, et al.
Pubblicazione: (2025) -
A Scalable Approach to Solving Simulation-Based Network Security Games
di: Lanier, Michael, et al.
Pubblicazione: (2026) -
Sliced Rényi Pufferfish Privacy: Directional Additive Noise Mechanism and Private Learning with Gradient Clipping
di: Zhang, Tao, et al.
Pubblicazione: (2025)