Training RL Agents for Multi-Objective Network Defense Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Molina-Markham, Andres, Robaina, Luis, Steinle, Sean, Trivedi, Akash, Tsui, Derek, Potteiger, Nicholas, Brandt, Lauren, Winder, Ransom, Ridley, Ahmad |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Designing Robust Cyber-Defense Agents with Evolving Behavior Trees
por: Potteiger, Nicholas, et al.
Publicado: (2024)
por: Potteiger, Nicholas, et al.
Publicado: (2024)
Out-of-Distribution Detection for Neurosymbolic Autonomous Cyber Agents
por: Samaddar, Ankita, et al.
Publicado: (2024)
por: Samaddar, Ankita, et al.
Publicado: (2024)
Proceedings of the 2nd International Workshop on Adaptive Cyber Defense
por: Carvalho, Marco, et al.
Publicado: (2023)
por: Carvalho, Marco, et al.
Publicado: (2023)
Interpreting Agent Behaviors in Reinforcement-Learning-Based Cyber-Battle Simulation Platforms
por: Claypoole, Jared, et al.
Publicado: (2025)
por: Claypoole, Jared, et al.
Publicado: (2025)
Interpretability-Guided Test-Time Adversarial Defense
por: Kulkarni, Akshay, et al.
Publicado: (2024)
por: Kulkarni, Akshay, et al.
Publicado: (2024)
The Autonomy Tax: Defense Training Breaks LLM Agents
por: Li, Shawn, et al.
Publicado: (2026)
por: Li, Shawn, et al.
Publicado: (2026)
Explainable Autonomous Cyber Defense using Adversarial Multi-Agent Reinforcement Learning
por: Zhang, Yiyao, et al.
Publicado: (2026)
por: Zhang, Yiyao, et al.
Publicado: (2026)
RL and Fingerprinting to Select Moving Target Defense Mechanisms for Zero-day Attacks in IoT
por: Celdrán, Alberto Huertas, et al.
Publicado: (2022)
por: Celdrán, Alberto Huertas, et al.
Publicado: (2022)
Dynamic Dual-level Defense Routing for Continual Adversarial Training
por: Wang, Wenxuan, et al.
Publicado: (2025)
por: Wang, Wenxuan, et al.
Publicado: (2025)
Attacks and Defenses Against LLM Fingerprinting
por: Kurian, Kevin, et al.
Publicado: (2025)
por: Kurian, Kevin, et al.
Publicado: (2025)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
por: Wang, Yihan, et al.
Publicado: (2025)
por: Wang, Yihan, et al.
Publicado: (2025)
AgentDyn: Are Your Agent Security Defenses Deployable in Real-World Dynamic Environments?
por: Li, Hao, et al.
Publicado: (2026)
por: Li, Hao, et al.
Publicado: (2026)
The Path To Autonomous Cyber Defense
por: Oesch, Sean, et al.
Publicado: (2024)
por: Oesch, Sean, et al.
Publicado: (2024)
Anti-Sensing: Defense against Unauthorized Radar-based Human Vital Sign Sensing with Physically Realizable Wearable Oscillators
por: Oshim, Md Farhan Tasnim, et al.
Publicado: (2025)
por: Oshim, Md Farhan Tasnim, et al.
Publicado: (2025)
AgentSentinel: An End-to-End and Real-Time Security Defense Framework for Computer-Use Agents
por: Hu, Haitao, et al.
Publicado: (2025)
por: Hu, Haitao, et al.
Publicado: (2025)
Defending Against Prompt Injection With a Few DefensiveTokens
por: Chen, Sizhe, et al.
Publicado: (2025)
por: Chen, Sizhe, et al.
Publicado: (2025)
Defensible Design for OpenClaw: Securing Autonomous Tool-Invoking Agents
por: Li, Zongwei, et al.
Publicado: (2026)
por: Li, Zongwei, et al.
Publicado: (2026)
Cross-Task Defense: Instruction-Tuning LLMs for Content Safety
por: Fu, Yu, et al.
Publicado: (2024)
por: Fu, Yu, et al.
Publicado: (2024)
RLShield: Practical Multi-Agent RL for Financial Cyber Defense with Attack-Surface MDPs and Real-Time Response Orchestration
por: Nayak, Srikumar
Publicado: (2026)
por: Nayak, Srikumar
Publicado: (2026)
PocketAgents: A Manifest-Driven Library of Autonomous Defense Agents
por: Barbieri, Sidnei, et al.
Publicado: (2026)
por: Barbieri, Sidnei, et al.
Publicado: (2026)
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
por: Zeng, Yifan, et al.
Publicado: (2024)
por: Zeng, Yifan, et al.
Publicado: (2024)
Contextualized Privacy Defense for LLM Agents
por: Wen, Yule, et al.
Publicado: (2026)
por: Wen, Yule, et al.
Publicado: (2026)
Decoding FL Defenses: Systemization, Pitfalls, and Remedies
por: Khan, Momin Ahmad, et al.
Publicado: (2025)
por: Khan, Momin Ahmad, et al.
Publicado: (2025)
Elevating Defenses: Bridging Adversarial Training and Watermarking for Model Resilience
por: Thakkar, Janvi, et al.
Publicado: (2023)
por: Thakkar, Janvi, et al.
Publicado: (2023)
Toward a Dynamic Intellectual Property Protection Model in High-Growth SMEs
por: Pitruzzello, Sam, et al.
Publicado: (2026)
por: Pitruzzello, Sam, et al.
Publicado: (2026)
Threat Intelligence Driven IP Protection for Entrepreneurial SMEs
por: Pitruzzello, Sam, et al.
Publicado: (2026)
por: Pitruzzello, Sam, et al.
Publicado: (2026)
Efficient RL-based Cache Vulnerability Exploration by Penalizing Useless Agent Actions
por: Nakanishi, Kanato, et al.
Publicado: (2025)
por: Nakanishi, Kanato, et al.
Publicado: (2025)
Jatmo: Prompt Injection Defense by Task-Specific Finetuning
por: Piet, Julien, et al.
Publicado: (2023)
por: Piet, Julien, et al.
Publicado: (2023)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
por: Zhang, Hanrong, et al.
Publicado: (2024)
por: Zhang, Hanrong, et al.
Publicado: (2024)
Memory Poisoning Attack and Defense on Memory Based LLM-Agents
por: Sunil, Balachandra Devarangadi, et al.
Publicado: (2026)
por: Sunil, Balachandra Devarangadi, et al.
Publicado: (2026)
Preserving security in a world with powerful AI Considerations for the future Defense Architecture
por: Generous, Nicholas, et al.
Publicado: (2025)
por: Generous, Nicholas, et al.
Publicado: (2025)
RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents
por: Yeke, Doguhuan, et al.
Publicado: (2026)
por: Yeke, Doguhuan, et al.
Publicado: (2026)
Multi-Agent LLM Governance for Safe Two-Timescale Reinforcement Learning in SDN-IoT Defense
por: Jamshidi, Saeid, et al.
Publicado: (2026)
por: Jamshidi, Saeid, et al.
Publicado: (2026)
ALRPHFS: Adversarially Learned Risk Patterns with Hierarchical Fast \& Slow Reasoning for Robust Agent Defense
por: Xiang, Shiyu, et al.
Publicado: (2025)
por: Xiang, Shiyu, et al.
Publicado: (2025)
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
por: Ying, Zonghao, et al.
Publicado: (2026)
por: Ying, Zonghao, et al.
Publicado: (2026)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
por: Kong, Dezhang, et al.
Publicado: (2025)
por: Kong, Dezhang, et al.
Publicado: (2025)
Organizational Learning in Industry 4.0: Applying Crossan's 4I Framework with Double Loop Learning
por: Akram, Nimra, et al.
Publicado: (2025)
por: Akram, Nimra, et al.
Publicado: (2025)
Rotated Robustness: A Training-Free Defense against Bit-Flip Attacks on Large Language Models
por: Liu, Deng, et al.
Publicado: (2026)
por: Liu, Deng, et al.
Publicado: (2026)
Evaluating the Robustness of the "Ensemble Everything Everywhere" Defense
por: Zhang, Jie, et al.
Publicado: (2024)
por: Zhang, Jie, et al.
Publicado: (2024)
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
por: Debenedetti, Edoardo, et al.
Publicado: (2024)
por: Debenedetti, Edoardo, et al.
Publicado: (2024)
Ejemplares similares
-
Designing Robust Cyber-Defense Agents with Evolving Behavior Trees
por: Potteiger, Nicholas, et al.
Publicado: (2024) -
Out-of-Distribution Detection for Neurosymbolic Autonomous Cyber Agents
por: Samaddar, Ankita, et al.
Publicado: (2024) -
Proceedings of the 2nd International Workshop on Adaptive Cyber Defense
por: Carvalho, Marco, et al.
Publicado: (2023) -
Interpreting Agent Behaviors in Reinforcement-Learning-Based Cyber-Battle Simulation Platforms
por: Claypoole, Jared, et al.
Publicado: (2025) -
Interpretability-Guided Test-Time Adversarial Defense
por: Kulkarni, Akshay, et al.
Publicado: (2024)