Optimal Defender Strategies for CAGE-2 using Causal Modeling and Tree Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hammar, Kim, Dhir, Neil, Stadler, Rolf |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Intrusion Prevention through Optimal Stopping
von: Hammar, Kim, et al.
Veröffentlicht: (2021)
von: Hammar, Kim, et al.
Veröffentlicht: (2021)
Learning Intrusion Prevention Policies through Optimal Stopping
von: Hammar, Kim, et al.
Veröffentlicht: (2021)
von: Hammar, Kim, et al.
Veröffentlicht: (2021)
Learning Optimal Defender Strategies for CAGE-2 using a POMDP Model
von: Le, Duc Huy, et al.
Veröffentlicht: (2025)
von: Le, Duc Huy, et al.
Veröffentlicht: (2025)
A System for Interactive Examination of Learned Security Policies
von: Hammar, Kim, et al.
Veröffentlicht: (2022)
von: Hammar, Kim, et al.
Veröffentlicht: (2022)
Finding Effective Security Strategies through Reinforcement Learning and Self-Play
von: Hammar, Kim, et al.
Veröffentlicht: (2020)
von: Hammar, Kim, et al.
Veröffentlicht: (2020)
Scalable Learning of Intrusion Responses through Recursive Decomposition
von: Hammar, Kim, et al.
Veröffentlicht: (2023)
von: Hammar, Kim, et al.
Veröffentlicht: (2023)
Learning Near-Optimal Intrusion Responses Against Dynamic Attackers
von: Hammar, Kim, et al.
Veröffentlicht: (2023)
von: Hammar, Kim, et al.
Veröffentlicht: (2023)
Automated Security Response through Online Learning with Adaptive Conjectures
von: Hammar, Kim, et al.
Veröffentlicht: (2024)
von: Hammar, Kim, et al.
Veröffentlicht: (2024)
General Autonomous Cybersecurity Defense: Learning Robust Policies for Dynamic Topologies and Diverse Attackers
von: Ramamurthy, Arun, et al.
Veröffentlicht: (2025)
von: Ramamurthy, Arun, et al.
Veröffentlicht: (2025)
Online Incident Response Planning under Model Misspecification through Bayesian Learning and Belief Quantization
von: Hammar, Kim, et al.
Veröffentlicht: (2025)
von: Hammar, Kim, et al.
Veröffentlicht: (2025)
Defending Against Unforeseen Failure Modes with Latent Adversarial Training
von: Casper, Stephen, et al.
Veröffentlicht: (2024)
von: Casper, Stephen, et al.
Veröffentlicht: (2024)
Your Agent Can Defend Itself against Backdoor Attacks
von: Changjiang, Li, et al.
Veröffentlicht: (2025)
von: Changjiang, Li, et al.
Veröffentlicht: (2025)
Efficient but Vulnerable: Benchmarking and Defending LLM Batch Prompting Attack
von: Yue, Murong, et al.
Veröffentlicht: (2025)
von: Yue, Murong, et al.
Veröffentlicht: (2025)
Learning to Defend by Attacking (and Vice-Versa): Transfer of Learning in Cybersecurity Games
von: Malloy, Tailia, et al.
Veröffentlicht: (2023)
von: Malloy, Tailia, et al.
Veröffentlicht: (2023)
Intrusion Tolerance for Networked Systems through Two-Level Feedback Control
von: Hammar, Kim, et al.
Veröffentlicht: (2024)
von: Hammar, Kim, et al.
Veröffentlicht: (2024)
Defending the Edge: Representative-Attention Defense against Backdoor Attacks in Federated Learning
von: Obioma, Chibueze Peace, et al.
Veröffentlicht: (2025)
von: Obioma, Chibueze Peace, et al.
Veröffentlicht: (2025)
Attackers Strike Back? Not Anymore -- An Ensemble of RL Defenders Awakens for APT Detection
von: Benabderrahmane, Sidahmed, et al.
Veröffentlicht: (2025)
von: Benabderrahmane, Sidahmed, et al.
Veröffentlicht: (2025)
Have You Poisoned My Data? Defending Neural Networks against Data Poisoning
von: De Gaspari, Fabio, et al.
Veröffentlicht: (2024)
von: De Gaspari, Fabio, et al.
Veröffentlicht: (2024)
Demystifying the Mythos or Disrupting Bugonomics? From Zero-Day Asymmetry to Defender Remediation Throughput
von: Pesoli, Alfredo, et al.
Veröffentlicht: (2026)
von: Pesoli, Alfredo, et al.
Veröffentlicht: (2026)
Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks
von: Halloran, John T., et al.
Veröffentlicht: (2026)
von: Halloran, John T., et al.
Veröffentlicht: (2026)
Trapping Attacker in Dilemma: Examining Internal Correlations and External Influences of Trigger for Defending GNN Backdoors
von: Yang, Fan, et al.
Veröffentlicht: (2026)
von: Yang, Fan, et al.
Veröffentlicht: (2026)
SPARD: Defending Harmful Fine-Tuning Attack via Safety Projection with Relevance-Diversity Data Selection
von: Chen, Shuhao, et al.
Veröffentlicht: (2026)
von: Chen, Shuhao, et al.
Veröffentlicht: (2026)
An Optimal Cascade Feature-Level Spatiotemporal Fusion Strategy for Anomaly Detection in CAN Bus
von: Fatahi, Mohammad, et al.
Veröffentlicht: (2025)
von: Fatahi, Mohammad, et al.
Veröffentlicht: (2025)
Concealing Backdoor Model Updates in Federated Learning by Trigger-Optimized Data Poisoning
von: Zhang, Yujie, et al.
Veröffentlicht: (2024)
von: Zhang, Yujie, et al.
Veröffentlicht: (2024)
CSLE: A Reinforcement Learning Platform for Autonomous Security Management
von: Hammar, Kim
Veröffentlicht: (2026)
von: Hammar, Kim
Veröffentlicht: (2026)
A Survey on Model Extraction Attacks and Defenses for Large Language Models
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
PLeak: Prompt Leaking Attacks against Large Language Model Applications
von: Hui, Bo, et al.
Veröffentlicht: (2024)
von: Hui, Bo, et al.
Veröffentlicht: (2024)
HSF: Defending against Jailbreak Attacks with Hidden State Filtering
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
IT Intrusion Detection Using Statistical Learning and Testbed Measurements
von: Wang, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoxuan, et al.
Veröffentlicht: (2024)
A Survey of Model Extraction Attacks and Defenses in Distributed Computing Environments
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
A Causal Perspective for Enhancing Jailbreak Attack and Defense
von: Pan, Licheng, et al.
Veröffentlicht: (2026)
von: Pan, Licheng, et al.
Veröffentlicht: (2026)
A Systematic Survey of Model Extraction Attacks and Defenses: State-of-the-Art and Perspectives
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
Unsafe LLM-Based Search: Quantitative Analysis and Mitigation of Safety Risks in AI Web Search
von: Luo, Zeren, et al.
Veröffentlicht: (2025)
von: Luo, Zeren, et al.
Veröffentlicht: (2025)
Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM
von: Cao, Bochuan, et al.
Veröffentlicht: (2023)
von: Cao, Bochuan, et al.
Veröffentlicht: (2023)
Mitigating Watermark Forgery in Generative Models via Randomized Key Selection
von: Aremu, Toluwani, et al.
Veröffentlicht: (2025)
von: Aremu, Toluwani, et al.
Veröffentlicht: (2025)
MetaDefense: Defending Finetuning-based Jailbreak Attack Before and During Generation
von: Jiang, Weisen, et al.
Veröffentlicht: (2025)
von: Jiang, Weisen, et al.
Veröffentlicht: (2025)
BountyBench: Dollar Impact of AI Agent Attackers and Defenders on Real-World Cybersecurity Systems
von: Zhang, Andy K., et al.
Veröffentlicht: (2025)
von: Zhang, Andy K., et al.
Veröffentlicht: (2025)
Rapid Plug-in Defenders
von: Wu, Kai, et al.
Veröffentlicht: (2023)
von: Wu, Kai, et al.
Veröffentlicht: (2023)
A Comprehensive Comparative Study of Individual ML Models and Ensemble Strategies for Network Intrusion Detection Systems
von: Bibers, Ismail, et al.
Veröffentlicht: (2024)
von: Bibers, Ismail, et al.
Veröffentlicht: (2024)
Differentially Private In-Context Learning with Nearest Neighbor Search
von: Koskela, Antti, et al.
Veröffentlicht: (2025)
von: Koskela, Antti, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Intrusion Prevention through Optimal Stopping
von: Hammar, Kim, et al.
Veröffentlicht: (2021) -
Learning Intrusion Prevention Policies through Optimal Stopping
von: Hammar, Kim, et al.
Veröffentlicht: (2021) -
Learning Optimal Defender Strategies for CAGE-2 using a POMDP Model
von: Le, Duc Huy, et al.
Veröffentlicht: (2025) -
A System for Interactive Examination of Learned Security Policies
von: Hammar, Kim, et al.
Veröffentlicht: (2022) -
Finding Effective Security Strategies through Reinforcement Learning and Self-Play
von: Hammar, Kim, et al.
Veröffentlicht: (2020)