Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Goel, Diksha, Moore, Kristen, Wang, Jeff, Kim, Minjune, Nguyen, Thanh Thi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimizing Cyber Defense in Dynamic Active Directories through Reinforcement Learning
by: Goel, Diksha, et al.
Published: (2024)
by: Goel, Diksha, et al.
Published: (2024)
CyberAlly: Leveraging LLMs and Knowledge Graphs to Empower Cyber Defenders
by: Kim, Minjune, et al.
Published: (2025)
by: Kim, Minjune, et al.
Published: (2025)
Explainable Autonomous Cyber Defense using Adversarial Multi-Agent Reinforcement Learning
by: Zhang, Yiyao, et al.
Published: (2026)
by: Zhang, Yiyao, et al.
Published: (2026)
CAMP in the Odyssey: Provably Robust Reinforcement Learning with Certified Radius Maximization
by: Wang, Derui, et al.
Published: (2025)
by: Wang, Derui, et al.
Published: (2025)
TempoNet: Learning Realistic Communication and Timing Patterns for Network Traffic Simulation
by: Moore, Kristen, et al.
Published: (2026)
by: Moore, Kristen, et al.
Published: (2026)
Machine Learning Driven Smishing Detection Framework for Mobile Security
by: Goel, Diksha, et al.
Published: (2024)
by: Goel, Diksha, et al.
Published: (2024)
Co-Evolutionary Defence of Active Directory Attack Graphs via GNN-Approximated Dynamic Programming
by: Goel, Diksha, et al.
Published: (2025)
by: Goel, Diksha, et al.
Published: (2025)
Distributional Black-Box Model Inversion Attack with Multi-Agent Reinforcement Learning
by: Bao, Huan, et al.
Published: (2024)
by: Bao, Huan, et al.
Published: (2024)
Interpreting Agent Behaviors in Reinforcement-Learning-Based Cyber-Battle Simulation Platforms
by: Claypoole, Jared, et al.
Published: (2025)
by: Claypoole, Jared, et al.
Published: (2025)
Learning to Communicate in Multi-Agent Reinforcement Learning for Autonomous Cyber Defence
by: Contractor, Faizan, et al.
Published: (2025)
by: Contractor, Faizan, et al.
Published: (2025)
Automated Cyber Defense with Generalizable Graph-based Reinforcement Learning Agents
by: King, Isaiah J., et al.
Published: (2025)
by: King, Isaiah J., et al.
Published: (2025)
Multi-Agent Reinforcement Learning for Maritime Operational Technology Cyber Security
by: Wilson, Alec, et al.
Published: (2024)
by: Wilson, Alec, et al.
Published: (2024)
From Promise to Peril: Rethinking Cybersecurity Red and Blue Teaming in the Age of LLMs
by: Abuadbba, Alsharif, et al.
Published: (2025)
by: Abuadbba, Alsharif, et al.
Published: (2025)
Detection and Prevention of Smishing Attacks
by: Goel, Diksha
Published: (2024)
by: Goel, Diksha
Published: (2024)
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025)
by: Cadet, Xavier, et al.
Published: (2025)
Entity-based Reinforcement Learning for Autonomous Cyber Defence
by: Thompson, Isaac Symes, et al.
Published: (2024)
by: Thompson, Isaac Symes, et al.
Published: (2024)
Learning Communication Between Heterogeneous Agents in Multi-Agent Reinforcement Learning for Autonomous Cyber Defence
by: Popa, Alex, et al.
Published: (2026)
by: Popa, Alex, et al.
Published: (2026)
Hierarchical Multi-agent Reinforcement Learning for Cyber Network Defense
by: Singh, Aditya Vikram, et al.
Published: (2024)
by: Singh, Aditya Vikram, et al.
Published: (2024)
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning
by: Domico, Kyle, et al.
Published: (2025)
by: Domico, Kyle, et al.
Published: (2025)
Black-Box Privacy Attacks on Shared Representations in Multitask Learning
by: Abascal, John, et al.
Published: (2025)
by: Abascal, John, et al.
Published: (2025)
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Pack-A-Mal: A Malware Analysis Framework for Open-Source Packages
by: Vu, Duc-Ly, et al.
Published: (2025)
by: Vu, Duc-Ly, et al.
Published: (2025)
Nearly Tight Black-Box Auditing of Differentially Private Machine Learning
by: Annamalai, Meenatchi Sundaram Muthu Selva, et al.
Published: (2024)
by: Annamalai, Meenatchi Sundaram Muthu Selva, et al.
Published: (2024)
Manipulating Recommender Systems: A Survey of Poisoning Attacks and Countermeasures
by: Nguyen, Thanh Toan, et al.
Published: (2024)
by: Nguyen, Thanh Toan, et al.
Published: (2024)
Hierarchical Adversarially-Resilient Multi-Agent Reinforcement Learning for Cyber-Physical Systems Security
by: Alqithami, Saad
Published: (2025)
by: Alqithami, Saad
Published: (2025)
Effective Multi-Stage Training Model For Edge Computing Devices In Intrusion Detection
by: Trong, Thua Huynh, et al.
Published: (2024)
by: Trong, Thua Huynh, et al.
Published: (2024)
Black-Box Detection of Language Model Watermarks
by: Gloaguen, Thibaud, et al.
Published: (2024)
by: Gloaguen, Thibaud, et al.
Published: (2024)
Uncertainty-Aware Federated Learning for Cyber-Resilient Microgrid Energy Management
by: Babayomi, Oluleke, et al.
Published: (2025)
by: Babayomi, Oluleke, et al.
Published: (2025)
DeepTaster: Adversarial Perturbation-Based Fingerprinting to Identify Proprietary Dataset Use in Deep Neural Networks
by: Park, Seonhye, et al.
Published: (2022)
by: Park, Seonhye, et al.
Published: (2022)
Toward a Multi-Layer ML-Based Security Framework for Industrial IoT
by: Bouferroum, Aymen, et al.
Published: (2026)
by: Bouferroum, Aymen, et al.
Published: (2026)
Can Current Detectors Catch Face-to-Voice Deepfake Attacks?
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
Evaluating Generalization Mechanisms in Autonomous Cyber Attack Agents
by: Lukáš, Ondřej, et al.
Published: (2026)
by: Lukáš, Ondřej, et al.
Published: (2026)
Multi-Objective Reinforcement Learning for Automated Resilient Cyber Defence
by: O'Driscoll, Ross, et al.
Published: (2024)
by: O'Driscoll, Ross, et al.
Published: (2024)
The Challenge of Identifying the Origin of Black-Box Large Language Models
by: Yang, Ziqing, et al.
Published: (2025)
by: Yang, Ziqing, et al.
Published: (2025)
BruSLeAttack: A Query-Efficient Score-Based Black-Box Sparse Adversarial Attack
by: Vo, Viet Quoc, et al.
Published: (2024)
by: Vo, Viet Quoc, et al.
Published: (2024)
AgenticVM: Agentic AI for Adaptive Software Vulnerability Management
by: Arifin, Asrul, et al.
Published: (2026)
by: Arifin, Asrul, et al.
Published: (2026)
Auditing Differential Privacy in the Black-Box Setting
by: Shi, Kaining, et al.
Published: (2025)
by: Shi, Kaining, et al.
Published: (2025)
Dissecting the Black Box: Circuit-Level Analysis of LLM Vulnerability Detection
by: Atiiq, Syafiq Al, et al.
Published: (2026)
by: Atiiq, Syafiq Al, et al.
Published: (2026)
Targeted Data Poisoning for Black-Box Audio Datasets Ownership Verification
by: Bouaziz, Wassim, et al.
Published: (2025)
by: Bouaziz, Wassim, et al.
Published: (2025)
Turning Black Box into White Box: Dataset Distillation Leaks
by: Chen, Huajie, et al.
Published: (2026)
by: Chen, Huajie, et al.
Published: (2026)
Similar Items
-
Optimizing Cyber Defense in Dynamic Active Directories through Reinforcement Learning
by: Goel, Diksha, et al.
Published: (2024) -
CyberAlly: Leveraging LLMs and Knowledge Graphs to Empower Cyber Defenders
by: Kim, Minjune, et al.
Published: (2025) -
Explainable Autonomous Cyber Defense using Adversarial Multi-Agent Reinforcement Learning
by: Zhang, Yiyao, et al.
Published: (2026) -
CAMP in the Odyssey: Provably Robust Reinforcement Learning with Certified Radius Maximization
by: Wang, Derui, et al.
Published: (2025) -
TempoNet: Learning Realistic Communication and Timing Patterns for Network Traffic Simulation
by: Moore, Kristen, et al.
Published: (2026)