From Promise to Peril: Rethinking Cybersecurity Red and Blue Teaming in the Age of LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Abuadbba, Alsharif, Hicks, Chris, Moore, Kristen, Mavroudis, Vasilios, Hasircioglu, Burak, Goel, Diksha, Jennings, Piers |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One Pic is All it Takes: Poisoning Visual Document Retrieval Augmented Generation with a Single Image
by: Shereen, Ezzeldin, et al.
Published: (2025)
by: Shereen, Ezzeldin, et al.
Published: (2025)
Mitigating Deep Reinforcement Learning Backdoors in the Neural Activation Space
by: Vyas, Sanyam, et al.
Published: (2024)
by: Vyas, Sanyam, et al.
Published: (2024)
Less is more? Rewards in RL for Cyber Defence
by: Bates, Elizabeth, et al.
Published: (2025)
by: Bates, Elizabeth, et al.
Published: (2025)
CybORG++: An Enhanced Gym for the Development of Autonomous Cyber Agents
by: Emerson, Harry, et al.
Published: (2024)
by: Emerson, Harry, et al.
Published: (2024)
Entity-based Reinforcement Learning for Autonomous Cyber Defence
by: Thompson, Isaac Symes, et al.
Published: (2024)
by: Thompson, Isaac Symes, et al.
Published: (2024)
SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity
by: McFadden, Shae, et al.
Published: (2026)
by: McFadden, Shae, et al.
Published: (2026)
Autonomous Network Defence using Reinforcement Learning
by: Foley, Myles, et al.
Published: (2024)
by: Foley, Myles, et al.
Published: (2024)
Zero-Trust Network Access (ZTNA)
by: Mavroudis, Vasilios
Published: (2024)
by: Mavroudis, Vasilios
Published: (2024)
What if we could hot swap our Biometrics?
by: Crowcroft, Jon, et al.
Published: (2025)
by: Crowcroft, Jon, et al.
Published: (2025)
Analysis of Publicly Accessible Operational Technology and Associated Risks
by: Rodda, Matthew, et al.
Published: (2025)
by: Rodda, Matthew, et al.
Published: (2025)
Quantifying Mix Network Privacy Erosion with Generative Models
by: Mavroudis, Vasilios, et al.
Published: (2025)
by: Mavroudis, Vasilios, et al.
Published: (2025)
APT-Agent: Automated Penetration Testing using Large Language Models
by: Li, William Guanting, et al.
Published: (2026)
by: Li, William Guanting, et al.
Published: (2026)
DeepiSign-G: Generic Watermark to Stamp Hidden DNN Parameters for Self-contained Tracking
by: Abuadbba, Alsharif, et al.
Published: (2024)
by: Abuadbba, Alsharif, et al.
Published: (2024)
Co-Evolutionary Defence of Active Directory Attack Graphs via GNN-Approximated Dynamic Programming
by: Goel, Diksha, et al.
Published: (2025)
by: Goel, Diksha, et al.
Published: (2025)
Detecting Misuse of Security APIs: A Systematic Review
by: Mousavi, Zahra, et al.
Published: (2023)
by: Mousavi, Zahra, et al.
Published: (2023)
Beyond Training-time Poisoning: Component-level and Post-training Backdoors in Deep Reinforcement Learning
by: Vyas, Sanyam, et al.
Published: (2025)
by: Vyas, Sanyam, et al.
Published: (2025)
Can Current Detectors Catch Face-to-Voice Deepfake Attacks?
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
HonestCyberEval: An AI Cyber Risk Benchmark for Automated Software Exploitation
by: Ristea, Dan, et al.
Published: (2024)
by: Ristea, Dan, et al.
Published: (2024)
Referential Security as a New Paradigm for AI Evaluations
by: Ristea, Dan, et al.
Published: (2026)
by: Ristea, Dan, et al.
Published: (2026)
An Investigation into Misuse of Java Security APIs by Large Language Models
by: Mousavi, Zahra, et al.
Published: (2024)
by: Mousavi, Zahra, et al.
Published: (2024)
An Attentive Graph Agent for Topology-Adaptive Cyber Defence
by: Sandoval, Ilya Orson, et al.
Published: (2025)
by: Sandoval, Ilya Orson, et al.
Published: (2025)
DeepTaster: Adversarial Perturbation-Based Fingerprinting to Identify Proprietary Dataset Use in Deep Neural Networks
by: Park, Seonhye, et al.
Published: (2022)
by: Park, Seonhye, et al.
Published: (2022)
Large Language Model Adversarial Landscape Through the Lens of Attack Objectives
by: Wang, Nan, et al.
Published: (2025)
by: Wang, Nan, et al.
Published: (2025)
ChatNVD: Advancing Cybersecurity Vulnerability Assessment with Large Language Models
by: Chopra, Shivansh, et al.
Published: (2024)
by: Chopra, Shivansh, et al.
Published: (2024)
NADD: Amplifying Noise for Effective Diffusion-based Adversarial Purification
by: Nguyen, David D., et al.
Published: (2026)
by: Nguyen, David D., et al.
Published: (2026)
On The Effectiveness of the UK NIS Regulations as a Mandatory Cybersecurity Reporting Regime
by: Ali, Junade, et al.
Published: (2026)
by: Ali, Junade, et al.
Published: (2026)
DRMD: Deep Reinforcement Learning for Malware Detection under Concept Drift
by: McFadden, Shae, et al.
Published: (2025)
by: McFadden, Shae, et al.
Published: (2025)
Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents
by: Goel, Diksha, et al.
Published: (2025)
by: Goel, Diksha, et al.
Published: (2025)
CyberAlly: Leveraging LLMs and Knowledge Graphs to Empower Cyber Defenders
by: Kim, Minjune, et al.
Published: (2025)
by: Kim, Minjune, et al.
Published: (2025)
Comprehensive Evaluation of Cloaking Backdoor Attacks on Object Detector in Real-World
by: Ma, Hua, et al.
Published: (2025)
by: Ma, Hua, et al.
Published: (2025)
EPhishCADE: A Privacy-Aware Multi-Dimensional Framework for Email Phishing Campaign Detection
by: Kang, Wei, et al.
Published: (2025)
by: Kang, Wei, et al.
Published: (2025)
Detection and Prevention of Smishing Attacks
by: Goel, Diksha
Published: (2024)
by: Goel, Diksha
Published: (2024)
Does Teaming-Up LLMs Improve Secure Code Generation? A Comprehensive Evaluation with Multi-LLMSecCodeEval
by: Sabir, Bushra, et al.
Published: (2026)
by: Sabir, Bushra, et al.
Published: (2026)
Human Society-Inspired Approaches to Agentic AI Security: The 4C Framework
by: Abuadbba, Alsharif, et al.
Published: (2026)
by: Abuadbba, Alsharif, et al.
Published: (2026)
SoK: Systematization and Benchmarking of Deepfake Detectors in a Unified Framework
by: Le, Binh M., et al.
Published: (2024)
by: Le, Binh M., et al.
Published: (2024)
PEEK: Phishing Evolution Framework for Phishing Generation and Evolving Pattern Analysis using Large Language Models
by: Chen, Fengchao, et al.
Published: (2024)
by: Chen, Fengchao, et al.
Published: (2024)
Benchmarking LLMs in an Embodied Environment for Blue Team Threat Hunting
by: Liu, Xiaoqun, et al.
Published: (2025)
by: Liu, Xiaoqun, et al.
Published: (2025)
Adversarial Attacks Against Automated Fact-Checking: A Survey
by: Liu, Fanzhen, et al.
Published: (2025)
by: Liu, Fanzhen, et al.
Published: (2025)
From Solitary Directives to Interactive Encouragement! LLM Secure Code Generation by Natural Language Prompting
by: Liu, Shigang, et al.
Published: (2024)
by: Liu, Shigang, et al.
Published: (2024)
SoK: Can Trajectory Generation Combine Privacy and Utility?
by: Buchholz, Erik, et al.
Published: (2024)
by: Buchholz, Erik, et al.
Published: (2024)
Similar Items
-
One Pic is All it Takes: Poisoning Visual Document Retrieval Augmented Generation with a Single Image
by: Shereen, Ezzeldin, et al.
Published: (2025) -
Mitigating Deep Reinforcement Learning Backdoors in the Neural Activation Space
by: Vyas, Sanyam, et al.
Published: (2024) -
Less is more? Rewards in RL for Cyber Defence
by: Bates, Elizabeth, et al.
Published: (2025) -
CybORG++: An Enhanced Gym for the Development of Autonomous Cyber Agents
by: Emerson, Harry, et al.
Published: (2024) -
Entity-based Reinforcement Learning for Autonomous Cyber Defence
by: Thompson, Isaac Symes, et al.
Published: (2024)