Cybersecurity AI: Evaluating Agentic Cybersecurity in Attack/Defense CTFs
Fuente:
arXiv
Guardado en:
| Autores principales: | Balassone, Francesco, Mayoral-Vilches, Víctor, Rass, Stefan, Pinzger, Martin, Perrone, Gaetano, Romano, Simon Pietro, Schartner, Peter |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Cybersecurity AI: A Game-Theoretic AI for Guiding Attack and Defense
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
Cybersecurity AI: The Dangerous Gap Between Automation and Autonomy
por: Mayoral-Vilches, Víctor
Publicado: (2025)
por: Mayoral-Vilches, Víctor
Publicado: (2025)
Cybersecurity AI: Humanoid Robots as Attack Vectors
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
Cybersecurity AI Benchmark (CAIBench): A Meta-Benchmark for Evaluating Cybersecurity AI Agents
por: Sanz-Gómez, María, et al.
Publicado: (2025)
por: Sanz-Gómez, María, et al.
Publicado: (2025)
The Cybersecurity of a Humanoid Robot
por: Mayoral-Vilches, Víctor
Publicado: (2025)
por: Mayoral-Vilches, Víctor
Publicado: (2025)
Offensive Robot Cybersecurity
por: Mayoral-Vilches, Víctor
Publicado: (2025)
por: Mayoral-Vilches, Víctor
Publicado: (2025)
Towards Cybersecurity SuperIntelligence (CSI): What's the best harness for cybersecurity?
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
CAI: An Open, Bug Bounty-Ready Cybersecurity AI
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
Cybersecurity AI: Hacking Consumer Robots in the AI Era
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
Cybersecurity AI: Hacking the AI Hackers via Prompt Injection
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
Towards Cybersecurity Superintelligence: from AI-guided humans to human-guided AI
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026)
Cybersecurity AI in OT: Insights from an AI Top-10 Ranker in the Dragos OT CTF 2025
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
Cybersecurity AI: The World's Top AI Agent for Security Capture-the-Flag (CTF)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
WebAssembly and Security: a review
por: Perrone, Gaetano, et al.
Publicado: (2024)
por: Perrone, Gaetano, et al.
Publicado: (2024)
CAI Fluency: A Framework for Cybersecurity AI Fluency
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025)
GPT-5 at CTFs: Case Studies From Top-Tier Cybersecurity Events
por: Reworr, et al.
Publicado: (2025)
por: Reworr, et al.
Publicado: (2025)
Dockerized Android: a container-based platform to build mobile Android scenarios for Cyber Ranges
por: Capone, Daniele, et al.
Publicado: (2022)
por: Capone, Daniele, et al.
Publicado: (2022)
PentestGPT: An LLM-empowered Automatic Penetration Testing Tool
por: Deng, Gelei, et al.
Publicado: (2023)
por: Deng, Gelei, et al.
Publicado: (2023)
ExploitWP2Docker: a Platform for Automating the Generation of Vulnerable WordPress Environments for Cyber Ranges
por: Caturano, Francesco, et al.
Publicado: (2022)
por: Caturano, Francesco, et al.
Publicado: (2022)
Cybersecurity Defenses: Exploration of CVE Types through Attack Descriptions
por: Othman, Refat, et al.
Publicado: (2024)
por: Othman, Refat, et al.
Publicado: (2024)
Generative AI in Cybersecurity
por: Metta, Shivani, et al.
Publicado: (2024)
por: Metta, Shivani, et al.
Publicado: (2024)
Leveraging AI to optimize website structure discovery during Penetration Testing
por: Antonelli, Diego, et al.
Publicado: (2021)
por: Antonelli, Diego, et al.
Publicado: (2021)
Space Cybersecurity Norms
por: Sharfman, Peter, et al.
Publicado: (2023)
por: Sharfman, Peter, et al.
Publicado: (2023)
Cybersecurity as a Service
por: Morris, John, et al.
Publicado: (2024)
por: Morris, John, et al.
Publicado: (2024)
AgentCyTE: Leveraging Agentic AI to Generate Cybersecurity Training & Experimentation Scenarios
por: Rodriguez, Ana M., et al.
Publicado: (2025)
por: Rodriguez, Ana M., et al.
Publicado: (2025)
AI-Driven Cybersecurity Threat Detection: Building Resilient Defense Systems Using Predictive Analytics
por: Das, Biswajit Chandra, et al.
Publicado: (2025)
por: Das, Biswajit Chandra, et al.
Publicado: (2025)
Review of Generative AI Methods in Cybersecurity
por: Yigit, Yagmur, et al.
Publicado: (2024)
por: Yigit, Yagmur, et al.
Publicado: (2024)
Does Johnny Get the Message? Evaluating Cybersecurity Notifications for Everyday Users
por: Jüttner, Victor, et al.
Publicado: (2025)
por: Jüttner, Victor, et al.
Publicado: (2025)
Agentic AI for Cybersecurity: A Meta-Cognitive Architecture for Governable Autonomy
por: Kojukhov, Andrei, et al.
Publicado: (2026)
por: Kojukhov, Andrei, et al.
Publicado: (2026)
AI-Driven Cybersecurity Threats: A Survey of Emerging Risks and Defensive Strategies
por: Erukude, Sai Teja, et al.
Publicado: (2026)
por: Erukude, Sai Teja, et al.
Publicado: (2026)
Understanding Human-AI Collaboration in Cybersecurity Competitions
por: Tang, Tingxuan, et al.
Publicado: (2026)
por: Tang, Tingxuan, et al.
Publicado: (2026)
The Evolution of Agentic AI in Cybersecurity: From Single LLM Reasoners to Multi-Agent Systems and Autonomous Pipelines
por: Vinay, Vaishali
Publicado: (2025)
por: Vinay, Vaishali
Publicado: (2025)
Evaluating LLM Generated Detection Rules in Cybersecurity
por: Bertiger, Anna, et al.
Publicado: (2025)
por: Bertiger, Anna, et al.
Publicado: (2025)
A Survey of Agentic AI and Cybersecurity: Challenges, Opportunities and Use-case Prototypes
por: Lazer, Sahaya Jestus, et al.
Publicado: (2026)
por: Lazer, Sahaya Jestus, et al.
Publicado: (2026)
Integrative Approaches in Cybersecurity and AI
por: Omar, Marwan
Publicado: (2024)
por: Omar, Marwan
Publicado: (2024)
Thwarting Cybersecurity Attacks with Explainable Concept Drift
por: Shaer, Ibrahim, et al.
Publicado: (2024)
por: Shaer, Ibrahim, et al.
Publicado: (2024)
Adaptive Deception Framework with Behavioral Analysis for Enhanced Cybersecurity Defense
por: AL-Zahrani, Basil Abdullah
Publicado: (2025)
por: AL-Zahrani, Basil Abdullah
Publicado: (2025)
What is Cybersecurity in Space?
por: Mattar, Charbel, et al.
Publicado: (2025)
por: Mattar, Charbel, et al.
Publicado: (2025)
Game-Theoretic Modeling of Stealthy Intrusion Defense against MDP-Based Attackers
por: Kouam, Willie, et al.
Publicado: (2026)
por: Kouam, Willie, et al.
Publicado: (2026)
EAGER: Edge-Aligned LLM Defense for Robust, Efficient, and Accurate Cybersecurity Question Answering
por: Gungor, Onat, et al.
Publicado: (2025)
por: Gungor, Onat, et al.
Publicado: (2025)
Ejemplares similares
-
Cybersecurity AI: A Game-Theoretic AI for Guiding Attack and Defense
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2026) -
Cybersecurity AI: The Dangerous Gap Between Automation and Autonomy
por: Mayoral-Vilches, Víctor
Publicado: (2025) -
Cybersecurity AI: Humanoid Robots as Attack Vectors
por: Mayoral-Vilches, Víctor, et al.
Publicado: (2025) -
Cybersecurity AI Benchmark (CAIBench): A Meta-Benchmark for Evaluating Cybersecurity AI Agents
por: Sanz-Gómez, María, et al.
Publicado: (2025) -
The Cybersecurity of a Humanoid Robot
por: Mayoral-Vilches, Víctor
Publicado: (2025)