Autonomous LLM Agents & CTFs: A Second Look
Fuente:
arXiv
Saved in:
| Main Authors: | Bouchari, Youness, Boffa, Matteo, Mellia, Marco, Drago, Idilio, Bui, Thanh Minh, Rossi, Dario |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Agentic Honeynet Configuration
by: Mirra, Federico, et al.
Published: (2026)
by: Mirra, Federico, et al.
Published: (2026)
Improving Generalization on Cybersecurity Tasks with Multi-Modal Contrastive Learning
by: Huang, Jianan, et al.
Published: (2026)
by: Huang, Jianan, et al.
Published: (2026)
AutoPenBench: Benchmarking Generative Agents for Penetration Testing
by: Gioacchini, Luca, et al.
Published: (2024)
by: Gioacchini, Luca, et al.
Published: (2024)
LogPrécis: Unleashing Language Models for Automated Malicious Log Analysis
by: Boffa, Matteo, et al.
Published: (2023)
by: Boffa, Matteo, et al.
Published: (2023)
CyberSleuth: Autonomous Blue-Team LLM Agent for Web Attack Forensics
by: Fumero, Stefano, et al.
Published: (2025)
by: Fumero, Stefano, et al.
Published: (2025)
Hacking CTFs with Plain Agents
by: Turtayev, Rustem, et al.
Published: (2024)
by: Turtayev, Rustem, et al.
Published: (2024)
ARBITER: AI-Driven Filtering for Role-Based Access Control
by: Lorenzo, Michele, et al.
Published: (2025)
by: Lorenzo, Michele, et al.
Published: (2025)
LLM Agents can Autonomously Hack Websites
by: Fang, Richard, et al.
Published: (2024)
by: Fang, Richard, et al.
Published: (2024)
AutoPentest: Enhancing Vulnerability Management With Autonomous LLM Agents
by: Henke, Julius
Published: (2025)
by: Henke, Julius
Published: (2025)
LLM Agents can Autonomously Exploit One-day Vulnerabilities
by: Fang, Richard, et al.
Published: (2024)
by: Fang, Richard, et al.
Published: (2024)
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks
by: Song, Ruoyu, et al.
Published: (2024)
by: Song, Ruoyu, et al.
Published: (2024)
Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
by: Deng, Xinhao, et al.
Published: (2026)
by: Deng, Xinhao, et al.
Published: (2026)
ARACNE: An LLM-Based Autonomous Shell Pentesting Agent
by: Nieponice, Tomas, et al.
Published: (2025)
by: Nieponice, Tomas, et al.
Published: (2025)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
by: Wu, Yixin, et al.
Published: (2025)
by: Wu, Yixin, et al.
Published: (2025)
Generic Multi-modal Representation Learning for Network Traffic Analysis
by: Gioacchini, Luca, et al.
Published: (2024)
by: Gioacchini, Luca, et al.
Published: (2024)
ChamaleoNet: Programmable Passive Probe for Enhanced Visibility on Erroneous Traffic
by: Wang, Zhihao, et al.
Published: (2025)
by: Wang, Zhihao, et al.
Published: (2025)
PocketAgents: A Manifest-Driven Library of Autonomous Defense Agents
by: Barbieri, Sidnei, et al.
Published: (2026)
by: Barbieri, Sidnei, et al.
Published: (2026)
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
by: Zhang, Yixiang, et al.
Published: (2026)
by: Zhang, Yixiang, et al.
Published: (2026)
A First Look At Efficient And Secure On-Device LLM Inference Against KV Leakage
by: Yang, Huan, et al.
Published: (2024)
by: Yang, Huan, et al.
Published: (2024)
LiaisonAgent: An Multi-Agent Framework for Autonomous Risk Investigation and Governance
by: Tang, Chuanming, et al.
Published: (2026)
by: Tang, Chuanming, et al.
Published: (2026)
APEX: Agent Payment Execution with Policy for Autonomous Agent API Access
by: Uddin, Mohd Safwan, et al.
Published: (2026)
by: Uddin, Mohd Safwan, et al.
Published: (2026)
CheatAgent: Attacking LLM-Empowered Recommender Systems via LLM Agent
by: Ning, Liang-bo, et al.
Published: (2025)
by: Ning, Liang-bo, et al.
Published: (2025)
Agentic JWT: A Secure Delegation Protocol for Autonomous AI Agents
by: Goswami, Abhishek
Published: (2025)
by: Goswami, Abhishek
Published: (2025)
Measuring Safety Alignment Effects in Autonomous Security Agents
by: David, Isaac, et al.
Published: (2026)
by: David, Isaac, et al.
Published: (2026)
Agent Audit: A Security Analysis System for LLM Agent Applications
by: Zhang, Haiyue, et al.
Published: (2026)
by: Zhang, Haiyue, et al.
Published: (2026)
A Framework for Formalizing LLM Agent Security
by: Siu, Vincent, et al.
Published: (2026)
by: Siu, Vincent, et al.
Published: (2026)
Caging the Agents: A Zero Trust Security Architecture for Autonomous AI in Healthcare
by: Maiti, Saikat
Published: (2026)
by: Maiti, Saikat
Published: (2026)
CybORG++: An Enhanced Gym for the Development of Autonomous Cyber Agents
by: Emerson, Harry, et al.
Published: (2024)
by: Emerson, Harry, et al.
Published: (2024)
Agent-Sentry: Bounding LLM Agents via Execution Provenance
by: Sequeira, Rohan, et al.
Published: (2026)
by: Sequeira, Rohan, et al.
Published: (2026)
LLM Agent Honeypot: Monitoring AI Hacking Agents in the Wild
by: Reworr, et al.
Published: (2024)
by: Reworr, et al.
Published: (2024)
ZeroDayBench: Evaluating LLM Agents on Unseen Zero-Day Vulnerabilities for Cyberdefense
by: Lau, Nancy, et al.
Published: (2026)
by: Lau, Nancy, et al.
Published: (2026)
Sequential Behavioral Watermarking for LLM Agents
by: An, Hyeseon, et al.
Published: (2026)
by: An, Hyeseon, et al.
Published: (2026)
RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents
by: Black, Sid, et al.
Published: (2025)
by: Black, Sid, et al.
Published: (2025)
Incalmo: An Autonomous LLM-assisted System for Red Teaming Multi-Host Networks
by: Singer, Brian, et al.
Published: (2025)
by: Singer, Brian, et al.
Published: (2025)
SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents
by: Cheng, Hao, et al.
Published: (2026)
by: Cheng, Hao, et al.
Published: (2026)
Before the Tool Call: Deterministic Pre-Action Authorization for Autonomous AI Agents
by: Uchibeke, Uchi
Published: (2026)
by: Uchibeke, Uchi
Published: (2026)
Poster: ClawdGo: Endogenous Security Awareness Training for Autonomous AI Agents
by: Li, Jiaqi, et al.
Published: (2026)
by: Li, Jiaqi, et al.
Published: (2026)
Hacking Back the AI-Hacker: Prompt Injection as a Defense Against LLM-driven Cyberattacks
by: Pasquini, Dario, et al.
Published: (2024)
by: Pasquini, Dario, et al.
Published: (2024)
LLM Agents Should Employ Security Principles
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
Unveiling Privacy Risks in LLM Agent Memory
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
Similar Items
-
Towards Agentic Honeynet Configuration
by: Mirra, Federico, et al.
Published: (2026) -
Improving Generalization on Cybersecurity Tasks with Multi-Modal Contrastive Learning
by: Huang, Jianan, et al.
Published: (2026) -
AutoPenBench: Benchmarking Generative Agents for Penetration Testing
by: Gioacchini, Luca, et al.
Published: (2024) -
LogPrécis: Unleashing Language Models for Automated Malicious Log Analysis
by: Boffa, Matteo, et al.
Published: (2023) -
CyberSleuth: Autonomous Blue-Team LLM Agent for Web Attack Forensics
by: Fumero, Stefano, et al.
Published: (2025)