PentestAgent: Incorporating LLM Agents to Automated Penetration Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Xiangmin, Wang, Lingzhi, Li, Zhenyuan, Chen, Yan, Zhao, Wencheng, Sun, Dawei, Wang, Jiashui, Ruan, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automated Penetration Testing with LLM Agents and Classical Planning
by: Wang, Lingzhi, et al.
Published: (2025)
by: Wang, Lingzhi, et al.
Published: (2025)
AEAS: Actionable Exploit Assessment System
by: Shen, Xiangmin, et al.
Published: (2025)
by: Shen, Xiangmin, et al.
Published: (2025)
From Sands to Mansions: Towards Automated Cyberattack Emulation with Classical Planning and Large Language Models
by: Wang, Lingzhi, et al.
Published: (2024)
by: Wang, Lingzhi, et al.
Published: (2024)
Incorporating Gradients to Rules: Towards Lightweight, Adaptive Provenance-based Intrusion Detection
by: Wang, Lingzhi, et al.
Published: (2024)
by: Wang, Lingzhi, et al.
Published: (2024)
Decoding the MITRE Engenuity ATT&CK Enterprise Evaluation: An Analysis of EDR Performance in Real-World Environments
by: Shen, Xiangmin, et al.
Published: (2024)
by: Shen, Xiangmin, et al.
Published: (2024)
AutoPentester: An LLM Agent-based Framework for Automated Pentesting
by: Ginige, Yasod, et al.
Published: (2025)
by: Ginige, Yasod, et al.
Published: (2025)
PentestGPT: An LLM-empowered Automatic Penetration Testing Tool
by: Deng, Gelei, et al.
Published: (2023)
by: Deng, Gelei, et al.
Published: (2023)
PentestMCP: A Toolkit for Agentic Penetration Testing
by: Ezetta, Zachary, et al.
Published: (2025)
by: Ezetta, Zachary, et al.
Published: (2025)
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
by: Yang, Ruozhao, et al.
Published: (2025)
by: Yang, Ruozhao, et al.
Published: (2025)
Beyond Input Guardrails: Reconstructing Cross-Agent Semantic Flows for Execution-Aware Attack Detection
by: Wei, Yangyang, et al.
Published: (2026)
by: Wei, Yangyang, et al.
Published: (2026)
PARIS: A Practical, Adaptive Trace-Fetching and Real-Time Malicious Behavior Detection System
by: Wang, Jian, et al.
Published: (2024)
by: Wang, Jian, et al.
Published: (2024)
ARACNE: An LLM-Based Autonomous Shell Pentesting Agent
by: Nieponice, Tomas, et al.
Published: (2025)
by: Nieponice, Tomas, et al.
Published: (2025)
AutoPentest: Enhancing Vulnerability Management With Autonomous LLM Agents
by: Henke, Julius
Published: (2025)
by: Henke, Julius
Published: (2025)
Shell or Nothing: Real-World Benchmarks and Memory-Activated Agents for Automated Penetration Testing
by: Mai, Wuyuao, et al.
Published: (2025)
by: Mai, Wuyuao, et al.
Published: (2025)
RapidPen: Fully Automated IP-to-Shell Penetration Testing with LLM-based Agents
by: Nakatani, Sho
Published: (2025)
by: Nakatani, Sho
Published: (2025)
Towards Automated Pentesting with Large Language Models
by: Bessa, Ricardo, et al.
Published: (2026)
by: Bessa, Ricardo, et al.
Published: (2026)
APT-Agent: Automated Penetration Testing using Large Language Models
by: Li, William Guanting, et al.
Published: (2026)
by: Li, William Guanting, et al.
Published: (2026)
PenHeal: A Two-Stage LLM Framework for Automated Pentesting and Optimal Remediation
by: Huang, Junjie, et al.
Published: (2024)
by: Huang, Junjie, et al.
Published: (2024)
Marlin: Knowledge-Driven Analysis of Provenance Graphs for Efficient and Robust Detection of Cyber Attacks
by: Li, Zhenyuan, et al.
Published: (2024)
by: Li, Zhenyuan, et al.
Published: (2024)
Vulnerability Mitigation System (VMS): LLM Agent and Evaluation Framework for Autonomous Penetration Testing
by: Abdulzada, Farzana
Published: (2025)
by: Abdulzada, Farzana
Published: (2025)
Ethics Statements in Autonomous Penetration-Testing Agent Research
by: Happe, Andreas, et al.
Published: (2025)
by: Happe, Andreas, et al.
Published: (2025)
PentestJudge: Judging Agent Behavior Against Operational Requirements
by: Caldwell, Shane, et al.
Published: (2025)
by: Caldwell, Shane, et al.
Published: (2025)
From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World
by: Conde, Pedro, et al.
Published: (2026)
by: Conde, Pedro, et al.
Published: (2026)
What Makes a Good LLM Agent for Real-world Penetration Testing?
by: Deng, Gelei, et al.
Published: (2026)
by: Deng, Gelei, et al.
Published: (2026)
Hackers or Hallucinators? A Comprehensive Analysis of LLM-Based Automated Penetration Testing
by: Peng, Jiaren, et al.
Published: (2026)
by: Peng, Jiaren, et al.
Published: (2026)
Multi-Agent Penetration Testing AI for the Web
by: David, Isaac, et al.
Published: (2025)
by: David, Isaac, et al.
Published: (2025)
AutoBackdoor: Automating Backdoor Attacks via LLM Agents
by: Li, Yige, et al.
Published: (2025)
by: Li, Yige, et al.
Published: (2025)
AWE: Adaptive Agents for Dynamic Web Penetration Testing
by: Jaswal, Akshat Singh, et al.
Published: (2026)
by: Jaswal, Akshat Singh, et al.
Published: (2026)
BreachSeek: A Multi-Agent Automated Penetration Tester
by: Alshehri, Ibrahim, et al.
Published: (2024)
by: Alshehri, Ibrahim, et al.
Published: (2024)
TraceAegis: Securing LLM-Based Agents via Hierarchical and Behavioral Anomaly Detection
by: Liu, Jiahao, et al.
Published: (2025)
by: Liu, Jiahao, et al.
Published: (2025)
Rethinking Side-Channel Analysis: Automated Discovery and Analysis of Side-Channel Leakage with LLM-Assisted Agents
by: Xu, Zhen, et al.
Published: (2026)
by: Xu, Zhen, et al.
Published: (2026)
Towards Automated Penetration Testing: Introducing LLM Benchmark, Analysis, and Improvements
by: Isozaki, Isamu, et al.
Published: (2024)
by: Isozaki, Isamu, et al.
Published: (2024)
KryptoPilot: An Open-World Knowledge-Augmented LLM Agent for Automated Cryptographic Exploitation
by: Liu, Xiaonan, et al.
Published: (2026)
by: Liu, Xiaonan, et al.
Published: (2026)
Lessons from Penetration Tests on Large-Scale Agent Systems
by: Eykholt, Kevin, et al.
Published: (2026)
by: Eykholt, Kevin, et al.
Published: (2026)
AutoPenBench: Benchmarking Generative Agents for Penetration Testing
by: Gioacchini, Luca, et al.
Published: (2024)
by: Gioacchini, Luca, et al.
Published: (2024)
Automated Penetration Testing: Formalization and Realization
by: Skandylas, Charilaos, et al.
Published: (2024)
by: Skandylas, Charilaos, et al.
Published: (2024)
LLM Agents for Automated Web Vulnerability Reproduction: Are We There Yet?
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
Penetration Testing for System Security: Methods and Practical Approaches
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
Reframing LLM Agent Security as an Agent-Human Interaction Problem
by: Wang, Peiran, et al.
Published: (2026)
by: Wang, Peiran, et al.
Published: (2026)
Reinforcement Learning for Automated Cybersecurity Penetration Testing
by: López-Montero, Daniel, et al.
Published: (2025)
by: López-Montero, Daniel, et al.
Published: (2025)
Similar Items
-
Automated Penetration Testing with LLM Agents and Classical Planning
by: Wang, Lingzhi, et al.
Published: (2025) -
AEAS: Actionable Exploit Assessment System
by: Shen, Xiangmin, et al.
Published: (2025) -
From Sands to Mansions: Towards Automated Cyberattack Emulation with Classical Planning and Large Language Models
by: Wang, Lingzhi, et al.
Published: (2024) -
Incorporating Gradients to Rules: Towards Lightweight, Adaptive Provenance-based Intrusion Detection
by: Wang, Lingzhi, et al.
Published: (2024) -
Decoding the MITRE Engenuity ATT&CK Enterprise Evaluation: An Analysis of EDR Performance in Real-World Environments
by: Shen, Xiangmin, et al.
Published: (2024)