PentestMCP: A Toolkit for Agentic Penetration Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Ezetta, Zachary, Feng, Wu-chang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
by: Yang, Ruozhao, et al.
Published: (2025)
by: Yang, Ruozhao, et al.
Published: (2025)
AutoPentester: An LLM Agent-based Framework for Automated Pentesting
by: Ginige, Yasod, et al.
Published: (2025)
by: Ginige, Yasod, et al.
Published: (2025)
Penetration Testing of Agentic AI: A Comparative Security Analysis Across Models and Frameworks
by: Nguyen, Viet K., et al.
Published: (2025)
by: Nguyen, Viet K., et al.
Published: (2025)
A Preliminary Study on Using Large Language Models in Software Pentesting
by: Shashwat, Kumar, et al.
Published: (2024)
by: Shashwat, Kumar, et al.
Published: (2024)
PentestJudge: Judging Agent Behavior Against Operational Requirements
by: Caldwell, Shane, et al.
Published: (2025)
by: Caldwell, Shane, et al.
Published: (2025)
AutoPentest: Enhancing Vulnerability Management With Autonomous LLM Agents
by: Henke, Julius
Published: (2025)
by: Henke, Julius
Published: (2025)
From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World
by: Conde, Pedro, et al.
Published: (2026)
by: Conde, Pedro, et al.
Published: (2026)
PentestAgent: Incorporating LLM Agents to Automated Penetration Testing
by: Shen, Xiangmin, et al.
Published: (2024)
by: Shen, Xiangmin, et al.
Published: (2024)
ARACNE: An LLM-Based Autonomous Shell Pentesting Agent
by: Nieponice, Tomas, et al.
Published: (2025)
by: Nieponice, Tomas, et al.
Published: (2025)
Multi-Agent Penetration Testing AI for the Web
by: David, Isaac, et al.
Published: (2025)
by: David, Isaac, et al.
Published: (2025)
Reinforcement Learning for Automated Cybersecurity Penetration Testing
by: López-Montero, Daniel, et al.
Published: (2025)
by: López-Montero, Daniel, et al.
Published: (2025)
Generative Artificial Intelligence-Supported Pentesting: A Comparison between Claude Opus, GPT-4, and Copilot
by: Martínez, Antonio López, et al.
Published: (2025)
by: Martínez, Antonio López, et al.
Published: (2025)
MCP-Guard: A Multi-Stage Defense-in-Depth Framework for Securing Model Context Protocol in Agentic AI
by: Xing, Wenpeng, et al.
Published: (2025)
by: Xing, Wenpeng, et al.
Published: (2025)
AWE: Adaptive Agents for Dynamic Web Penetration Testing
by: Jaswal, Akshat Singh, et al.
Published: (2026)
by: Jaswal, Akshat Singh, et al.
Published: (2026)
MCP-ITP: An Automated Framework for Implicit Tool Poisoning in MCP
by: Li, Ruiqi, et al.
Published: (2026)
by: Li, Ruiqi, et al.
Published: (2026)
PentestGPT: An LLM-empowered Automatic Penetration Testing Tool
by: Deng, Gelei, et al.
Published: (2023)
by: Deng, Gelei, et al.
Published: (2023)
MCP Guardian: A Security-First Layer for Safeguarding MCP-Based AI System
by: Kumar, Sonu, et al.
Published: (2025)
by: Kumar, Sonu, et al.
Published: (2025)
MCP-in-SoS: Risk assessment framework for open-source MCP servers
by: Kumar, Pratyay, et al.
Published: (2026)
by: Kumar, Pratyay, et al.
Published: (2026)
Lessons from Penetration Tests on Large-Scale Agent Systems
by: Eykholt, Kevin, et al.
Published: (2026)
by: Eykholt, Kevin, et al.
Published: (2026)
AutoPenBench: Benchmarking Generative Agents for Penetration Testing
by: Gioacchini, Luca, et al.
Published: (2024)
by: Gioacchini, Luca, et al.
Published: (2024)
Hacking, The Lazy Way: LLM Augmented Pentesting
by: Goyal, Dhruva, et al.
Published: (2024)
by: Goyal, Dhruva, et al.
Published: (2024)
Pen-Strategist: A Reasoning Framework for Penetration Testing Strategy Formation and Analysis
by: Ginige, Yasod, et al.
Published: (2026)
by: Ginige, Yasod, et al.
Published: (2026)
BlackIce: A Containerized Red Teaming Toolkit for AI Security Testing
by: Kaplan, Caelin, et al.
Published: (2025)
by: Kaplan, Caelin, et al.
Published: (2025)
Towards Automated Penetration Testing: Introducing LLM Benchmark, Analysis, and Improvements
by: Isozaki, Isamu, et al.
Published: (2024)
by: Isozaki, Isamu, et al.
Published: (2024)
APT-Agent: Automated Penetration Testing using Large Language Models
by: Li, William Guanting, et al.
Published: (2026)
by: Li, William Guanting, et al.
Published: (2026)
MCPGuard : Automatically Detecting Vulnerabilities in MCP Servers
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
AutoPT: How Far Are We from the End2End Automated Web Penetration Testing?
by: Wu, Benlong, et al.
Published: (2024)
by: Wu, Benlong, et al.
Published: (2024)
RapidPen: Fully Automated IP-to-Shell Penetration Testing with LLM-based Agents
by: Nakatani, Sho
Published: (2025)
by: Nakatani, Sho
Published: (2025)
Evaluation of Reinforcement Learning for Autonomous Penetration Testing using A3C, Q-learning and DQN
by: Becker, Norman, et al.
Published: (2024)
by: Becker, Norman, et al.
Published: (2024)
Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models
by: Belkhiter, Yannis, et al.
Published: (2026)
by: Belkhiter, Yannis, et al.
Published: (2026)
Cochise: A Reference Harness for Autonomous Penetration Testing
by: Happe, Andreas, et al.
Published: (2026)
by: Happe, Andreas, et al.
Published: (2026)
xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models
by: Luong, Phung Duc, et al.
Published: (2025)
by: Luong, Phung Duc, et al.
Published: (2025)
Incorporation of Verifier Functionality in the Software for Operations and Network Attack Results Review and the Autonomous Penetration Testing System
by: Milbrath, Jordan, et al.
Published: (2024)
by: Milbrath, Jordan, et al.
Published: (2024)
Simplified and Secure MCP Gateways for Enterprise AI Integration
by: Brett, Ivo
Published: (2025)
by: Brett, Ivo
Published: (2025)
From Tool Orchestration to Code Execution: A Study of MCP Design Choices
by: Felendler, Yuval, et al.
Published: (2026)
by: Felendler, Yuval, et al.
Published: (2026)
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
by: Lin, Justin W., et al.
Published: (2025)
by: Lin, Justin W., et al.
Published: (2025)
How Reliable Are AI Attackers Against a Fixed Vulnerable Target? A 400-Run Empirical Study of LLM Penetration Testing Consistency
by: Erdem, Galip Tolga
Published: (2026)
by: Erdem, Galip Tolga
Published: (2026)
BreachSeek: A Multi-Agent Automated Penetration Tester
by: Alshehri, Ibrahim, et al.
Published: (2024)
by: Alshehri, Ibrahim, et al.
Published: (2024)
We Urgently Need Privilege Management in MCP: A Measurement of API Usage in MCP Ecosystems
by: Li, Zhihao, et al.
Published: (2025)
by: Li, Zhihao, et al.
Published: (2025)
Hackers or Hallucinators? A Comprehensive Analysis of LLM-Based Automated Penetration Testing
by: Peng, Jiaren, et al.
Published: (2026)
by: Peng, Jiaren, et al.
Published: (2026)
Similar Items
-
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
by: Yang, Ruozhao, et al.
Published: (2025) -
AutoPentester: An LLM Agent-based Framework for Automated Pentesting
by: Ginige, Yasod, et al.
Published: (2025) -
Penetration Testing of Agentic AI: A Comparative Security Analysis Across Models and Frameworks
by: Nguyen, Viet K., et al.
Published: (2025) -
A Preliminary Study on Using Large Language Models in Software Pentesting
by: Shashwat, Kumar, et al.
Published: (2024) -
PentestJudge: Judging Agent Behavior Against Operational Requirements
by: Caldwell, Shane, et al.
Published: (2025)