Automation-Exploit: A Multi-Agent LLM Framework for Adaptive Offensive Security with Digital Twin-Based Risk-Mitigated Exploitation
Fuente:
arXiv
Saved in:
| Main Authors: | Andreucci, Biagio, Castiglione, Arcangelo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What is the AGI in Offensive Security ?
by: Cho, Youngwoong
Published: (2026)
by: Cho, Youngwoong
Published: (2026)
Dynamic Risk Assessments for Offensive Cybersecurity Agents
by: Wei, Boyi, et al.
Published: (2025)
by: Wei, Boyi, et al.
Published: (2025)
From CVE Entries to Verifiable Exploits: An Automated Multi-Agent Framework for Reproducing CVEs
by: Ullah, Saad, et al.
Published: (2025)
by: Ullah, Saad, et al.
Published: (2025)
KryptoPilot: An Open-World Knowledge-Augmented LLM Agent for Automated Cryptographic Exploitation
by: Liu, Xiaonan, et al.
Published: (2026)
by: Liu, Xiaonan, et al.
Published: (2026)
Exploiting AI for Attacks: On the Interplay between Adversarial AI and Offensive AI
by: Schröer, Saskia Laura, et al.
Published: (2025)
by: Schröer, Saskia Laura, et al.
Published: (2025)
Self-Adaptive Multi-Agent LLM-Based Security Pattern Selection for IoT Systems
by: Jamshidi, Saeid, et al.
Published: (2026)
by: Jamshidi, Saeid, et al.
Published: (2026)
Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
by: He, Pengfei, et al.
Published: (2026)
by: He, Pengfei, et al.
Published: (2026)
Artificial Intelligence as the New Hacker: Developing Agents for Offensive Security
by: Valencia, Leroy Jacob
Published: (2024)
by: Valencia, Leroy Jacob
Published: (2024)
An Empirical Evaluation of LLMs for Solving Offensive Security Challenges
by: Shao, Minghao, et al.
Published: (2024)
by: Shao, Minghao, et al.
Published: (2024)
Adaptive Exploit Generation against Security Devices and Security APIs
by: Künnemann, Robert, et al.
Published: (2024)
by: Künnemann, Robert, et al.
Published: (2024)
Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark
by: Shao, Minghao, et al.
Published: (2025)
by: Shao, Minghao, et al.
Published: (2025)
D-CIPHER: Dynamic Collaborative Intelligent Multi-Agent System with Planner and Heterogeneous Executors for Offensive Security
by: Udeshi, Meet, et al.
Published: (2025)
by: Udeshi, Meet, et al.
Published: (2025)
Benchmarking Practices in LLM-driven Offensive Security: Testbeds, Metrics, and Experiment Design
by: Happe, Andreas, et al.
Published: (2025)
by: Happe, Andreas, et al.
Published: (2025)
Taxonomy, Evaluation and Exploitation of IPI-Centric LLM Agent Defense Frameworks
by: Ji, Zimo, et al.
Published: (2025)
by: Ji, Zimo, et al.
Published: (2025)
Prompt Engineering vs. Fine-Tuning for LLM-Based Vulnerability Detection in Solana and Algorand Smart Contracts
by: Boi, Biagio, et al.
Published: (2025)
by: Boi, Biagio, et al.
Published: (2025)
Poster: Machine Learning for Vulnerability Detection as Target Oracle in Automated Fuzz Driver Generation
by: Castiglione, Gianpietro, et al.
Published: (2025)
by: Castiglione, Gianpietro, et al.
Published: (2025)
Integrating Cybersecurity Frameworks into IT Security: A Comprehensive Analysis of Threat Mitigation Strategies and Adaptive Technologies
by: Lokare, Amit, et al.
Published: (2025)
by: Lokare, Amit, et al.
Published: (2025)
The Infinite Mutation Engine? Measuring Polymorphism in LLM-Generated Offensive Code
by: Hortea, Gabriel, et al.
Published: (2026)
by: Hortea, Gabriel, et al.
Published: (2026)
Offensive Security for AI Systems: Concepts, Practices, and Applications
by: Harguess, Josh, et al.
Published: (2025)
by: Harguess, Josh, et al.
Published: (2025)
Mapping the Exploitation Surface: A 10,000-Trial Taxonomy of What Makes LLM Agents Exploit Vulnerabilities
by: Mouzouni, Charafeddine
Published: (2026)
by: Mouzouni, Charafeddine
Published: (2026)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
by: Rani, Nanda, et al.
Published: (2026)
by: Rani, Nanda, et al.
Published: (2026)
A Framework for Formalizing LLM Agent Security
by: Siu, Vincent, et al.
Published: (2026)
by: Siu, Vincent, et al.
Published: (2026)
Advancing Security with Digital Twins: A Comprehensive Survey
by: Airehenbuwa, Blessing, et al.
Published: (2025)
by: Airehenbuwa, Blessing, et al.
Published: (2025)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
PhishLumos: An Adaptive Multi-Agent System for Proactive Phishing Campaign Mitigation
by: Chiba, Daiki, et al.
Published: (2025)
by: Chiba, Daiki, et al.
Published: (2025)
Machine Learning for Healthcare-IoT Security: A Review and Risk Mitigation
by: Khatun, Mirza Akhi, et al.
Published: (2024)
by: Khatun, Mirza Akhi, et al.
Published: (2024)
HonestCyberEval: An AI Cyber Risk Benchmark for Automated Software Exploitation
by: Ristea, Dan, et al.
Published: (2024)
by: Ristea, Dan, et al.
Published: (2024)
PhishDebate: An LLM-Based Multi-Agent Framework for Phishing Website Detection
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
The Trust Paradox in LLM-Based Multi-Agent Systems: When Collaboration Becomes a Security Vulnerability
by: Xu, Zijie, et al.
Published: (2025)
by: Xu, Zijie, et al.
Published: (2025)
VeriSBOM: Secure and Verifiable SBOM Sharing Via Zero-Knowledge Proofs
by: Castiglione, Gianpietro, et al.
Published: (2026)
by: Castiglione, Gianpietro, et al.
Published: (2026)
Offensive Robot Cybersecurity
by: Mayoral-Vilches, Víctor
Published: (2025)
by: Mayoral-Vilches, Víctor
Published: (2025)
PoC-Adapt: Semantic-Aware Automated Vulnerability Reproduction with LLM Multi-Agents and Reinforcement Learning-Driven Adaptive Policy
by: Duy, Phan The, et al.
Published: (2026)
by: Duy, Phan The, et al.
Published: (2026)
Securing RAG: A Risk Assessment and Mitigation Framework
by: Ammann, Lukas, et al.
Published: (2025)
by: Ammann, Lukas, et al.
Published: (2025)
Don't Trust Your Upstream: Exploiting LLM Multi-Agent System via Topology-Guided Adversarial Propagation
by: Liang, Ruichao, et al.
Published: (2025)
by: Liang, Ruichao, et al.
Published: (2025)
Vulnerability Mitigation System (VMS): LLM Agent and Evaluation Framework for Autonomous Penetration Testing
by: Abdulzada, Farzana
Published: (2025)
by: Abdulzada, Farzana
Published: (2025)
W3ID: A Quantum Computing-Secure Digital Identity System Redefining Standards for Web3 and Digital Twins
by: Yun, Joseph, et al.
Published: (2025)
by: Yun, Joseph, et al.
Published: (2025)
HySecTwin: A Knowledge-Driven Digital Twin Framework Augmented with Hybrid Reasoning for Cyber-Physical Systems
by: Holmes, David, et al.
Published: (2026)
by: Holmes, David, et al.
Published: (2026)
Atomicity for Agents: Exposing, Exploiting, and Mitigating TOCTOU Vulnerabilities in Browser-Use Agents
by: Jiang, Linxi, et al.
Published: (2026)
by: Jiang, Linxi, et al.
Published: (2026)
JPRO: Automated Multimodal Jailbreaking via Multi-Agent Collaboration Framework
by: Zhou, Yuxuan, et al.
Published: (2025)
by: Zhou, Yuxuan, et al.
Published: (2025)
Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
by: Deng, Xinhao, et al.
Published: (2026)
by: Deng, Xinhao, et al.
Published: (2026)
Similar Items
-
What is the AGI in Offensive Security ?
by: Cho, Youngwoong
Published: (2026) -
Dynamic Risk Assessments for Offensive Cybersecurity Agents
by: Wei, Boyi, et al.
Published: (2025) -
From CVE Entries to Verifiable Exploits: An Automated Multi-Agent Framework for Reproducing CVEs
by: Ullah, Saad, et al.
Published: (2025) -
KryptoPilot: An Open-World Knowledge-Augmented LLM Agent for Automated Cryptographic Exploitation
by: Liu, Xiaonan, et al.
Published: (2026) -
Exploiting AI for Attacks: On the Interplay between Adversarial AI and Offensive AI
by: Schröer, Saskia Laura, et al.
Published: (2025)