HackSynth: LLM Agent and Evaluation Framework for Autonomous Penetration Testing
Fuente:
arXiv
Salvato in:
| Autori principali: | Muzsai, Lajos, Imolai, David, Lukács, András |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving LLM Agents with Reinforcement Learning on Cryptographic CTF Challenges
di: Muzsai, Lajos, et al.
Pubblicazione: (2025)
di: Muzsai, Lajos, et al.
Pubblicazione: (2025)
PACT: Reducing Alert Fatigue in Low-Prevalence SOC Streams with Triggered Active Learning
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)
Explainable Attention-Based LSTM Framework for Early Detection of AI-Assisted Ransomware via File System Behavioral Analysis
di: Nayak, Prabhudarshi, et al.
Pubblicazione: (2026)
di: Nayak, Prabhudarshi, et al.
Pubblicazione: (2026)
Expanding the Attack Scenarios of SAE J1939: A Comprehensive Analysis of Established and Novel Vulnerabilities in Transport Protocol
di: Lee, Hwejae, et al.
Pubblicazione: (2024)
di: Lee, Hwejae, et al.
Pubblicazione: (2024)
"Abuse Risks are Often Inherent to Product Features": Exploring AI Vendors' Bug Bounty and Responsible Disclosure Policies
di: Piao, Yangheran, et al.
Pubblicazione: (2025)
di: Piao, Yangheran, et al.
Pubblicazione: (2025)
Operationalizing Cybersecurity Governance for Mitigation Planning with Attack-Path Modeling and Reinforcement Learning
di: Huff, Philip, et al.
Pubblicazione: (2026)
di: Huff, Philip, et al.
Pubblicazione: (2026)
Quantum Machine Learning for Cyber-Physical Anomaly Detection in Unmanned Aerial Vehicles: A Leakage-Free Evaluation with Proxy-Audited Feature Sets
di: Paredes, Carlos A. Durán, et al.
Pubblicazione: (2026)
di: Paredes, Carlos A. Durán, et al.
Pubblicazione: (2026)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
di: Dang, Kieu, et al.
Pubblicazione: (2025)
di: Dang, Kieu, et al.
Pubblicazione: (2025)
Control Physiology: An Agent-Based Model of FAIR-CAM Dynamics
di: Jones, Jack, et al.
Pubblicazione: (2026)
di: Jones, Jack, et al.
Pubblicazione: (2026)
Tool Receipts, Not Zero-Knowledge Proofs: Practical Hallucination Detection for AI Agents
di: Basu, Abhinaba
Pubblicazione: (2026)
di: Basu, Abhinaba
Pubblicazione: (2026)
Cross-Domain Malware Detection via Probability-Level Fusion of Lightweight Gradient Boosting Models
di: Mohamed, Omar Khalid Ali
Pubblicazione: (2025)
di: Mohamed, Omar Khalid Ali
Pubblicazione: (2025)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
di: Mitchell, Richard Joseph
Pubblicazione: (2026)
di: Mitchell, Richard Joseph
Pubblicazione: (2026)
Developing a Strong CPS Defender: An Evolutionary Approach
di: Hu, Qingyuan, et al.
Pubblicazione: (2025)
di: Hu, Qingyuan, et al.
Pubblicazione: (2025)
Exploratory Analysis of Cyberattack Patterns on E-Commerce Platforms Using Statistical Methods
di: Adeniya, Fatimo Adenike
Pubblicazione: (2025)
di: Adeniya, Fatimo Adenike
Pubblicazione: (2025)
False Security Confidence in Benign LLM Code Generation
di: Ren, Xiaolei
Pubblicazione: (2026)
di: Ren, Xiaolei
Pubblicazione: (2026)
Identity Deepfake Threats to Biometric Authentication Systems: Public and Expert Perspectives
di: He, Shijing, et al.
Pubblicazione: (2025)
di: He, Shijing, et al.
Pubblicazione: (2025)
RAR: Setting Knowledge Tripwires for Retrieval Augmented Rejection
di: Buonocore, Tommaso Mario, et al.
Pubblicazione: (2025)
di: Buonocore, Tommaso Mario, et al.
Pubblicazione: (2025)
Multi-Agent Honeypot-Based Request-Response Context Dataset for Improved SQL Injection Detection Performance
di: Yu, Hao, et al.
Pubblicazione: (2026)
di: Yu, Hao, et al.
Pubblicazione: (2026)
Revisiting Third-Party Library Detection: A Ground Truth Dataset and Its Implications Across Security Tasks
di: Gu, Jintao, et al.
Pubblicazione: (2025)
di: Gu, Jintao, et al.
Pubblicazione: (2025)
RealVuln: Benchmarking Rule-Based, General-Purpose LLM, and Security-Specialized Scanners on Real-World Code
di: Pellew, John, et al.
Pubblicazione: (2026)
di: Pellew, John, et al.
Pubblicazione: (2026)
A Method for Quantifying Human Risk and a Blueprint for LLM Integration
di: Canale, Giuseppe
Pubblicazione: (2025)
di: Canale, Giuseppe
Pubblicazione: (2025)
VOLTRON: Detecting Unknown Malware Using Graph-Based Zero-Shot Learning
di: Akdeniz, M. Tahir, et al.
Pubblicazione: (2025)
di: Akdeniz, M. Tahir, et al.
Pubblicazione: (2025)
Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks
di: Merves, Tyler H., et al.
Pubblicazione: (2026)
di: Merves, Tyler H., et al.
Pubblicazione: (2026)
A Dual-Path Generative Framework for Zero-Day Fraud Detection in Banking Systems
di: Ismail, Nasim Abdirahman, et al.
Pubblicazione: (2026)
di: Ismail, Nasim Abdirahman, et al.
Pubblicazione: (2026)
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
di: Ahi, Kiarash, et al.
Pubblicazione: (2026)
di: Ahi, Kiarash, et al.
Pubblicazione: (2026)
Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study
di: Xu, Luyao, et al.
Pubblicazione: (2026)
di: Xu, Luyao, et al.
Pubblicazione: (2026)
Refute-or-Promote: An Adversarial Stage-Gated Multi-Agent Review Methodology for High-Precision LLM-Assisted Defect Discovery
di: Agarwal, Abhinav
Pubblicazione: (2026)
di: Agarwal, Abhinav
Pubblicazione: (2026)
Illuminating the Black Box: Real-Time Monitoring of Backdoor Unlearning in CNNs via Explainable AI
di: Hoang, Tien Dat
Pubblicazione: (2025)
di: Hoang, Tien Dat
Pubblicazione: (2025)
An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations
di: Fatouros, George, et al.
Pubblicazione: (2026)
di: Fatouros, George, et al.
Pubblicazione: (2026)
SCAFDS: Edge-Feature Graph Attention for Interbank Fraud Detection with Attribution-Grounded SAR Generation
di: Uddin, Mohammad Nasir
Pubblicazione: (2026)
di: Uddin, Mohammad Nasir
Pubblicazione: (2026)
Safety, Security, and Cognitive Risks in World Models
di: Parmar, Manoj
Pubblicazione: (2026)
di: Parmar, Manoj
Pubblicazione: (2026)
h4rm3l: A language for Composable Jailbreak Attack Synthesis
di: Doumbouya, Moussa Koulako Bala, et al.
Pubblicazione: (2024)
di: Doumbouya, Moussa Koulako Bala, et al.
Pubblicazione: (2024)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
DRIFT: Drift-Resilient Invariant-Feature Transformer for DGA Detection
di: Lee, Chaeyoung, et al.
Pubblicazione: (2026)
di: Lee, Chaeyoung, et al.
Pubblicazione: (2026)
Benchmarking Autonomous Agents against Temporal, Spatial, and Semantic Evasions
di: Ma, Jianan, et al.
Pubblicazione: (2026)
di: Ma, Jianan, et al.
Pubblicazione: (2026)
Send to which account? Evaluation of an LLM-based Scambaiting System
di: Siadati, Hossein, et al.
Pubblicazione: (2025)
di: Siadati, Hossein, et al.
Pubblicazione: (2025)
From nuclear safety to LLM security: Applying non-probabilistic risk management strategies to build safe and secure LLM-powered systems
di: Gutfraind, Alexander, et al.
Pubblicazione: (2025)
di: Gutfraind, Alexander, et al.
Pubblicazione: (2025)
Constant-Size Cryptographic Evidence Structures for Regulated AI Workflows
di: Kao, Leo
Pubblicazione: (2025)
di: Kao, Leo
Pubblicazione: (2025)
Sola-Visibility-ISPM: Benchmarking Agentic AI for Identity Security Posture Management Visibility
di: Engelberg, Gal, et al.
Pubblicazione: (2026)
di: Engelberg, Gal, et al.
Pubblicazione: (2026)
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
di: Gu, Yongtong, et al.
Pubblicazione: (2026)
di: Gu, Yongtong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Improving LLM Agents with Reinforcement Learning on Cryptographic CTF Challenges
di: Muzsai, Lajos, et al.
Pubblicazione: (2025) -
PACT: Reducing Alert Fatigue in Low-Prevalence SOC Streams with Triggered Active Learning
di: Ndichu, Samuel, et al.
Pubblicazione: (2026) -
Explainable Attention-Based LSTM Framework for Early Detection of AI-Assisted Ransomware via File System Behavioral Analysis
di: Nayak, Prabhudarshi, et al.
Pubblicazione: (2026) -
Expanding the Attack Scenarios of SAE J1939: A Comprehensive Analysis of Established and Novel Vulnerabilities in Transport Protocol
di: Lee, Hwejae, et al.
Pubblicazione: (2024) -
"Abuse Risks are Often Inherent to Product Features": Exploring AI Vendors' Bug Bounty and Responsible Disclosure Policies
di: Piao, Yangheran, et al.
Pubblicazione: (2025)