Construction and Evaluation of LLM-based agents for Semi-Autonomous penetration testing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kobayashi, Masaya, Fuchi, Masane, Zanashir, Amar, Yoneda, Tomonori, Takagi, Tomohiro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ResBit: Residual Bit Vector for Categorical Values
von: Fuchi, Masane, et al.
Veröffentlicht: (2023)
von: Fuchi, Masane, et al.
Veröffentlicht: (2023)
Erasing with Precision: Evaluating Specific Concept Erasure from Text-to-Image Generative Models
von: Fuchi, Masane, et al.
Veröffentlicht: (2025)
von: Fuchi, Masane, et al.
Veröffentlicht: (2025)
RecTable: Fast Modeling Tabular Data with Rectified Flow
von: Fuchi, Masane, et al.
Veröffentlicht: (2025)
von: Fuchi, Masane, et al.
Veröffentlicht: (2025)
Erasing Concepts from Text-to-Image Diffusion Models with Few-shot Unlearning
von: Fuchi, Masane, et al.
Veröffentlicht: (2024)
von: Fuchi, Masane, et al.
Veröffentlicht: (2024)
Vulnerability Mitigation System (VMS): LLM Agent and Evaluation Framework for Autonomous Penetration Testing
von: Abdulzada, Farzana
Veröffentlicht: (2025)
von: Abdulzada, Farzana
Veröffentlicht: (2025)
Autonomous Adversary: Red-Teaming in the age of LLM
von: Mamun, Mohammad, et al.
Veröffentlicht: (2026)
von: Mamun, Mohammad, et al.
Veröffentlicht: (2026)
Evaluating LLM-based Personal Information Extraction and Countermeasures
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
Asymmetry Vulnerability and Physical Attacks on Online Map Construction for Autonomous Driving
von: Lou, Yang, et al.
Veröffentlicht: (2025)
von: Lou, Yang, et al.
Veröffentlicht: (2025)
Red-MIRROR: Agentic LLM-based Autonomous Penetration Testing with Reflective Verification and Knowledge-augmented Interaction
von: Khang, Tran Vy, et al.
Veröffentlicht: (2026)
von: Khang, Tran Vy, et al.
Veröffentlicht: (2026)
AI Kill Switch for malicious web-based LLM agent
von: Lee, Sechan, et al.
Veröffentlicht: (2025)
von: Lee, Sechan, et al.
Veröffentlicht: (2025)
Hypervisor-based Double Extortion Ransomware Detection Method Using Kitsune Network Features
von: Hirano, Manabu, et al.
Veröffentlicht: (2025)
von: Hirano, Manabu, et al.
Veröffentlicht: (2025)
TFHE-Coder: Evaluating LLM-agentic Fully Homomorphic Encryption Code Generation
von: Kumar, Mayank, et al.
Veröffentlicht: (2025)
von: Kumar, Mayank, et al.
Veröffentlicht: (2025)
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks
von: Song, Ruoyu, et al.
Veröffentlicht: (2024)
von: Song, Ruoyu, et al.
Veröffentlicht: (2024)
CyberSleuth: Autonomous Blue-Team LLM Agent for Web Attack Forensics
von: Fumero, Stefano, et al.
Veröffentlicht: (2025)
von: Fumero, Stefano, et al.
Veröffentlicht: (2025)
Semi-Supervised Learning for Anomaly Detection in Blockchain-based Supply Chains
von: Son, Do Hai, et al.
Veröffentlicht: (2024)
von: Son, Do Hai, et al.
Veröffentlicht: (2024)
Securing Private Federated Learning in a Malicious Setting: A Scalable TEE-Based Approach with Client Auditing
von: Takagi, Shun, et al.
Veröffentlicht: (2025)
von: Takagi, Shun, et al.
Veröffentlicht: (2025)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
von: Wang, Yihan, et al.
Veröffentlicht: (2025)
von: Wang, Yihan, et al.
Veröffentlicht: (2025)
Learning with Errors over Group Rings Constructed by Semi-direct Product
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023)
A Study of Semi-Fungible Token based Wi-Fi Access Control
von: Ye, Litao, et al.
Veröffentlicht: (2025)
von: Ye, Litao, et al.
Veröffentlicht: (2025)
Evaluating LLM Generated Detection Rules in Cybersecurity
von: Bertiger, Anna, et al.
Veröffentlicht: (2025)
von: Bertiger, Anna, et al.
Veröffentlicht: (2025)
The Evolution of Agentic AI in Cybersecurity: From Single LLM Reasoners to Multi-Agent Systems and Autonomous Pipelines
von: Vinay, Vaishali
Veröffentlicht: (2025)
von: Vinay, Vaishali
Veröffentlicht: (2025)
Autonomous LLM Agent Worms: Cross-Platform Propagation, Automated Discovery and Temporal Re-Entry Defense
von: Zha, Mingming, et al.
Veröffentlicht: (2026)
von: Zha, Mingming, et al.
Veröffentlicht: (2026)
Counterfactual Evaluation for Blind Attack Detection in LLM-based Evaluation Systems
von: Liu, Lijia, et al.
Veröffentlicht: (2025)
von: Liu, Lijia, et al.
Veröffentlicht: (2025)
Input-Output History Feedback Controller for Encrypted Control with Leveled Fully Homomorphic Encryption
von: Teranishi, Kaoru, et al.
Veröffentlicht: (2021)
von: Teranishi, Kaoru, et al.
Veröffentlicht: (2021)
LLM-Assisted AHP for Explainable Cyber Range Evaluation
von: Kampourakis, Vyron, et al.
Veröffentlicht: (2025)
von: Kampourakis, Vyron, et al.
Veröffentlicht: (2025)
Evaluating the Vulnerability Landscape of LLM-Generated Smart Contracts
von: Do, Hoang Long, et al.
Veröffentlicht: (2026)
von: Do, Hoang Long, et al.
Veröffentlicht: (2026)
Security awareness in LLM agents: the NDAI zone case
von: Bottazzi, Enrico, et al.
Veröffentlicht: (2026)
von: Bottazzi, Enrico, et al.
Veröffentlicht: (2026)
Semi-Compressed CRYSTALS-Kyber
von: Liu, Shuiyin, et al.
Veröffentlicht: (2024)
von: Liu, Shuiyin, et al.
Veröffentlicht: (2024)
LLM Agents can Autonomously Hack Websites
von: Fang, Richard, et al.
Veröffentlicht: (2024)
von: Fang, Richard, et al.
Veröffentlicht: (2024)
Evasive Ransomware Attacks Using Low-level Behavioral Adversarial Examples
von: Hirano, Manabu, et al.
Veröffentlicht: (2025)
von: Hirano, Manabu, et al.
Veröffentlicht: (2025)
ROFBS$α$: Real Time Backup System Decoupled from ML Based Ransomware Detection
von: Higuchi, Kosuke, et al.
Veröffentlicht: (2025)
von: Higuchi, Kosuke, et al.
Veröffentlicht: (2025)
Impact of File-Open Hook Points on Backup Ratio in ROFBS on XFS
von: Higuchi, Kosuke, et al.
Veröffentlicht: (2026)
von: Higuchi, Kosuke, et al.
Veröffentlicht: (2026)
FuncPoison: Poisoning Function Library to Hijack Multi-agent Autonomous Driving Systems
von: Long, Yuzhen, et al.
Veröffentlicht: (2025)
von: Long, Yuzhen, et al.
Veröffentlicht: (2025)
RLSpoofer: A Lightweight Evaluator for LLM Watermark Spoofing Resilience
von: Huang, Hanbo, et al.
Veröffentlicht: (2026)
von: Huang, Hanbo, et al.
Veröffentlicht: (2026)
Systematic Categorization, Construction and Evaluation of New Attacks against Multi-modal Mobile GUI Agents
von: Yang, Yulong, et al.
Veröffentlicht: (2024)
von: Yang, Yulong, et al.
Veröffentlicht: (2024)
An LLM Agent-based Framework for Whaling Countermeasures
von: Miyamoto, Daisuke, et al.
Veröffentlicht: (2026)
von: Miyamoto, Daisuke, et al.
Veröffentlicht: (2026)
Odoo-based Subcontract Inter-site Access Control Mechanism for Construction Projects
von: Ho, Huy Hung, et al.
Veröffentlicht: (2025)
von: Ho, Huy Hung, et al.
Veröffentlicht: (2025)
SoK: Security of the Image Processing Pipeline for Camera-based Sensing in Autonomous Vehicles
von: Kühr, Michael, et al.
Veröffentlicht: (2024)
von: Kühr, Michael, et al.
Veröffentlicht: (2024)
CTFusion: A CTF-based Benchmark for LLM Agent Evaluation
von: Lee, Dongjun, et al.
Veröffentlicht: (2026)
von: Lee, Dongjun, et al.
Veröffentlicht: (2026)
Autonomous LLM Agents & CTFs: A Second Look
von: Bouchari, Youness, et al.
Veröffentlicht: (2026)
von: Bouchari, Youness, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ResBit: Residual Bit Vector for Categorical Values
von: Fuchi, Masane, et al.
Veröffentlicht: (2023) -
Erasing with Precision: Evaluating Specific Concept Erasure from Text-to-Image Generative Models
von: Fuchi, Masane, et al.
Veröffentlicht: (2025) -
RecTable: Fast Modeling Tabular Data with Rectified Flow
von: Fuchi, Masane, et al.
Veröffentlicht: (2025) -
Erasing Concepts from Text-to-Image Diffusion Models with Few-shot Unlearning
von: Fuchi, Masane, et al.
Veröffentlicht: (2024) -
Vulnerability Mitigation System (VMS): LLM Agent and Evaluation Framework for Autonomous Penetration Testing
von: Abdulzada, Farzana
Veröffentlicht: (2025)