PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Huihui, Shi, Jieke, Chen, Junkai, Zhang, Ting, Li, Yikun, Yang, Chengran, Ouh, Eng Lieh, Shar, Lwin Khin, Lo, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantics-Aligned, Curriculum-Driven, and Reasoning-Enhanced Vulnerability Repair Framework
von: Yang, Chengran, et al.
Veröffentlicht: (2025)
von: Yang, Chengran, et al.
Veröffentlicht: (2025)
Beyond Function-Level Analysis: Context-Aware Reasoning for Inter-Procedural Vulnerability Detection
von: Li, Yikun, et al.
Veröffentlicht: (2026)
von: Li, Yikun, et al.
Veröffentlicht: (2026)
An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning
von: Li, Yikun, et al.
Veröffentlicht: (2026)
von: Li, Yikun, et al.
Veröffentlicht: (2026)
Benchmarking Large Language Models for Multi-Language Software Vulnerability Detection
von: Zhang, Ting, et al.
Veröffentlicht: (2025)
von: Zhang, Ting, et al.
Veröffentlicht: (2025)
Back to the Basics: Rethinking Issue-Commit Linking with LLM-Assisted Retrieval
von: Huang, Huihui, et al.
Veröffentlicht: (2025)
von: Huang, Huihui, et al.
Veröffentlicht: (2025)
Revisiting Vulnerability Patch Identification on Data in the Wild
von: Irsan, Ivana Clairine, et al.
Veröffentlicht: (2026)
von: Irsan, Ivana Clairine, et al.
Veröffentlicht: (2026)
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation
von: Weyssow, Martin, et al.
Veröffentlicht: (2025)
von: Weyssow, Martin, et al.
Veröffentlicht: (2025)
Let the Trial Begin: A Mock-Court Approach to Vulnerability Detection using LLM-Based Agents
von: Widyasari, Ratnadira, et al.
Veröffentlicht: (2025)
von: Widyasari, Ratnadira, et al.
Veröffentlicht: (2025)
PatchSeeker: Mapping NVD Records to their Vulnerability-fixing Commits with LLM Generated Commits and Embeddings
von: Nguyen, Huu Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Hung, et al.
Veröffentlicht: (2025)
VulCoCo: A Simple Yet Effective Method for Detecting Vulnerable Code Clones
von: Bui, Tan, et al.
Veröffentlicht: (2025)
von: Bui, Tan, et al.
Veröffentlicht: (2025)
Mapping NVD Records to Their Vulnerability-fixing Commits: How Hard is It?
von: Nguyen, Huu Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Hung, et al.
Veröffentlicht: (2025)
Out of Distribution, Out of Luck: How Well Can LLMs Trained on Vulnerability Datasets Detect Top 25 CWE Weaknesses?
von: Li, Yikun, et al.
Veröffentlicht: (2025)
von: Li, Yikun, et al.
Veröffentlicht: (2025)
LIDL: LLM Integration Defect Localization via Knowledge Graph-Enhanced Multi-Agent Analysis
von: Tan, Gou, et al.
Veröffentlicht: (2026)
von: Tan, Gou, et al.
Veröffentlicht: (2026)
CleanVul: Automatic Function-Level Vulnerability Detection in Code Commits Using LLM Heuristics
von: Li, Yikun, et al.
Veröffentlicht: (2024)
von: Li, Yikun, et al.
Veröffentlicht: (2024)
VLM-Fuzz: Vision Language Model Assisted Recursive Depth-first Search Exploration for Effective UI Testing of Android Apps
von: Demissie, Biniam Fisseha, et al.
Veröffentlicht: (2025)
von: Demissie, Biniam Fisseha, et al.
Veröffentlicht: (2025)
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
von: Chen, Junkai, et al.
Veröffentlicht: (2025)
von: Chen, Junkai, et al.
Veröffentlicht: (2025)
Virtualization-based Penetration Testing Study for Detecting Accessibility Abuse Vulnerabilities in Banking Apps in East and Southeast Asia
von: Minn, Wei, et al.
Veröffentlicht: (2026)
von: Minn, Wei, et al.
Veröffentlicht: (2026)
ACECode: A Reinforcement Learning Framework for Aligning Code Efficiency and Correctness in Code Language Models
von: Yang, Chengran, et al.
Veröffentlicht: (2024)
von: Yang, Chengran, et al.
Veröffentlicht: (2024)
Fixseeker: An Empirical Driven Graph-based Approach for Detecting Silent Vulnerability Fixes in Open Source Software
von: Cheng, Yiran, et al.
Veröffentlicht: (2025)
von: Cheng, Yiran, et al.
Veröffentlicht: (2025)
CovAgent: Overcoming the 30% Curse of Mobile Application Coverage with Agentic AI and Dynamic Instrumentation
von: Minn, Wei, et al.
Veröffentlicht: (2026)
von: Minn, Wei, et al.
Veröffentlicht: (2026)
Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead
von: Cheng, Yiran, et al.
Veröffentlicht: (2025)
von: Cheng, Yiran, et al.
Veröffentlicht: (2025)
VERCATION: Precise Vulnerable Open-source Software Version Identification based on Static Analysis and LLM
von: Cheng, Yiran, et al.
Veröffentlicht: (2024)
von: Cheng, Yiran, et al.
Veröffentlicht: (2024)
Think Like Human Developers: Harnessing Community Knowledge for Structured Code Reasoning
von: Yang, Chengran, et al.
Veröffentlicht: (2025)
von: Yang, Chengran, et al.
Veröffentlicht: (2025)
AgentSZZ: Teaching the LLM Agent to Play Detective with Bug-Inducing Commits
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026)
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026)
Curiosity-Driven Testing for Sequential Decision-Making Process
von: He, Junda, et al.
Veröffentlicht: (2025)
von: He, Junda, et al.
Veröffentlicht: (2025)
Runtime Anomaly Detection for Drones: An Integrated Rule-Mining and Unsupervised-Learning Approach
von: Tan, Ivan, et al.
Veröffentlicht: (2025)
von: Tan, Ivan, et al.
Veröffentlicht: (2025)
Bamboo: LLM-Driven Discovery of API-Permission Mappings in the Android Framework
von: Hu, Han, et al.
Veröffentlicht: (2025)
von: Hu, Han, et al.
Veröffentlicht: (2025)
Finding Memory Leaks in C/C++ Programs via Neuro-Symbolic Augmented Static Analysis
von: Huang, Huihui, et al.
Veröffentlicht: (2026)
von: Huang, Huihui, et al.
Veröffentlicht: (2026)
Efficient and Green Large Language Models for Software Engineering: Literature Review, Vision, and the Road Ahead
von: Shi, Jieke, et al.
Veröffentlicht: (2024)
von: Shi, Jieke, et al.
Veröffentlicht: (2024)
MUCOCO: Automated Consistency Testing of Code LLMs
von: Chou, Chua Jin, et al.
Veröffentlicht: (2026)
von: Chou, Chua Jin, et al.
Veröffentlicht: (2026)
Compiling Code LLMs into Lightweight Executables
von: Shi, Jieke, et al.
Veröffentlicht: (2026)
von: Shi, Jieke, et al.
Veröffentlicht: (2026)
TitanCA: Lessons from Orchestrating LLM Agents to Discover 100+ CVEs
von: Zhang, Ting, et al.
Veröffentlicht: (2026)
von: Zhang, Ting, et al.
Veröffentlicht: (2026)
Ecosystem of Large Language Models for Code
von: Yang, Zhou, et al.
Veröffentlicht: (2024)
von: Yang, Zhou, et al.
Veröffentlicht: (2024)
BugsInPy: A Database of Existing Bugs in Python Programs to Enable Controlled Testing and Debugging Studies
von: Widyasari, Ratnadira, et al.
Veröffentlicht: (2024)
von: Widyasari, Ratnadira, et al.
Veröffentlicht: (2024)
Automated Penetration Testing: Formalization and Realization
von: Skandylas, Charilaos, et al.
Veröffentlicht: (2024)
von: Skandylas, Charilaos, et al.
Veröffentlicht: (2024)
Synthesizing Efficient and Permissive Programmatic Runtime Shields for Neural Policies
von: Shi, Jieke, et al.
Veröffentlicht: (2024)
von: Shi, Jieke, et al.
Veröffentlicht: (2024)
Automated Repair of TEE Partitioning Issues via DSL-Guided and LLM-Assisted Patching
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
von: Li, Dawei, et al.
Veröffentlicht: (2026)
von: Li, Dawei, et al.
Veröffentlicht: (2026)
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
Beyond the Tip of the Iceberg: Understanding SATD in Dockerfiles through the Lens of Co-evolution
von: Minn, Wei, et al.
Veröffentlicht: (2026)
von: Minn, Wei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Semantics-Aligned, Curriculum-Driven, and Reasoning-Enhanced Vulnerability Repair Framework
von: Yang, Chengran, et al.
Veröffentlicht: (2025) -
Beyond Function-Level Analysis: Context-Aware Reasoning for Inter-Procedural Vulnerability Detection
von: Li, Yikun, et al.
Veröffentlicht: (2026) -
An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning
von: Li, Yikun, et al.
Veröffentlicht: (2026) -
Benchmarking Large Language Models for Multi-Language Software Vulnerability Detection
von: Zhang, Ting, et al.
Veröffentlicht: (2025) -
Back to the Basics: Rethinking Issue-Commit Linking with LLM-Assisted Retrieval
von: Huang, Huihui, et al.
Veröffentlicht: (2025)