From CVE Entries to Verifiable Exploits: An Automated Multi-Agent Framework for Reproducing CVEs
Fuente:
arXiv
Salvato in:
| Autori principali: | Ullah, Saad, Balasubramanian, Praneeth, Guo, Wenbo, Burnett, Amanda, Pearce, Hammond, Kruegel, Christopher, Vigna, Giovanni, Stringhini, Gianluca |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLMs Cannot Reliably Identify and Reason About Security Vulnerabilities (Yet?): A Comprehensive Evaluation, Framework, and Benchmarks
di: Ullah, Saad, et al.
Pubblicazione: (2023)
di: Ullah, Saad, et al.
Pubblicazione: (2023)
Multi-Agent Taint Specification Extraction for Vulnerability Detection
di: Ghebremichael, Jonah, et al.
Pubblicazione: (2026)
di: Ghebremichael, Jonah, et al.
Pubblicazione: (2026)
MADCAT: Combating Malware Detection Under Concept Drift with Test-Time Adaptation
di: Roh, Eunjin, et al.
Pubblicazione: (2025)
di: Roh, Eunjin, et al.
Pubblicazione: (2025)
LOKI: Proactively Discovering Online Scam Websites by Mining Toxic Search Queries
di: Paudel, Pujan, et al.
Pubblicazione: (2025)
di: Paudel, Pujan, et al.
Pubblicazione: (2025)
MalwarePT: A Binary-Level Foundation Model for Malware Analysis
di: Vasan, Saastha, et al.
Pubblicazione: (2026)
di: Vasan, Saastha, et al.
Pubblicazione: (2026)
When AI Meets the Web: Prompt Injection Risks in Third-Party AI Chatbot Plugins
di: Kaya, Yigitcan, et al.
Pubblicazione: (2025)
di: Kaya, Yigitcan, et al.
Pubblicazione: (2025)
Remote Keylogging Attacks in Multi-user VR Applications
di: Su, Zihao, et al.
Pubblicazione: (2024)
di: Su, Zihao, et al.
Pubblicazione: (2024)
The Secret Life of CVEs
di: Przymus, Piotr, et al.
Pubblicazione: (2025)
di: Przymus, Piotr, et al.
Pubblicazione: (2025)
TitanCA: Lessons from Orchestrating LLM Agents to Discover 100+ CVEs
di: Zhang, Ting, et al.
Pubblicazione: (2026)
di: Zhang, Ting, et al.
Pubblicazione: (2026)
Turning CVEs into Educational Labs:Insights and Challenges
di: Tafese, Trueye
Pubblicazione: (2025)
di: Tafese, Trueye
Pubblicazione: (2025)
Efficacy of EPSS in High Severity CVEs found in KEV
di: Parla, Rianna
Pubblicazione: (2024)
di: Parla, Rianna
Pubblicazione: (2024)
CVEs With a CVSS Score Greater Than or Equal to 9
di: Sinterhauf, Lena, et al.
Pubblicazione: (2026)
di: Sinterhauf, Lena, et al.
Pubblicazione: (2026)
AutoPatch: Multi-Agent Framework for Patching Real-World CVE Vulnerabilities
di: Seo, Minjae, et al.
Pubblicazione: (2025)
di: Seo, Minjae, et al.
Pubblicazione: (2025)
Can LLMs Classify CVEs? Investigating LLMs Capabilities in Computing CVSS Vectors
di: Marchiori, Francesco, et al.
Pubblicazione: (2025)
di: Marchiori, Francesco, et al.
Pubblicazione: (2025)
TrojanPuzzle: Covertly Poisoning Code-Suggestion Models
di: Aghakhani, Hojjat, et al.
Pubblicazione: (2023)
di: Aghakhani, Hojjat, et al.
Pubblicazione: (2023)
MalCVE: Malware Detection and CVE Association Using Large Language Models
di: Cristea, Eduard Andrei, et al.
Pubblicazione: (2025)
di: Cristea, Eduard Andrei, et al.
Pubblicazione: (2025)
Uncovering CWE-CVE-CPE Relations with Threat Knowledge Graphs
di: Shi, Zhenpeng, et al.
Pubblicazione: (2023)
di: Shi, Zhenpeng, et al.
Pubblicazione: (2023)
LLM4CVE: Enabling Iterative Automated Vulnerability Repair with Large Language Models
di: Fakih, Mohamad, et al.
Pubblicazione: (2025)
di: Fakih, Mohamad, et al.
Pubblicazione: (2025)
Reproducible Builds and Insights from an Independent Verifier for Arch Linux
di: Drexel, Joshua, et al.
Pubblicazione: (2025)
di: Drexel, Joshua, et al.
Pubblicazione: (2025)
Autonomous LLM Agent Worms: Cross-Platform Propagation, Automated Discovery and Temporal Re-Entry Defense
di: Zha, Mingming, et al.
Pubblicazione: (2026)
di: Zha, Mingming, et al.
Pubblicazione: (2026)
Enabling Contextual Soft Moderation on Social Media through Contrastive Textual Deviation
di: Paudel, Pujan, et al.
Pubblicazione: (2024)
di: Paudel, Pujan, et al.
Pubblicazione: (2024)
Silent Consent, Persistent Risk: Android Permission Groups and Custom Permissions
di: Akanji, Olawale Amos, et al.
Pubblicazione: (2026)
di: Akanji, Olawale Amos, et al.
Pubblicazione: (2026)
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
di: Tang, Yuheng, et al.
Pubblicazione: (2026)
di: Tang, Yuheng, et al.
Pubblicazione: (2026)
LASHED: LLMs And Static Hardware Analysis for Early Detection of RTL Bugs
di: Ahmad, Baleegh, et al.
Pubblicazione: (2025)
di: Ahmad, Baleegh, et al.
Pubblicazione: (2025)
Reproducibility in Event-Log Research: A Parametrised Generator and Benchmark for Event-based Signatures
di: Khan, Saad, et al.
Pubblicazione: (2026)
di: Khan, Saad, et al.
Pubblicazione: (2026)
CVE-Bench: A Benchmark for AI Agents' Ability to Exploit Real-World Web Application Vulnerabilities
di: Zhu, Yuxuan, et al.
Pubblicazione: (2025)
di: Zhu, Yuxuan, et al.
Pubblicazione: (2025)
Invisible Image Watermarks Are Provably Removable Using Generative AI
di: Zhao, Xuandong, et al.
Pubblicazione: (2023)
di: Zhao, Xuandong, et al.
Pubblicazione: (2023)
ThreatLinker: An NLP-based Methodology to Automatically Estimate CVE Relevance for CAPEC Attack Patterns
di: Ciavotta, Andrea, et al.
Pubblicazione: (2025)
di: Ciavotta, Andrea, et al.
Pubblicazione: (2025)
CVE Breadcrumbs: Tracking Vulnerabilities Through Versioned Apache Libraries
di: Garcia, Derek, et al.
Pubblicazione: (2025)
di: Garcia, Derek, et al.
Pubblicazione: (2025)
Cybersecurity Defenses: Exploration of CVE Types through Attack Descriptions
di: Othman, Refat, et al.
Pubblicazione: (2024)
di: Othman, Refat, et al.
Pubblicazione: (2024)
Decentralized Vulnerability Disclosure via Permissioned Blockchain: A Secure, Transparent Alternative to Centralized CVE Management
di: Amirov, Novruz, et al.
Pubblicazione: (2025)
di: Amirov, Novruz, et al.
Pubblicazione: (2025)
Fixing Hardware Security Bugs with Large Language Models
di: Ahmad, Baleegh, et al.
Pubblicazione: (2023)
di: Ahmad, Baleegh, et al.
Pubblicazione: (2023)
TUBERAIDER: Attributing Coordinated Hate Attacks on YouTube Videos to their Source Communities
di: Saeed, Mohammad Hammas, et al.
Pubblicazione: (2023)
di: Saeed, Mohammad Hammas, et al.
Pubblicazione: (2023)
The Cost of Convenience: Identifying, Analyzing, and Mitigating Predatory Loan Applications on Android
di: Akanji, Olawale Amos, et al.
Pubblicazione: (2026)
di: Akanji, Olawale Amos, et al.
Pubblicazione: (2026)
Linux Kernel Recency Matters, CVE Severity Doesn't, and History Fades
di: Przymus, Piotr, et al.
Pubblicazione: (2026)
di: Przymus, Piotr, et al.
Pubblicazione: (2026)
HackerSignal: A Large-Scale Multi-Source Dataset Linking Hacker Community Discourse to the CVE Vulnerability Lifecycle
di: Ampel, Benjamin M., et al.
Pubblicazione: (2026)
di: Ampel, Benjamin M., et al.
Pubblicazione: (2026)
Automated CVE Analysis: Harnessing Machine Learning In Designing Question-Answering Models For Cybersecurity Information Extraction
di: Faruk, Tanjim Bin
Pubblicazione: (2024)
di: Faruk, Tanjim Bin
Pubblicazione: (2024)
Automation-Exploit: A Multi-Agent LLM Framework for Adaptive Offensive Security with Digital Twin-Based Risk-Mitigated Exploitation
di: Andreucci, Biagio, et al.
Pubblicazione: (2026)
di: Andreucci, Biagio, et al.
Pubblicazione: (2026)
Combinatorial Privacy: Private Multi-Party Bitstream Grand Sum by Hiding in Birkhoff Polytopes
di: Vepakomma, Praneeth
Pubblicazione: (2026)
di: Vepakomma, Praneeth
Pubblicazione: (2026)
Takedown: How It's Done in Modern Coding Agent Exploits
di: Lee, Eunkyu, et al.
Pubblicazione: (2025)
di: Lee, Eunkyu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LLMs Cannot Reliably Identify and Reason About Security Vulnerabilities (Yet?): A Comprehensive Evaluation, Framework, and Benchmarks
di: Ullah, Saad, et al.
Pubblicazione: (2023) -
Multi-Agent Taint Specification Extraction for Vulnerability Detection
di: Ghebremichael, Jonah, et al.
Pubblicazione: (2026) -
MADCAT: Combating Malware Detection Under Concept Drift with Test-Time Adaptation
di: Roh, Eunjin, et al.
Pubblicazione: (2025) -
LOKI: Proactively Discovering Online Scam Websites by Mining Toxic Search Queries
di: Paudel, Pujan, et al.
Pubblicazione: (2025) -
MalwarePT: A Binary-Level Foundation Model for Malware Analysis
di: Vasan, Saastha, et al.
Pubblicazione: (2026)