Understanding the Effectiveness of Large Language Models in Detecting Security Vulnerabilities
Fuente:
arXiv
Saved in:
| Main Authors: | Khare, Avishree, Dutta, Saikat, Li, Ziyang, Solko-Breslin, Alaia, Alur, Rajeev, Naik, Mayur |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IRIS: LLM-Assisted Static Analysis for Detecting Security Vulnerabilities
by: Li, Ziyang, et al.
Published: (2024)
by: Li, Ziyang, et al.
Published: (2024)
QLCoder: A Query Synthesizer For Static Analysis of Security Vulnerabilities
by: Wang, Claire, et al.
Published: (2025)
by: Wang, Claire, et al.
Published: (2025)
DeepCode AI Fix: Fixing Security Vulnerabilities with Large Language Models
by: Berabi, Berkay, et al.
Published: (2024)
by: Berabi, Berkay, et al.
Published: (2024)
Demystifying Invariant Effectiveness for Securing Smart Contracts
by: Chen, Zhiyang, et al.
Published: (2024)
by: Chen, Zhiyang, et al.
Published: (2024)
Symbolic Execution Meets Multi-LLM Orchestration: Detecting Memory Vulnerabilities in Incomplete Rust CVE Snippets
by: Abdelrazek, Zeyad, et al.
Published: (2026)
by: Abdelrazek, Zeyad, et al.
Published: (2026)
The Secrets Must Not Flow: Scaling Security Verification to Large Codebases (extended version)
by: Arquint, Linard, et al.
Published: (2025)
by: Arquint, Linard, et al.
Published: (2025)
From Solitary Directives to Interactive Encouragement! LLM Secure Code Generation by Natural Language Prompting
by: Liu, Shigang, et al.
Published: (2024)
by: Liu, Shigang, et al.
Published: (2024)
Large Language Models for Code: Security Hardening and Adversarial Testing
by: He, Jingxuan, et al.
Published: (2023)
by: He, Jingxuan, et al.
Published: (2023)
Logicbreaks: A Framework for Understanding Subversion of Rule-based Inference
by: Xue, Anton, et al.
Published: (2024)
by: Xue, Anton, et al.
Published: (2024)
Yaksha-Prashna: Understanding eBPF Bytecode Network Function Behavior
by: Singh, Animesh, et al.
Published: (2026)
by: Singh, Animesh, et al.
Published: (2026)
Static Deadlock Detection for Rust Programs
by: Zhang, Yu, et al.
Published: (2024)
by: Zhang, Yu, et al.
Published: (2024)
Large Language Models Cannot Reliably Detect Vulnerabilities in JavaScript: The First Systematic Benchmark and Evaluation
by: Fei, Qingyuan, et al.
Published: (2025)
by: Fei, Qingyuan, et al.
Published: (2025)
Belobog: Move Language Fuzzing Framework For Real-World Smart Contracts
by: Kong, Ziqiao, et al.
Published: (2025)
by: Kong, Ziqiao, et al.
Published: (2025)
YASA: Scalable Multi-Language Taint Analysis on the Unified AST at Ant Group
by: Wang, Yayi, et al.
Published: (2026)
by: Wang, Yayi, et al.
Published: (2026)
Artemis: Toward Accurate Detection of Server-Side Request Forgeries through LLM-Assisted Inter-Procedural Path-Sensitive Taint Analysis
by: Ji, Yuchen, et al.
Published: (2025)
by: Ji, Yuchen, et al.
Published: (2025)
Argus: Reorchestrating Static Analysis via a Multi-Agent Ensemble for Full-Chain Security Vulnerability Detection
by: Liang, Zi, et al.
Published: (2026)
by: Liang, Zi, et al.
Published: (2026)
Fuzzing: Randomness? Reasoning! Efficient Directed Fuzzing via Large Language Models
by: Feng, Xiaotao, et al.
Published: (2025)
by: Feng, Xiaotao, et al.
Published: (2025)
Vulnerability Detection with Interprocedural Context in Multiple Languages: Assessing Effectiveness and Cost of Modern LLMs
by: Lira, Kevin, et al.
Published: (2026)
by: Lira, Kevin, et al.
Published: (2026)
HardTaint: Production-Run Dynamic Taint Analysis via Selective Hardware Tracing
by: Zhang, Yiyu, et al.
Published: (2024)
by: Zhang, Yiyu, et al.
Published: (2024)
Arguzz: Testing zkVMs for Soundness and Completeness Bugs
by: Hochrainer, Christoph, et al.
Published: (2025)
by: Hochrainer, Christoph, et al.
Published: (2025)
SmartInv: Multimodal Learning for Smart Contract Invariant Inference
by: Wang, Sally Junsong, et al.
Published: (2024)
by: Wang, Sally Junsong, et al.
Published: (2024)
Dynamic Taint Tracking using Partial Instrumentation for Java Applications
by: Thakur, Manoj RameshChandra
Published: (2024)
by: Thakur, Manoj RameshChandra
Published: (2024)
Fuzzing Processing Pipelines for Zero-Knowledge Circuits
by: Hochrainer, Christoph, et al.
Published: (2024)
by: Hochrainer, Christoph, et al.
Published: (2024)
Extracting Protocol Format as State Machine via Controlled Static Loop Analysis
by: Shi, Qingkai, et al.
Published: (2023)
by: Shi, Qingkai, et al.
Published: (2023)
Beyond Imprecise Distance Metrics: Trace-Guided Directed Greybox Fuzzing via LLM-Predicted Call Stacks
by: Zhang, Yifan, et al.
Published: (2025)
by: Zhang, Yifan, et al.
Published: (2025)
A Broad Comparative Evaluation of Software Debloating Tools
by: Brown, Michael D., et al.
Published: (2023)
by: Brown, Michael D., et al.
Published: (2023)
iResolveX: Multi-Layered Indirect Call Resolution via Static Reasoning and Learning-Augmented Refinement
by: Santra, Monika, et al.
Published: (2026)
by: Santra, Monika, et al.
Published: (2026)
Hornet Node and the Hornet DSL: A Minimal, Executable Specification for Bitcoin Consensus
by: Sharp, Toby
Published: (2025)
by: Sharp, Toby
Published: (2025)
Accurate and Extensible Symbolic Execution of Binary Code based on Formal ISA Semantics
by: Tempel, Sören, et al.
Published: (2024)
by: Tempel, Sören, et al.
Published: (2024)
Translating C To Rust: Lessons from a User Study
by: Li, Ruishi, et al.
Published: (2024)
by: Li, Ruishi, et al.
Published: (2024)
An Empirical Study of Bitwise Operators Intuitiveness through Performance Metrics
by: Joshi, Shubham
Published: (2025)
by: Joshi, Shubham
Published: (2025)
UBfuzz: Finding Bugs in Sanitizer Implementations
by: Li, Shaohua, et al.
Published: (2024)
by: Li, Shaohua, et al.
Published: (2024)
OpenTracer: A Dynamic Transaction Trace Analyzer for Smart Contract Invariant Generation and Beyond
by: Chen, Zhiyang, et al.
Published: (2024)
by: Chen, Zhiyang, et al.
Published: (2024)
HALURust: Exploiting Hallucinations of Large Language Models to Detect Vulnerabilities in Rust
by: Luo, Yu, et al.
Published: (2025)
by: Luo, Yu, et al.
Published: (2025)
SAGA: Detecting Security Vulnerabilities Using Static Aspect Analysis
by: Marquer, Yoann, et al.
Published: (2026)
by: Marquer, Yoann, et al.
Published: (2026)
An Empirical Study on the Effectiveness of Large Language Models for Binary Code Understanding
by: Shang, Xiuwei, et al.
Published: (2025)
by: Shang, Xiuwei, et al.
Published: (2025)
Improving Smart Contract Security with Contrastive Learning-based Vulnerability Detection
by: Chen, Yizhou, et al.
Published: (2024)
by: Chen, Yizhou, et al.
Published: (2024)
Vulnerability Detection in Popular Programming Languages with Language Models
by: Atiiq, Syafiq Al, et al.
Published: (2024)
by: Atiiq, Syafiq Al, et al.
Published: (2024)
Security Vulnerability Detection with Multitask Self-Instructed Fine-Tuning of Large Language Models
by: Yang, Aidan Z. H., et al.
Published: (2024)
by: Yang, Aidan Z. H., et al.
Published: (2024)
On the Effectiveness of Function-Level Vulnerability Detectors for Inter-Procedural Vulnerabilities
by: Li, Zhen, et al.
Published: (2024)
by: Li, Zhen, et al.
Published: (2024)
Similar Items
-
IRIS: LLM-Assisted Static Analysis for Detecting Security Vulnerabilities
by: Li, Ziyang, et al.
Published: (2024) -
QLCoder: A Query Synthesizer For Static Analysis of Security Vulnerabilities
by: Wang, Claire, et al.
Published: (2025) -
DeepCode AI Fix: Fixing Security Vulnerabilities with Large Language Models
by: Berabi, Berkay, et al.
Published: (2024) -
Demystifying Invariant Effectiveness for Securing Smart Contracts
by: Chen, Zhiyang, et al.
Published: (2024) -
Symbolic Execution Meets Multi-LLM Orchestration: Detecting Memory Vulnerabilities in Incomplete Rust CVE Snippets
by: Abdelrazek, Zeyad, et al.
Published: (2026)