CleanVul: Automatic Function-Level Vulnerability Detection in Code Commits Using LLM Heuristics
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yikun, Zhang, Ting, Widyasari, Ratnadira, Tun, Yan Naing, Nguyen, Huu Hung, Bui, Tan, Irsan, Ivana Clairine, Cheng, Yiran, Lan, Xiang, Ang, Han Wei, Liauw, Frank, Weyssow, Martin, Kang, Hong Jin, Ouh, Eng Lieh, Shar, Lwin Khin, Lo, David |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Let the Trial Begin: A Mock-Court Approach to Vulnerability Detection using LLM-Based Agents
by: Widyasari, Ratnadira, et al.
Published: (2025)
by: Widyasari, Ratnadira, et al.
Published: (2025)
Back to the Basics: Rethinking Issue-Commit Linking with LLM-Assisted Retrieval
by: Huang, Huihui, et al.
Published: (2025)
by: Huang, Huihui, et al.
Published: (2025)
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation
by: Weyssow, Martin, et al.
Published: (2025)
by: Weyssow, Martin, et al.
Published: (2025)
Revisiting Vulnerability Patch Identification on Data in the Wild
by: Irsan, Ivana Clairine, et al.
Published: (2026)
by: Irsan, Ivana Clairine, et al.
Published: (2026)
VulCoCo: A Simple Yet Effective Method for Detecting Vulnerable Code Clones
by: Bui, Tan, et al.
Published: (2025)
by: Bui, Tan, et al.
Published: (2025)
Mapping NVD Records to Their Vulnerability-fixing Commits: How Hard is It?
by: Nguyen, Huu Hung, et al.
Published: (2025)
by: Nguyen, Huu Hung, et al.
Published: (2025)
PatchSeeker: Mapping NVD Records to their Vulnerability-fixing Commits with LLM Generated Commits and Embeddings
by: Nguyen, Huu Hung, et al.
Published: (2025)
by: Nguyen, Huu Hung, et al.
Published: (2025)
TitanCA: Lessons from Orchestrating LLM Agents to Discover 100+ CVEs
by: Zhang, Ting, et al.
Published: (2026)
by: Zhang, Ting, et al.
Published: (2026)
JavaVFC: Java Vulnerability Fixing Commits from Open-source Software
by: Bui, Tan, et al.
Published: (2024)
by: Bui, Tan, et al.
Published: (2024)
Benchmarking Large Language Models for Multi-Language Software Vulnerability Detection
by: Zhang, Ting, et al.
Published: (2025)
by: Zhang, Ting, et al.
Published: (2025)
Debt Behind the AI Boom: A Large-Scale Empirical Study of AI-Generated Code in the Wild
by: Liu, Yue, et al.
Published: (2026)
by: Liu, Yue, et al.
Published: (2026)
VLM-Fuzz: Vision Language Model Assisted Recursive Depth-first Search Exploration for Effective UI Testing of Android Apps
by: Demissie, Biniam Fisseha, et al.
Published: (2025)
by: Demissie, Biniam Fisseha, et al.
Published: (2025)
PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
by: Huang, Huihui, et al.
Published: (2026)
by: Huang, Huihui, et al.
Published: (2026)
Out of Distribution, Out of Luck: How Well Can LLMs Trained on Vulnerability Datasets Detect Top 25 CWE Weaknesses?
by: Li, Yikun, et al.
Published: (2025)
by: Li, Yikun, et al.
Published: (2025)
An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning
by: Li, Yikun, et al.
Published: (2026)
by: Li, Yikun, et al.
Published: (2026)
Semantics-Aligned, Curriculum-Driven, and Reasoning-Enhanced Vulnerability Repair Framework
by: Yang, Chengran, et al.
Published: (2025)
by: Yang, Chengran, et al.
Published: (2025)
CovAgent: Overcoming the 30% Curse of Mobile Application Coverage with Agentic AI and Dynamic Instrumentation
by: Minn, Wei, et al.
Published: (2026)
by: Minn, Wei, et al.
Published: (2026)
Beyond Function-Level Analysis: Context-Aware Reasoning for Inter-Procedural Vulnerability Detection
by: Li, Yikun, et al.
Published: (2026)
by: Li, Yikun, et al.
Published: (2026)
Revisiting Sentiment Analysis for Software Engineering in the Era of Large Language Models
by: Zhang, Ting, et al.
Published: (2023)
by: Zhang, Ting, et al.
Published: (2023)
CUPID: Leveraging ChatGPT for More Accurate Duplicate Bug Report Detection
by: Zhang, Ting, et al.
Published: (2023)
by: Zhang, Ting, et al.
Published: (2023)
Security Modelling for Cyber-Physical Systems: A Systematic Literature Review
by: Huang, Shaofei, et al.
Published: (2024)
by: Huang, Shaofei, et al.
Published: (2024)
Bayesian and Multi-Objective Decision Support for Real-Time Incident Mitigation in Critical Infrastructure
by: Huang, Shaofei, et al.
Published: (2025)
by: Huang, Shaofei, et al.
Published: (2025)
From Incomplete Architecture to Quantified Risk: Multimodal LLM-Driven Security Assessment for Cyber-Physical Systems
by: Huang, Shaofei, et al.
Published: (2026)
by: Huang, Shaofei, et al.
Published: (2026)
ACTISM: Threat-informed Dynamic Security Modelling for Automotive Systems
by: Huang, Shaofei, et al.
Published: (2024)
by: Huang, Shaofei, et al.
Published: (2024)
Beyond the Tip of the Iceberg: Understanding SATD in Dockerfiles through the Lens of Co-evolution
by: Minn, Wei, et al.
Published: (2026)
by: Minn, Wei, et al.
Published: (2026)
Virtualization-based Penetration Testing Study for Detecting Accessibility Abuse Vulnerabilities in Banking Apps in East and Southeast Asia
by: Minn, Wei, et al.
Published: (2026)
by: Minn, Wei, et al.
Published: (2026)
Beyond ChatGPT: Enhancing Software Quality Assurance Tasks with Diverse LLMs and Validation Techniques
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
The Price of Prompting: Profiling Energy Use in Large Language Models Inference
by: Husom, Erik Johannes, et al.
Published: (2024)
by: Husom, Erik Johannes, et al.
Published: (2024)
Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead
by: Cheng, Yiran, et al.
Published: (2025)
by: Cheng, Yiran, et al.
Published: (2025)
Demystifying Faulty Code with LLM: Step-by-Step Reasoning for Explainable Fault Localization
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
Deterministic vs. LLM-Controlled Orchestration for COBOL-to-Python Modernization
by: Lwin, Naing Oo, et al.
Published: (2026)
by: Lwin, Naing Oo, et al.
Published: (2026)
Runtime Anomaly Detection for Drones: An Integrated Rule-Mining and Unsupervised-Learning Approach
by: Tan, Ivan, et al.
Published: (2025)
by: Tan, Ivan, et al.
Published: (2025)
Shelving it rather than Ditching it: Dynamically Debloating DEX and Native Methods of Android Applications without APK Modification
by: Zhang, Zicheng, et al.
Published: (2025)
by: Zhang, Zicheng, et al.
Published: (2025)
LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks
by: Zhou, Xin, et al.
Published: (2025)
by: Zhou, Xin, et al.
Published: (2025)
Differentiated Security Architecture for Secure and Efficient Infotainment Data Communication in IoV Networks
by: Fan, Jiani, et al.
Published: (2024)
by: Fan, Jiani, et al.
Published: (2024)
Fixseeker: An Empirical Driven Graph-based Approach for Detecting Silent Vulnerability Fixes in Open Source Software
by: Cheng, Yiran, et al.
Published: (2025)
by: Cheng, Yiran, et al.
Published: (2025)
VERCATION: Precise Vulnerable Open-source Software Version Identification based on Static Analysis and LLM
by: Cheng, Yiran, et al.
Published: (2024)
by: Cheng, Yiran, et al.
Published: (2024)
BugsInPy: A Database of Existing Bugs in Python Programs to Enable Controlled Testing and Debugging Studies
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
by: Widyasari, Ratnadira, et al.
Published: (2023)
by: Widyasari, Ratnadira, et al.
Published: (2023)
Evaluating SZZ Implementations: An Empirical Study on the Linux Kernel
by: Lyu, Yunbo, et al.
Published: (2023)
by: Lyu, Yunbo, et al.
Published: (2023)
Similar Items
-
Let the Trial Begin: A Mock-Court Approach to Vulnerability Detection using LLM-Based Agents
by: Widyasari, Ratnadira, et al.
Published: (2025) -
Back to the Basics: Rethinking Issue-Commit Linking with LLM-Assisted Retrieval
by: Huang, Huihui, et al.
Published: (2025) -
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation
by: Weyssow, Martin, et al.
Published: (2025) -
Revisiting Vulnerability Patch Identification on Data in the Wild
by: Irsan, Ivana Clairine, et al.
Published: (2026) -
VulCoCo: A Simple Yet Effective Method for Detecting Vulnerable Code Clones
by: Bui, Tan, et al.
Published: (2025)