Sifting the Noise: A Comparative Study of LLM Agents in Vulnerability False Positive Filtering
Fuente:
arXiv
Saved in:
| Main Authors: | Xiong, Yunpeng, Zhang, Ting |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Empirical Study of False Negatives and Positives of Static Code Analyzers From the Perspective of Historical Issues
by: Cui, Han, et al.
Published: (2024)
by: Cui, Han, et al.
Published: (2024)
How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection
by: Chen, Maofei, et al.
Published: (2026)
by: Chen, Maofei, et al.
Published: (2026)
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
by: Aggarwal, Pooja, et al.
Published: (2024)
by: Aggarwal, Pooja, et al.
Published: (2024)
LLM-based Vulnerability Detection at Project Scale: An Empirical Study
by: Li, Fengjie, et al.
Published: (2026)
by: Li, Fengjie, et al.
Published: (2026)
Minimizing False Positives in Static Bug Detection via LLM-Enhanced Path Feasibility Analysis
by: Du, Xueying, et al.
Published: (2025)
by: Du, Xueying, et al.
Published: (2025)
Sifting through the Chaff: On Utilizing Execution Feedback for Ranking the Generated Code Candidates
by: Sun, Zhihong, et al.
Published: (2024)
by: Sun, Zhihong, et al.
Published: (2024)
QASecClaw: A Multi-Agent LLM Approach for False Positive Reduction in Static Application Security Testing
by: Ameen, Mohd Ruhul, et al.
Published: (2026)
by: Ameen, Mohd Ruhul, et al.
Published: (2026)
AgentFM: Role-Aware Failure Management for Distributed Databases with LLM-Driven Multi-Agents
by: Zhang, Lingzhe, et al.
Published: (2025)
by: Zhang, Lingzhe, et al.
Published: (2025)
VulnResolver: A Hybrid Agent Framework for LLM-Based Automated Vulnerability Issue Resolution
by: Zhang, Mingming, et al.
Published: (2026)
by: Zhang, Mingming, et al.
Published: (2026)
Characterizing and Mitigating False-Positive Bug Reports in the Linux Kernel
by: Tian, Jiashuo, et al.
Published: (2026)
by: Tian, Jiashuo, et al.
Published: (2026)
FP-Predictor - False Positive Prediction for Static Analysis Reports
by: Ohlmer, Tom, et al.
Published: (2026)
by: Ohlmer, Tom, et al.
Published: (2026)
Reducing False Positives in Static Bug Detection with LLMs: An Empirical Study in Industry
by: Du, Xueying, et al.
Published: (2026)
by: Du, Xueying, et al.
Published: (2026)
FaultLine: Automated Proof-of-Vulnerability Generation Using LLM Agents
by: Nitin, Vikram, et al.
Published: (2025)
by: Nitin, Vikram, et al.
Published: (2025)
LLM-Driven Adaptive Source-Sink Identification and False Positive Mitigation for Static Analysis
by: Lin, Shiyin
Published: (2025)
by: Lin, Shiyin
Published: (2025)
Bridging the Gap: A Comparative Study of Academic and Developer Approaches to Smart Contract Vulnerabilities
by: Salzano, Francesco, et al.
Published: (2025)
by: Salzano, Francesco, et al.
Published: (2025)
Revisiting Vulnerability Patch Localization: An Empirical Study and LLM-Based Solution
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
Detecting State Manipulation Vulnerabilities in Smart Contracts Using LLM and Static Analysis
by: Wu, Hao, et al.
Published: (2025)
by: Wu, Hao, et al.
Published: (2025)
Mitigating False Positives in Static Memory Safety Analysis of Rust Programs via Reinforcement Learning
by: P, Akilesh, et al.
Published: (2026)
by: P, Akilesh, et al.
Published: (2026)
Multi-Agent End-to-End Vulnerability Management for Mitigating Recurring Vulnerabilities
by: Zheng, Zelong, et al.
Published: (2026)
by: Zheng, Zelong, et al.
Published: (2026)
Beyond Translation Accuracy: Addressing False Failures in LLM-Based Code Translation
by: Rabbi, Fazle, et al.
Published: (2026)
by: Rabbi, Fazle, et al.
Published: (2026)
Large Language Model for Vulnerability Detection: Emerging Results and Future Directions
by: Zhou, Xin, et al.
Published: (2024)
by: Zhou, Xin, et al.
Published: (2024)
Inducing Vulnerable Code Generation in LLM Coding Assistants
by: Zeng, Binqi, et al.
Published: (2025)
by: Zeng, Binqi, et al.
Published: (2025)
LLM Agents for Automated Web Vulnerability Reproduction: Are We There Yet?
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
A Case Study of LLM for Automated Vulnerability Repair: Assessing Impact of Reasoning and Patch Validation Feedback
by: Kulsum, Ummay, et al.
Published: (2024)
by: Kulsum, Ummay, et al.
Published: (2024)
An Empirical Study of Bugs in Modern LLM Agent Frameworks
by: Zhu, Xinxue, et al.
Published: (2026)
by: Zhu, Xinxue, et al.
Published: (2026)
A Dual-Loop Agent Framework for Automated Vulnerability Reproduction
by: Liu, Bin, et al.
Published: (2026)
by: Liu, Bin, et al.
Published: (2026)
VulAgent: Hypothesis-Validation based Multi-Agent Vulnerability Detection
by: Wang, Ziliang, et al.
Published: (2025)
by: Wang, Ziliang, et al.
Published: (2025)
Human or LLM? A Comparative Study on Accessible Code Generation Capability
by: Suh, Hyunjae, et al.
Published: (2025)
by: Suh, Hyunjae, et al.
Published: (2025)
An Empirical Study of Vulnerable Package Dependencies in LLM Repositories
by: Liu, Shuhan, et al.
Published: (2025)
by: Liu, Shuhan, et al.
Published: (2025)
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
by: Islam, Niful, et al.
Published: (2026)
by: Islam, Niful, et al.
Published: (2026)
Fixing Smart Contract Vulnerabilities: A Comparative Analysis of Literature and Developer's Practices
by: Salzano, Francesco, et al.
Published: (2024)
by: Salzano, Francesco, et al.
Published: (2024)
Code Vulnerability Detection: A Comparative Analysis of Emerging Large Language Models
by: Sultana, Shaznin, et al.
Published: (2024)
by: Sultana, Shaznin, et al.
Published: (2024)
Stop Comparing LLM Agents Without Disclosing the Harness
by: Zhang, Yunbei, et al.
Published: (2026)
by: Zhang, Yunbei, et al.
Published: (2026)
Agent That Debugs: Dynamic State-Guided Vulnerability Repair
by: Liu, Zhengyao, et al.
Published: (2025)
by: Liu, Zhengyao, et al.
Published: (2025)
MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation
by: Wang, Guanyu, et al.
Published: (2024)
by: Wang, Guanyu, et al.
Published: (2024)
Why Agentic-PRs Get Rejected: A Comparative Study of Coding Agents
by: Nakashima, Sota, et al.
Published: (2026)
by: Nakashima, Sota, et al.
Published: (2026)
M2CVD: Enhancing Vulnerability Semantic through Multi-Model Collaboration for Code Vulnerability Detection
by: Wang, Ziliang, et al.
Published: (2024)
by: Wang, Ziliang, et al.
Published: (2024)
Code Change Intention, Development Artifact and History Vulnerability: Putting Them Together for Vulnerability Fix Detection by LLM
by: Yang, Xu, et al.
Published: (2025)
by: Yang, Xu, et al.
Published: (2025)
Benchmarking and Studying the LLM-based Agent System in End-to-End Software Development
by: Zeng, Zhengran, et al.
Published: (2025)
by: Zeng, Zhengran, et al.
Published: (2025)
Multitask-based Evaluation of Open-Source LLM on Software Vulnerability
by: Yin, Xin, et al.
Published: (2024)
by: Yin, Xin, et al.
Published: (2024)
Similar Items
-
An Empirical Study of False Negatives and Positives of Static Code Analyzers From the Perspective of Historical Issues
by: Cui, Han, et al.
Published: (2024) -
How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection
by: Chen, Maofei, et al.
Published: (2026) -
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
by: Aggarwal, Pooja, et al.
Published: (2024) -
LLM-based Vulnerability Detection at Project Scale: An Empirical Study
by: Li, Fengjie, et al.
Published: (2026) -
Minimizing False Positives in Static Bug Detection via LLM-Enhanced Path Feasibility Analysis
by: Du, Xueying, et al.
Published: (2025)