One Bug, Hundreds Behind: LLMs for Large-Scale Bug Discovery
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Qiushi, Xiao, Yue, Kirat, Dhilung, Eykholt, Kevin, Jang, Jiyong, Schales, Douglas Lee |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Lessons from Penetration Tests on Large-Scale Agent Systems
di: Eykholt, Kevin, et al.
Pubblicazione: (2026)
di: Eykholt, Kevin, et al.
Pubblicazione: (2026)
Automated Duplicate Bug Report Detection in Large Open Bug Repositories
di: Laney, Clare E., et al.
Pubblicazione: (2025)
di: Laney, Clare E., et al.
Pubblicazione: (2025)
Bug Analysis Towards Bug Resolution Time Prediction
di: Ozkan, Hasan Yagiz, et al.
Pubblicazione: (2024)
di: Ozkan, Hasan Yagiz, et al.
Pubblicazione: (2024)
Can GPT-O1 Kill All Bugs? An Evaluation of GPT-Family LLMs on QuixBugs
di: Hu, Haichuan, et al.
Pubblicazione: (2024)
di: Hu, Haichuan, et al.
Pubblicazione: (2024)
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
di: Acharya, Jagrit, et al.
Pubblicazione: (2025)
di: Acharya, Jagrit, et al.
Pubblicazione: (2025)
ImproBR: Bug Report Improver Using LLMs
di: Akyol, Emre Furkan, et al.
Pubblicazione: (2026)
di: Akyol, Emre Furkan, et al.
Pubblicazione: (2026)
SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs
di: Pham, Minh V. T., et al.
Pubblicazione: (2025)
di: Pham, Minh V. T., et al.
Pubblicazione: (2025)
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
di: Sonwane, Atharv, et al.
Pubblicazione: (2025)
di: Sonwane, Atharv, et al.
Pubblicazione: (2025)
PyResBugs: A Dataset of Residual Python Bugs for Natural Language-Driven Fault Injection
di: Cotroneo, Domenico, et al.
Pubblicazione: (2025)
di: Cotroneo, Domenico, et al.
Pubblicazione: (2025)
Automated Bug Report Prioritization in Large Open-Source Projects
di: Pierson, Riley, et al.
Pubblicazione: (2025)
di: Pierson, Riley, et al.
Pubblicazione: (2025)
Bugs in Large Language Models Generated Code: An Empirical Study
di: Tambon, Florian, et al.
Pubblicazione: (2024)
di: Tambon, Florian, et al.
Pubblicazione: (2024)
Coding in a Bubble? Evaluating LLMs in Resolving Context Adaptation Bugs During Code Adaptation
di: Zhang, Tanghaoran, et al.
Pubblicazione: (2026)
di: Zhang, Tanghaoran, et al.
Pubblicazione: (2026)
Benchmarking Mythos-Linked Bug Rediscovery
di: David, Isaac, et al.
Pubblicazione: (2026)
di: David, Isaac, et al.
Pubblicazione: (2026)
RLocator: Reinforcement Learning for Bug Localization
di: Chakraborty, Partha, et al.
Pubblicazione: (2023)
di: Chakraborty, Partha, et al.
Pubblicazione: (2023)
Reducing False Positives in Static Bug Detection with LLMs: An Empirical Study in Industry
di: Du, Xueying, et al.
Pubblicazione: (2026)
di: Du, Xueying, et al.
Pubblicazione: (2026)
Bug In the Code Stack: Can LLMs Find Bugs in Large Python Code Stacks
di: Lee, Hokyung, et al.
Pubblicazione: (2024)
di: Lee, Hokyung, et al.
Pubblicazione: (2024)
Are Large Language Models Memorizing Bug Benchmarks?
di: Ramos, Daniel, et al.
Pubblicazione: (2024)
di: Ramos, Daniel, et al.
Pubblicazione: (2024)
BugBlitz-AI: An Intelligent QA Assistant
di: Yao, Yi, et al.
Pubblicazione: (2024)
di: Yao, Yi, et al.
Pubblicazione: (2024)
A Survey of Bugs in AI-Generated Code
di: Gao, Ruofan, et al.
Pubblicazione: (2025)
di: Gao, Ruofan, et al.
Pubblicazione: (2025)
Improving MPI Error Detection and Repair with Large Language Models and Bug References
di: Piersall, Scott, et al.
Pubblicazione: (2026)
di: Piersall, Scott, et al.
Pubblicazione: (2026)
BugSpotter: Automated Generation of Code Debugging Exercises
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2024)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2024)
BLAgent: Agentic RAG for File-Level Bug Localization
di: Mamun, Md Afif Al, et al.
Pubblicazione: (2026)
di: Mamun, Md Afif Al, et al.
Pubblicazione: (2026)
Model See, Model Do? Exposure-Aware Evaluation of Bug-vs-Fix Preference in Code LLMs
di: Al-Kaswan, Ali, et al.
Pubblicazione: (2026)
di: Al-Kaswan, Ali, et al.
Pubblicazione: (2026)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
di: Vulićević, Jelena Ilić
Pubblicazione: (2026)
di: Vulićević, Jelena Ilić
Pubblicazione: (2026)
Agents in the Sandbox: End-to-End Crash Bug Reproduction for Minecraft
di: Yapağcı, Eray, et al.
Pubblicazione: (2025)
di: Yapağcı, Eray, et al.
Pubblicazione: (2025)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
di: Meng, Xiangxin, et al.
Pubblicazione: (2024)
di: Meng, Xiangxin, et al.
Pubblicazione: (2024)
RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols
di: Zheng, Mingwei, et al.
Pubblicazione: (2025)
di: Zheng, Mingwei, et al.
Pubblicazione: (2025)
Dynamic Cogeneration of Bug Reproduction Test in Agentic Program Repair
di: Cheng, Runxiang, et al.
Pubblicazione: (2026)
di: Cheng, Runxiang, et al.
Pubblicazione: (2026)
Go-Oracle: Automated Test Oracle for Go Concurrency Bugs
di: Tsimpourlas, Foivos, et al.
Pubblicazione: (2024)
di: Tsimpourlas, Foivos, et al.
Pubblicazione: (2024)
Past, Present, and Future of Bug Tracking in the Generative AI Era
di: Torun, Utku Boran, et al.
Pubblicazione: (2025)
di: Torun, Utku Boran, et al.
Pubblicazione: (2025)
MarsCode Agent: AI-native Automated Bug Fixing
di: Liu, Yizhou, et al.
Pubblicazione: (2024)
di: Liu, Yizhou, et al.
Pubblicazione: (2024)
Agentic Bug Reproduction for Effective Automated Program Repair at Google
di: Cheng, Runxiang, et al.
Pubblicazione: (2025)
di: Cheng, Runxiang, et al.
Pubblicazione: (2025)
Self-Bootstrapping Automated Program Repair: Using LLMs to Generate and Evaluate Synthetic Training Data for Bug Repair
di: de-Fitero-Dominguez, David, et al.
Pubblicazione: (2025)
di: de-Fitero-Dominguez, David, et al.
Pubblicazione: (2025)
HLSDebugger: Identification and Correction of Logic Bugs in HLS Code with LLM Solutions
di: Wang, Jing, et al.
Pubblicazione: (2025)
di: Wang, Jing, et al.
Pubblicazione: (2025)
Agentic Property-Based Testing: Finding Bugs Across the Python Ecosystem
di: Maaz, Muhammad, et al.
Pubblicazione: (2025)
di: Maaz, Muhammad, et al.
Pubblicazione: (2025)
Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
di: Nirujan, Hinduja, et al.
Pubblicazione: (2026)
di: Nirujan, Hinduja, et al.
Pubblicazione: (2026)
Faster Configuration Performance Bug Testing with Neural Dual-level Prioritization
di: Ma, Youpeng, et al.
Pubblicazione: (2025)
di: Ma, Youpeng, et al.
Pubblicazione: (2025)
Fine-Tuning Code Language Models to Detect Cross-Language Bugs
di: Li, Zengyang, et al.
Pubblicazione: (2025)
di: Li, Zengyang, et al.
Pubblicazione: (2025)
Agora: Toward Autonomous Bug Detection in Production-Level Consensus Protocols with LLM Agents
di: Liu, Xiang, et al.
Pubblicazione: (2026)
di: Liu, Xiang, et al.
Pubblicazione: (2026)
Bug Destiny Prediction in Large Open-Source Software Repositories through Sentiment Analysis and BERT Topic Modeling
di: Pope, Sophie C., et al.
Pubblicazione: (2025)
di: Pope, Sophie C., et al.
Pubblicazione: (2025)
Documenti analoghi
-
Lessons from Penetration Tests on Large-Scale Agent Systems
di: Eykholt, Kevin, et al.
Pubblicazione: (2026) -
Automated Duplicate Bug Report Detection in Large Open Bug Repositories
di: Laney, Clare E., et al.
Pubblicazione: (2025) -
Bug Analysis Towards Bug Resolution Time Prediction
di: Ozkan, Hasan Yagiz, et al.
Pubblicazione: (2024) -
Can GPT-O1 Kill All Bugs? An Evaluation of GPT-Family LLMs on QuixBugs
di: Hu, Haichuan, et al.
Pubblicazione: (2024) -
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
di: Acharya, Jagrit, et al.
Pubblicazione: (2025)