SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs
Fuente:
arXiv
Saved in:
| Main Authors: | Pham, Minh V. T., Phan, Huy N., Phan, Hoang N., Chi, Cuong Le, Nguyen, Tien N., Bui, Nghi D. Q. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CodeFlow: Program Behavior Prediction with Dynamic Dependencies Learning
by: Le, Cuong Chi, et al.
Published: (2024)
by: Le, Cuong Chi, et al.
Published: (2024)
RepoHyper: Search-Expand-Refine on Semantic Graphs for Repository-Level Code Completion
by: Phan, Huy N., et al.
Published: (2024)
by: Phan, Huy N., et al.
Published: (2024)
When Names Disappear: Revealing What LLMs Actually Understand About Code
by: Le, Cuong Chi, et al.
Published: (2025)
by: Le, Cuong Chi, et al.
Published: (2025)
VisualCoder: Guiding Large Language Models in Code Execution with Fine-grained Multimodal Chain-of-Thought Reasoning
by: Le, Cuong Chi, et al.
Published: (2024)
by: Le, Cuong Chi, et al.
Published: (2024)
SpecMind: Cognitively Inspired, Interactive Multi-Turn Framework for Postcondition Inference
by: Le, Cuong Chi, et al.
Published: (2026)
by: Le, Cuong Chi, et al.
Published: (2026)
HyperAgent: Generalist Software Engineering Agents to Solve Coding Tasks at Scale
by: Phan, Huy Nhat, et al.
Published: (2024)
by: Phan, Huy Nhat, et al.
Published: (2024)
TestWeaver: Execution-aware, Feedback-driven Regression Testing Generation with Large Language Models
by: Le, Cuong Chi, et al.
Published: (2025)
by: Le, Cuong Chi, et al.
Published: (2025)
Enhancing Program Repair with Specification Guidance and Intermediate Behavioral Signals
by: Le-Anh, Minh, et al.
Published: (2026)
by: Le-Anh, Minh, et al.
Published: (2026)
AgileCoder: Dynamic Collaborative Agents for Software Development based on Agile Methodology
by: Nguyen, Minh Huynh, et al.
Published: (2024)
by: Nguyen, Minh Huynh, et al.
Published: (2024)
Semantic Evolution over Populations for LLM-Guided Automated Program Repair
by: Le, Cuong Chi, et al.
Published: (2026)
by: Le, Cuong Chi, et al.
Published: (2026)
Documentation-Guided Agentic Codebase Migration from C to Rust
by: Le-Anh, Minh, et al.
Published: (2026)
by: Le-Anh, Minh, et al.
Published: (2026)
CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases
by: Hoang, Anh Nguyen, et al.
Published: (2025)
by: Hoang, Anh Nguyen, et al.
Published: (2025)
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
by: Sonwane, Atharv, et al.
Published: (2025)
by: Sonwane, Atharv, et al.
Published: (2025)
BugsInPy: A Database of Existing Bugs in Python Programs to Enable Controlled Testing and Debugging Studies
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
PanicFI: An Infrastructure for Fixing Panic Bugs in Real-World Rust Programs
by: Ni, Yunbo, et al.
Published: (2024)
by: Ni, Yunbo, et al.
Published: (2024)
BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?
by: Chen, Guoxin, et al.
Published: (2026)
by: Chen, Guoxin, et al.
Published: (2026)
PerfBench: Can Agents Resolve Real-World Performance Bugs?
by: Garg, Spandan, et al.
Published: (2025)
by: Garg, Spandan, et al.
Published: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
by: Manh, Dung Nguyen, et al.
Published: (2024)
by: Manh, Dung Nguyen, et al.
Published: (2024)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
by: Hanna, Carol, et al.
Published: (2025)
by: Hanna, Carol, et al.
Published: (2025)
When Retriever Meets Generator: A Joint Model for Code Comment Generation
by: Le, Tien P. T., et al.
Published: (2025)
by: Le, Tien P. T., et al.
Published: (2025)
GitBug-Actions: Building Reproducible Bug-Fix Benchmarks with GitHub Actions
by: Saavedra, Nuno, et al.
Published: (2023)
by: Saavedra, Nuno, et al.
Published: (2023)
SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents
by: Mündler, Niels, et al.
Published: (2024)
by: Mündler, Niels, et al.
Published: (2024)
Functional Overlap Reranking for Neural Code Generation
by: To, Hung Quoc, et al.
Published: (2023)
by: To, Hung Quoc, et al.
Published: (2023)
A Preliminary Study of Fixed Flaky Tests in Rust Projects on GitHub
by: Schroeder, Tom, et al.
Published: (2025)
by: Schroeder, Tom, et al.
Published: (2025)
Envisioning the Next-Generation AI Coding Assistants: Insights & Proposals
by: Nghiem, Khanh, et al.
Published: (2024)
by: Nghiem, Khanh, et al.
Published: (2024)
SWE-Universe: Scale Real-World Verifiable Environments to Millions
by: Chen, Mouxiang, et al.
Published: (2026)
by: Chen, Mouxiang, et al.
Published: (2026)
PreciseBugCollector: Extensible, Executable and Precise Bug-fix Collection
by: Ye, He, et al.
Published: (2023)
by: Ye, He, et al.
Published: (2023)
Bug Whispering: Towards Audio Bug Reporting
by: Masserini, Elena, et al.
Published: (2025)
by: Masserini, Elena, et al.
Published: (2025)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
by: Li, Dawei, et al.
Published: (2026)
by: Li, Dawei, et al.
Published: (2026)
HAFix: History-Augmented Large Language Models for Bug Fixing
by: Shi, Yu, et al.
Published: (2025)
by: Shi, Yu, et al.
Published: (2025)
Multifaceted Hero Developers and Bug-Fixing Outcomes Across Severity
by: Kumar, Amit, et al.
Published: (2026)
by: Kumar, Amit, et al.
Published: (2026)
Developers' Perception: Fixed Bugs Often Overlooked as Quality Contributions
by: Alifanov, Vitaly, et al.
Published: (2024)
by: Alifanov, Vitaly, et al.
Published: (2024)
BugScope: Learn to Find Bugs Like Human
by: Guo, Jinyao, et al.
Published: (2025)
by: Guo, Jinyao, et al.
Published: (2025)
DocChecker: Bootstrapping Code Large Language Model for Detecting and Resolving Code-Comment Inconsistencies
by: Dau, Anh T. V., et al.
Published: (2023)
by: Dau, Anh T. V., et al.
Published: (2023)
The Limits of Long-Context Reasoning in Automated Bug Fixing
by: Raju, Ravi, et al.
Published: (2026)
by: Raju, Ravi, et al.
Published: (2026)
On the Impacts of Contexts on Repository-Level Code Generation
by: Hai, Nam Le, et al.
Published: (2024)
by: Hai, Nam Le, et al.
Published: (2024)
Towards Understanding the Bugs in Solidity Compiler
by: Ma, Haoyang, et al.
Published: (2024)
by: Ma, Haoyang, et al.
Published: (2024)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
by: Huynh, Hieu, et al.
Published: (2025)
by: Huynh, Hieu, et al.
Published: (2025)
BugGen: A Self-Correcting Multi-Agent LLM Pipeline for Realistic RTL Bug Synthesis
by: Jasper, Surya, et al.
Published: (2025)
by: Jasper, Surya, et al.
Published: (2025)
Similar Items
-
CodeFlow: Program Behavior Prediction with Dynamic Dependencies Learning
by: Le, Cuong Chi, et al.
Published: (2024) -
RepoHyper: Search-Expand-Refine on Semantic Graphs for Repository-Level Code Completion
by: Phan, Huy N., et al.
Published: (2024) -
When Names Disappear: Revealing What LLMs Actually Understand About Code
by: Le, Cuong Chi, et al.
Published: (2025) -
VisualCoder: Guiding Large Language Models in Code Execution with Fine-grained Multimodal Chain-of-Thought Reasoning
by: Le, Cuong Chi, et al.
Published: (2024) -
SpecMind: Cognitively Inspired, Interactive Multi-Turn Framework for Postcondition Inference
by: Le, Cuong Chi, et al.
Published: (2026)