SWE-bench-java: A GitHub Issue Resolving Benchmark for Java
Fuente:
arXiv
Saved in:
| Main Authors: | Zan, Daoguang, Huang, Zhirong, Yu, Ailun, Lin, Shaoxin, Shi, Yifan, Liu, Wei, Chen, Dong, Qi, Zongshuai, Yu, Hao, Yu, Lei, Ran, Dezhi, Zeng, Muhan, Shen, Bo, Bian, Pan, Liang, Guangtai, Guan, Bei, Huang, Pengjie, Xie, Tao, Wang, Yongji, Wang, Qianxiang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
by: Jimenez, Carlos E., et al.
Published: (2023)
by: Jimenez, Carlos E., et al.
Published: (2023)
Improving Natural Language Capability of Code Large Language Model
by: Li, Wei, et al.
Published: (2024)
by: Li, Wei, et al.
Published: (2024)
CodeV: Issue Resolving with Visual Data
by: Zhang, Linhao, et al.
Published: (2024)
by: Zhang, Linhao, et al.
Published: (2024)
Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving
by: Zan, Daoguang, et al.
Published: (2025)
by: Zan, Daoguang, et al.
Published: (2025)
CodeR: Issue Resolving with Multi-Agent and Task Graphs
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
Characterizing the Failure Modes of LLMs in Resolving Real-World GitHub Issues
by: Jiang, Yanjie, et al.
Published: (2026)
by: Jiang, Yanjie, et al.
Published: (2026)
GraphCoder: Enhancing Repository-Level Code Completion via Code Context Graph-based Retrieval and Language Model
by: Liu, Wei, et al.
Published: (2024)
by: Liu, Wei, et al.
Published: (2024)
LEAN-GitHub: Compiling GitHub LEAN repositories for a versatile LEAN prover
by: Wu, Zijian, et al.
Published: (2024)
by: Wu, Zijian, et al.
Published: (2024)
A GAN-based data poisoning framework against anomaly detection in vertical federated learning
by: Chen, Xiaolin, et al.
Published: (2024)
by: Chen, Xiaolin, et al.
Published: (2024)
SWE-Mirror: Scaling Issue-Resolving Datasets by Mirroring Issues Across Repositories
by: Wang, Junhao, et al.
Published: (2025)
by: Wang, Junhao, et al.
Published: (2025)
Imago GitHub training
by: Elsey, Jonathan
Published: (2026)
by: Elsey, Jonathan
Published: (2026)
SWE-Fixer: Training Open-Source LLMs for Effective and Efficient GitHub Issue Resolution
by: Xie, Chengxing, et al.
Published: (2025)
by: Xie, Chengxing, et al.
Published: (2025)
GitHub Proxy Server: A tool for supporting massive data collection on GitHub
by: Borges, Hudson Silva, et al.
Published: (2025)
by: Borges, Hudson Silva, et al.
Published: (2025)
CAM: A Collection of Snapshots of GitHub Java Repositories Together with Metrics
by: Bugayenko, Yegor
Published: (2024)
by: Bugayenko, Yegor
Published: (2024)
CodeS: Natural Language to Code Repository via Multi-Layer Sketch
by: Zan, Daoguang, et al.
Published: (2024)
by: Zan, Daoguang, et al.
Published: (2024)
SWE-bench Goes Live!
by: Zhang, Linghao, et al.
Published: (2025)
by: Zhang, Linghao, et al.
Published: (2025)
METXico Innovation CODE GitHub
by: Xool-Tamayo, Jorge, et al.
Published: (2023)
by: Xool-Tamayo, Jorge, et al.
Published: (2023)
Guidelines for Developing Bots for GitHub
by: Wessel, Mairieli, et al.
Published: (2022)
by: Wessel, Mairieli, et al.
Published: (2022)
Prioritising GitHub Priority Labels
by: Caddy, James, et al.
Published: (2024)
by: Caddy, James, et al.
Published: (2024)
How Do Java Developers Reuse StackOverflow Answers in Their GitHub Projects?
by: Chen, Juntong, et al.
Published: (2023)
by: Chen, Juntong, et al.
Published: (2023)
"My GitHub Sponsors profile is live!" Investigating the Impact of Twitter/X Mentions on GitHub Sponsors
by: Fan, Youmei, et al.
Published: (2024)
by: Fan, Youmei, et al.
Published: (2024)
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
by: Tao, Wei, et al.
Published: (2024)
by: Tao, Wei, et al.
Published: (2024)
Weaponizing the Commons: A Taxonomy and Detection Framework of Abuse on GitHub
by: Cheng, Yuli, et al.
Published: (2026)
by: Cheng, Yuli, et al.
Published: (2026)
SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution
by: Li, Han, et al.
Published: (2025)
by: Li, Han, et al.
Published: (2025)
How Complex is a GitHub Issue?
by: Anonymous
Published: (2025)
by: Anonymous
Published: (2025)
The Impact of Sanctions on GitHub Developers and Activities
by: Fan, Youmei, et al.
Published: (2024)
by: Fan, Youmei, et al.
Published: (2024)
Fingerprinting AI Coding Agents on GitHub
by: Ghaleb, Taher A.
Published: (2026)
by: Ghaleb, Taher A.
Published: (2026)
Where Is Self-admitted Code Generated by Large Language Models on GitHub?
by: Yu, Xiao, et al.
Published: (2024)
by: Yu, Xiao, et al.
Published: (2024)
SwingArena: Competitive Programming Arena for Long-context GitHub Issue Solving
by: Xu, Wendong, et al.
Published: (2025)
by: Xu, Wendong, et al.
Published: (2025)
Automating the Detection of Code Vulnerabilities by Analyzing GitHub Issues
by: Cipollone, Daniele, et al.
Published: (2025)
by: Cipollone, Daniele, et al.
Published: (2025)
Security Weaknesses of Copilot-Generated Code in GitHub Projects: An Empirical Study
by: Fu, Yujia, et al.
Published: (2023)
by: Fu, Yujia, et al.
Published: (2023)
GitHub Marketplace for Automation and Innovation in Software Production
by: Saroar, SK Golam, et al.
Published: (2024)
by: Saroar, SK Golam, et al.
Published: (2024)
Introducing Traceability in GitHub for Medical Software Development
by: Stirbu, Vlad, et al.
Published: (2021)
by: Stirbu, Vlad, et al.
Published: (2021)
The Landscape of Toxicity: An Empirical Investigation of Toxicity on GitHub
by: Sarker, Jaydeb, et al.
Published: (2025)
by: Sarker, Jaydeb, et al.
Published: (2025)
Chaos Engineering in the Wild: Findings from GitHub
by: Owotogbe, Joshua, et al.
Published: (2025)
by: Owotogbe, Joshua, et al.
Published: (2025)
GitHub Copilot: the perfect Code compLeeter?
by: Siroš, Ilja, et al.
Published: (2024)
by: Siroš, Ilja, et al.
Published: (2024)
AIDev: Studying AI Coding Agents on GitHub
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Classifying Issues in Open-source GitHub Repositories
by: Raaj, Amir Hossain, et al.
Published: (2025)
by: Raaj, Amir Hossain, et al.
Published: (2025)
An Empirical Study of the Evolution of GitHub Actions Workflows
by: Mazrae, Pooya Rostami, et al.
Published: (2026)
by: Mazrae, Pooya Rostami, et al.
Published: (2026)
Visual Analysis of GitHub Issues to Gain Insights
by: Proma, Rifat Ara, et al.
Published: (2024)
by: Proma, Rifat Ara, et al.
Published: (2024)
Similar Items
-
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
by: Jimenez, Carlos E., et al.
Published: (2023) -
Improving Natural Language Capability of Code Large Language Model
by: Li, Wei, et al.
Published: (2024) -
CodeV: Issue Resolving with Visual Data
by: Zhang, Linhao, et al.
Published: (2024) -
Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving
by: Zan, Daoguang, et al.
Published: (2025) -
CodeR: Issue Resolving with Multi-Agent and Task Graphs
by: Chen, Dong, et al.
Published: (2024)