OmniGIRL: A Multilingual and Multimodal Benchmark for GitHub Issue Resolution
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Lianghong, Tao, Wei, Jiang, Runhan, Wang, Yanlin, Chen, Jiachi, Liu, Xilin, Ma, Yuchi, Mao, Mingzhi, Zhang, Hongyu, Zheng, Zibin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
by: Tao, Wei, et al.
Published: (2024)
by: Tao, Wei, et al.
Published: (2024)
RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
by: Wang, Yanli, et al.
Published: (2024)
by: Wang, Yanli, et al.
Published: (2024)
An Empirical Study of ChatGPT-Related Projects and Their Issues on GitHub
by: Lin, Zheng, et al.
Published: (2024)
by: Lin, Zheng, et al.
Published: (2024)
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks
by: Guo, Lianghong, et al.
Published: (2025)
by: Guo, Lianghong, et al.
Published: (2025)
SimpleDevQA: Benchmarking Large Language Models on Development Knowledge QA
by: Zhang, Jing, et al.
Published: (2025)
by: Zhang, Jing, et al.
Published: (2025)
When to Stop? Towards Efficient Code Generation in LLMs with Excess Token Prevention
by: Guo, Lianghong, et al.
Published: (2024)
by: Guo, Lianghong, et al.
Published: (2024)
"My GitHub Sponsors profile is live!" Investigating the Impact of Twitter/X Mentions on GitHub Sponsors
by: Fan, Youmei, et al.
Published: (2024)
by: Fan, Youmei, et al.
Published: (2024)
GitHub Proxy Server: A tool for supporting massive data collection on GitHub
by: Borges, Hudson Silva, et al.
Published: (2025)
by: Borges, Hudson Silva, et al.
Published: (2025)
Classifying Issues in Open-source GitHub Repositories
by: Raaj, Amir Hossain, et al.
Published: (2025)
by: Raaj, Amir Hossain, et al.
Published: (2025)
Guidelines for Developing Bots for GitHub
by: Wessel, Mairieli, et al.
Published: (2022)
by: Wessel, Mairieli, et al.
Published: (2022)
Prioritising GitHub Priority Labels
by: Caddy, James, et al.
Published: (2024)
by: Caddy, James, et al.
Published: (2024)
GitBug-Actions: Building Reproducible Bug-Fix Benchmarks with GitHub Actions
by: Saavedra, Nuno, et al.
Published: (2023)
by: Saavedra, Nuno, et al.
Published: (2023)
Can GitHub Issues Help in App Review Classifications?
by: Abedini, Yasaman, et al.
Published: (2023)
by: Abedini, Yasaman, et al.
Published: (2023)
What Makes a GitHub Issue Ready for Copilot?
by: Sayagh, Mohammed
Published: (2025)
by: Sayagh, Mohammed
Published: (2025)
Top General Performance = Top Domain Performance? DomainCodeBench: A Multi-domain Code Generation Benchmark
by: Zheng, Dewu, et al.
Published: (2024)
by: Zheng, Dewu, et al.
Published: (2024)
SWE-bench-java: A GitHub Issue Resolving Benchmark for Java
by: Zan, Daoguang, et al.
Published: (2024)
by: Zan, Daoguang, et al.
Published: (2024)
What Drives Issue Resolution Speed? An Empirical Study of Scientific Workflow Systems on GitHub
by: Alam, Khairul, et al.
Published: (2025)
by: Alam, Khairul, et al.
Published: (2025)
Characterizing the Failure Modes of LLMs in Resolving Real-World GitHub Issues
by: Jiang, Yanjie, et al.
Published: (2026)
by: Jiang, Yanjie, et al.
Published: (2026)
Visual Analysis of GitHub Issues to Gain Insights
by: Proma, Rifat Ara, et al.
Published: (2024)
by: Proma, Rifat Ara, et al.
Published: (2024)
The Impact of Sanctions on GitHub Developers and Activities
by: Fan, Youmei, et al.
Published: (2024)
by: Fan, Youmei, et al.
Published: (2024)
Fingerprinting AI Coding Agents on GitHub
by: Ghaleb, Taher A.
Published: (2026)
by: Ghaleb, Taher A.
Published: (2026)
How Complex is a GitHub Issue?
by: Anonymous
Published: (2025)
by: Anonymous
Published: (2025)
GitHub Marketplace for Automation and Innovation in Software Production
by: Saroar, SK Golam, et al.
Published: (2024)
by: Saroar, SK Golam, et al.
Published: (2024)
Introducing Traceability in GitHub for Medical Software Development
by: Stirbu, Vlad, et al.
Published: (2021)
by: Stirbu, Vlad, et al.
Published: (2021)
The Landscape of Toxicity: An Empirical Investigation of Toxicity on GitHub
by: Sarker, Jaydeb, et al.
Published: (2025)
by: Sarker, Jaydeb, et al.
Published: (2025)
Chaos Engineering in the Wild: Findings from GitHub
by: Owotogbe, Joshua, et al.
Published: (2025)
by: Owotogbe, Joshua, et al.
Published: (2025)
An Empirical Study of the Evolution of GitHub Actions Workflows
by: Mazrae, Pooya Rostami, et al.
Published: (2026)
by: Mazrae, Pooya Rostami, et al.
Published: (2026)
GitHub Actions: The Impact on the Pull Request Process
by: Wessel, Mairieli, et al.
Published: (2022)
by: Wessel, Mairieli, et al.
Published: (2022)
Agentic Much? Adoption of Coding Agents on GitHub
by: Robbes, Romain, et al.
Published: (2026)
by: Robbes, Romain, et al.
Published: (2026)
Understanding and Predicting Derailment in Toxic Conversations on GitHub
by: Imran, Mia Mohammad, et al.
Published: (2025)
by: Imran, Mia Mohammad, et al.
Published: (2025)
EffiReasonTrans: RL-Optimized Reasoning for Code Translation
by: Wang, Yanlin, et al.
Published: (2025)
by: Wang, Yanlin, et al.
Published: (2025)
Advances and Frontiers of LLM-based Issue Resolution in Software Engineering: A Comprehensive Survey
by: Li, Caihua, et al.
Published: (2026)
by: Li, Caihua, et al.
Published: (2026)
On the GitHub Actions Language: Usage, Evolution, and Workflow Reliability
by: Bardsiri, Aref Talebzadeh, et al.
Published: (2026)
by: Bardsiri, Aref Talebzadeh, et al.
Published: (2026)
Designing for Cognitive Diversity: Improving the GitHub Experience for Newcomers
by: Santos, Italo, et al.
Published: (2023)
by: Santos, Italo, et al.
Published: (2023)
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery
by: Zhao, Jiale, et al.
Published: (2026)
by: Zhao, Jiale, et al.
Published: (2026)
DRAINCODE: Stealthy Energy Consumption Attacks on Retrieval-Augmented Code Generation via Context Poisoning
by: Wang, Yanlin, et al.
Published: (2026)
by: Wang, Yanlin, et al.
Published: (2026)
GitHub Copilot: the perfect Code compLeeter?
by: Siroš, Ilja, et al.
Published: (2024)
by: Siroš, Ilja, et al.
Published: (2024)
AIDev: Studying AI Coding Agents on GitHub
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Unpacking Security Scanners for GitHub Actions Workflows
by: Fares, Madjda, et al.
Published: (2026)
by: Fares, Madjda, et al.
Published: (2026)
Similar Items
-
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
by: Tao, Wei, et al.
Published: (2024) -
RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
by: Wang, Yanli, et al.
Published: (2024) -
An Empirical Study of ChatGPT-Related Projects and Their Issues on GitHub
by: Lin, Zheng, et al.
Published: (2024) -
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
by: Wang, Yanlin, et al.
Published: (2024) -
SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks
by: Guo, Lianghong, et al.
Published: (2025)