Diffusion is a code repair operator and generator
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Mukul, Verbruggen, Gust, Le, Vu, Gulwani, Sumit |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tabularis Formatus: Predictive Formatting for Tables
von: Singh, Mukul, et al.
Veröffentlicht: (2025)
von: Singh, Mukul, et al.
Veröffentlicht: (2025)
Semantically Aligned Question and Code Generation for Automated Insight Generation
von: Singha, Ananya, et al.
Veröffentlicht: (2024)
von: Singha, Ananya, et al.
Veröffentlicht: (2024)
Do Code Models Suffer from the Dunning-Kruger Effect?
von: Singh, Mukul, et al.
Veröffentlicht: (2025)
von: Singh, Mukul, et al.
Veröffentlicht: (2025)
LLM-Guided Compositional Program Synthesis
von: Khan, Ruhma, et al.
Veröffentlicht: (2025)
von: Khan, Ruhma, et al.
Veröffentlicht: (2025)
An Empirical Study of Validating Synthetic Data for Formula Generation
von: Singh, Usneek, et al.
Veröffentlicht: (2024)
von: Singh, Usneek, et al.
Veröffentlicht: (2024)
An evaluation of LLM code generation capabilities through graded exercises
von: Jiménez, Álvaro Barbero
Veröffentlicht: (2024)
von: Jiménez, Álvaro Barbero
Veröffentlicht: (2024)
Comparing large language models and human programmers for generating programming code
von: Hou, Wenpin, et al.
Veröffentlicht: (2024)
von: Hou, Wenpin, et al.
Veröffentlicht: (2024)
TableTalk: Scaffolding Spreadsheet Development with a Language Agent
von: Liang, Jenny T., et al.
Veröffentlicht: (2025)
von: Liang, Jenny T., et al.
Veröffentlicht: (2025)
Unmasking the giant: A comprehensive evaluation of ChatGPT's proficiency in coding algorithms and data structures
von: Arefin, Sayed Erfan, et al.
Veröffentlicht: (2023)
von: Arefin, Sayed Erfan, et al.
Veröffentlicht: (2023)
TeamUp: Semantic Project Matching and Team Formation for Learning at Scale
von: Gulwani, Dhruv, et al.
Veröffentlicht: (2026)
von: Gulwani, Dhruv, et al.
Veröffentlicht: (2026)
Exploring Interaction Patterns for Debugging: Enhancing Conversational Capabilities of AI-assistants
von: Chopra, Bhavya, et al.
Veröffentlicht: (2024)
von: Chopra, Bhavya, et al.
Veröffentlicht: (2024)
IRepair: An Intent-Aware Approach to Repair Data-Driven Errors in Large Language Models
von: Imtiaz, Sayem Mohammad, et al.
Veröffentlicht: (2025)
von: Imtiaz, Sayem Mohammad, et al.
Veröffentlicht: (2025)
SQLong: Enhanced NL2SQL for Longer Contexts with LLMs
von: Nguyen, Dai Quoc, et al.
Veröffentlicht: (2025)
von: Nguyen, Dai Quoc, et al.
Veröffentlicht: (2025)
Beyond Correctness: Benchmarking Multi-dimensional Code Generation for Large Language Models
von: Zheng, Jiasheng, et al.
Veröffentlicht: (2024)
von: Zheng, Jiasheng, et al.
Veröffentlicht: (2024)
SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
von: Sonwane, Atharv, et al.
Veröffentlicht: (2025)
von: Sonwane, Atharv, et al.
Veröffentlicht: (2025)
code_transformed: The Influence of Large Language Models on Code
von: Xu, Yuliang, et al.
Veröffentlicht: (2025)
von: Xu, Yuliang, et al.
Veröffentlicht: (2025)
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2025)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2025)
OpenHands: An Open Platform for AI Software Developers as Generalist Agents
von: Wang, Xingyao, et al.
Veröffentlicht: (2024)
von: Wang, Xingyao, et al.
Veröffentlicht: (2024)
CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks
von: Jiang, Hongchao, et al.
Veröffentlicht: (2025)
von: Jiang, Hongchao, et al.
Veröffentlicht: (2025)
Pull Requests as a Training Signal for Repo-Level Code Editing
von: Zhu, Qinglin, et al.
Veröffentlicht: (2026)
von: Zhu, Qinglin, et al.
Veröffentlicht: (2026)
BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation
von: Lai, Peng, et al.
Veröffentlicht: (2026)
von: Lai, Peng, et al.
Veröffentlicht: (2026)
CangjieBench: Benchmarking LLMs on a Low-Resource General-Purpose Programming Language
von: Cheng, Junhang, et al.
Veröffentlicht: (2026)
von: Cheng, Junhang, et al.
Veröffentlicht: (2026)
Software Mention Recognition with a Three-Stage Framework Based on BERTology Models at SOMD 2024
von: Thi, Thuy Nguyen, et al.
Veröffentlicht: (2024)
von: Thi, Thuy Nguyen, et al.
Veröffentlicht: (2024)
Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Application
von: Oliveira, William
Veröffentlicht: (2026)
von: Oliveira, William
Veröffentlicht: (2026)
Library Drift: Diagnosing and Fixing a Silent Failure Mode in Self-Evolving LLM Skill Libraries
von: Zhang, Xing, et al.
Veröffentlicht: (2026)
von: Zhang, Xing, et al.
Veröffentlicht: (2026)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
von: Zhong, Li, et al.
Veröffentlicht: (2024)
von: Zhong, Li, et al.
Veröffentlicht: (2024)
Text2BIM: Generating Building Models Using a Large Language Model-based Multi-Agent Framework
von: Du, Changyu, et al.
Veröffentlicht: (2024)
von: Du, Changyu, et al.
Veröffentlicht: (2024)
An Empirical Exploration of ChatGPT's Ability to Support Problem Formulation Tasks for Mission Engineering and a Documentation of its Performance Variability
von: Ofsa, Max, et al.
Veröffentlicht: (2025)
von: Ofsa, Max, et al.
Veröffentlicht: (2025)
Granite Code Models: A Family of Open Foundation Models for Code Intelligence
von: Mishra, Mayank, et al.
Veröffentlicht: (2024)
von: Mishra, Mayank, et al.
Veröffentlicht: (2024)
State-of-the-art Small Language Coder Model: Mify-Coder
von: Parmar, Abhinav, et al.
Veröffentlicht: (2025)
von: Parmar, Abhinav, et al.
Veröffentlicht: (2025)
debug-gym: A Text-Based Environment for Interactive Debugging
von: Yuan, Xingdi, et al.
Veröffentlicht: (2025)
von: Yuan, Xingdi, et al.
Veröffentlicht: (2025)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
An Empirical Study on Failures in Automated Issue Solving
von: Liu, Simiao, et al.
Veröffentlicht: (2025)
von: Liu, Simiao, et al.
Veröffentlicht: (2025)
RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation
von: Luo, Jane, et al.
Veröffentlicht: (2025)
von: Luo, Jane, et al.
Veröffentlicht: (2025)
Dissecting the SWE-Bench Leaderboards: Profiling Submitters and Architectures of LLM- and Agent-Based Repair Systems
von: Martinez, Matias, et al.
Veröffentlicht: (2025)
von: Martinez, Matias, et al.
Veröffentlicht: (2025)
Towards Automated Smart Contract Generation: Evaluation, Benchmarking, and Retrieval-Augmented Repair
von: Chen, Zaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Zaoyu, et al.
Veröffentlicht: (2025)
SEW: Self-Evolving Agentic Workflows for Automated Code Generation
von: Liu, Siwei, et al.
Veröffentlicht: (2025)
von: Liu, Siwei, et al.
Veröffentlicht: (2025)
E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task
von: Liu, Jingyao, et al.
Veröffentlicht: (2025)
von: Liu, Jingyao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Tabularis Formatus: Predictive Formatting for Tables
von: Singh, Mukul, et al.
Veröffentlicht: (2025) -
Semantically Aligned Question and Code Generation for Automated Insight Generation
von: Singha, Ananya, et al.
Veröffentlicht: (2024) -
Do Code Models Suffer from the Dunning-Kruger Effect?
von: Singh, Mukul, et al.
Veröffentlicht: (2025) -
LLM-Guided Compositional Program Synthesis
von: Khan, Ruhma, et al.
Veröffentlicht: (2025) -
An Empirical Study of Validating Synthetic Data for Formula Generation
von: Singh, Usneek, et al.
Veröffentlicht: (2024)