AuPair: Golden Example Pairs for Code Repair
Fuente:
arXiv
Guardado en:
| Autores principales: | Mavalankar, Aditi, Mansoor, Hassan, Marinho, Zita, Samsikova, Masha, Schaul, Tom |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback
por: Hajizadeh, Samira, et al.
Publicado: (2026)
por: Hajizadeh, Samira, et al.
Publicado: (2026)
The Art of Repair: Optimizing Iterative Program Repair with Instruction-Tuned Models
por: Ruiz, Fernando Vallecillos, et al.
Publicado: (2025)
por: Ruiz, Fernando Vallecillos, et al.
Publicado: (2025)
Is Programming by Example solved by LLMs?
por: Li, Wen-Ding, et al.
Publicado: (2024)
por: Li, Wen-Ding, et al.
Publicado: (2024)
Automated Unity Game Template Generation from GDDs via NLP and Multi-Modal LLMs
por: Hassan, Amna
Publicado: (2025)
por: Hassan, Amna
Publicado: (2025)
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
por: Yang, Dayu, et al.
Publicado: (2025)
por: Yang, Dayu, et al.
Publicado: (2025)
The Impact of Fine-tuning Large Language Models on Automated Program Repair
por: Macháček, Roman, et al.
Publicado: (2025)
por: Macháček, Roman, et al.
Publicado: (2025)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
por: Fan, Lishui, et al.
Publicado: (2025)
por: Fan, Lishui, et al.
Publicado: (2025)
Let the Code LLM Edit Itself When You Edit the Code
por: He, Zhenyu, et al.
Publicado: (2024)
por: He, Zhenyu, et al.
Publicado: (2024)
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
por: Guo, Jiawei, et al.
Publicado: (2024)
por: Guo, Jiawei, et al.
Publicado: (2024)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
por: Wang, Xinchen, et al.
Publicado: (2025)
por: Wang, Xinchen, et al.
Publicado: (2025)
How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data
por: Wang, Yejie, et al.
Publicado: (2024)
por: Wang, Yejie, et al.
Publicado: (2024)
Selective Prompt Anchoring for Code Generation
por: Tian, Yuan, et al.
Publicado: (2024)
por: Tian, Yuan, et al.
Publicado: (2024)
Synergizing LLMs and Knowledge Graphs: A Novel Approach to Software Repository-Related Question Answering
por: Abedu, Samuel, et al.
Publicado: (2024)
por: Abedu, Samuel, et al.
Publicado: (2024)
RovoDev Code Reviewer: A Large-Scale Online Evaluation of LLM-based Code Review Automation at Atlassian
por: Tantithamthavorn, Kla, et al.
Publicado: (2026)
por: Tantithamthavorn, Kla, et al.
Publicado: (2026)
Scaling Test-Time Compute for Agentic Coding
por: Kim, Joongwon, et al.
Publicado: (2026)
por: Kim, Joongwon, et al.
Publicado: (2026)
Rethinking Repetition Problems of LLMs in Code Generation
por: Dong, Yihong, et al.
Publicado: (2025)
por: Dong, Yihong, et al.
Publicado: (2025)
From I/O to Code with Discovery Agent
por: Dong, Yihong, et al.
Publicado: (2026)
por: Dong, Yihong, et al.
Publicado: (2026)
Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning
por: Ficek, Aleksander, et al.
Publicado: (2025)
por: Ficek, Aleksander, et al.
Publicado: (2025)
A Survey on Code Generation with LLM-based Agents
por: Dong, Yihong, et al.
Publicado: (2025)
por: Dong, Yihong, et al.
Publicado: (2025)
Towards Practical Defect-Focused Automated Code Review
por: Lu, Junyi, et al.
Publicado: (2025)
por: Lu, Junyi, et al.
Publicado: (2025)
code_transformed: The Influence of Large Language Models on Code
por: Xu, Yuliang, et al.
Publicado: (2025)
por: Xu, Yuliang, et al.
Publicado: (2025)
CODEMENV: Benchmarking Large Language Models on Code Migration
por: Cheng, Keyuan, et al.
Publicado: (2025)
por: Cheng, Keyuan, et al.
Publicado: (2025)
Vibe Checker: Aligning Code Evaluation with Human Preference
por: Zhong, Ming, et al.
Publicado: (2025)
por: Zhong, Ming, et al.
Publicado: (2025)
Improving Code Generation by Training with Natural Language Feedback
por: Chen, Angelica, et al.
Publicado: (2023)
por: Chen, Angelica, et al.
Publicado: (2023)
DeepCRCEval: Revisiting the Evaluation of Code Review Comment Generation
por: Lu, Junyi, et al.
Publicado: (2024)
por: Lu, Junyi, et al.
Publicado: (2024)
Evaluation of LLMs on Syntax-Aware Code Fill-in-the-Middle Tasks
por: Gong, Linyuan, et al.
Publicado: (2024)
por: Gong, Linyuan, et al.
Publicado: (2024)
Investigating the Efficacy of Large Language Models for Code Clone Detection
por: Khajezade, Mohamad, et al.
Publicado: (2024)
por: Khajezade, Mohamad, et al.
Publicado: (2024)
An Approach for Auto Generation of Labeling Functions for Software Engineering Chatbots
por: Alor, Ebube, et al.
Publicado: (2024)
por: Alor, Ebube, et al.
Publicado: (2024)
Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation
por: Yu, Zhuohao, et al.
Publicado: (2024)
por: Yu, Zhuohao, et al.
Publicado: (2024)
Confucius Code Agent: Scalable Agent Scaffolding for Real-World Codebases
por: Wong, Sherman, et al.
Publicado: (2025)
por: Wong, Sherman, et al.
Publicado: (2025)
The Larger the Better? Improved LLM Code-Generation via Budget Reallocation
por: Hassid, Michael, et al.
Publicado: (2024)
por: Hassid, Michael, et al.
Publicado: (2024)
Does Few-Shot Learning Help LLM Performance in Code Synthesis?
por: Xu, Derek, et al.
Publicado: (2024)
por: Xu, Derek, et al.
Publicado: (2024)
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
por: Hu, Ruida, et al.
Publicado: (2025)
por: Hu, Ruida, et al.
Publicado: (2025)
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
por: Liu, Jiacheng, et al.
Publicado: (2026)
por: Liu, Jiacheng, et al.
Publicado: (2026)
StRuCom: A Novel Dataset of Structured Code Comments in Russian
por: Dziuba, Maria, et al.
Publicado: (2025)
por: Dziuba, Maria, et al.
Publicado: (2025)
HGAdapter: Hypergraph-based Adapters in Language Models for Code Summarization and Clone Detection
por: Yang, Guang, et al.
Publicado: (2025)
por: Yang, Guang, et al.
Publicado: (2025)
DDPT: Diffusion-Driven Prompt Tuning for Large Language Model Code Generation
por: Li, Jinyang, et al.
Publicado: (2025)
por: Li, Jinyang, et al.
Publicado: (2025)
CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision
por: Lu, Yifei, et al.
Publicado: (2025)
por: Lu, Yifei, et al.
Publicado: (2025)
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
por: Yang, Dayu, et al.
Publicado: (2025)
por: Yang, Dayu, et al.
Publicado: (2025)
What can Large Language Models Capture about Code Functional Equivalence?
por: Maveli, Nickil, et al.
Publicado: (2024)
por: Maveli, Nickil, et al.
Publicado: (2024)
Ejemplares similares
-
EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback
por: Hajizadeh, Samira, et al.
Publicado: (2026) -
The Art of Repair: Optimizing Iterative Program Repair with Instruction-Tuned Models
por: Ruiz, Fernando Vallecillos, et al.
Publicado: (2025) -
Is Programming by Example solved by LLMs?
por: Li, Wen-Ding, et al.
Publicado: (2024) -
Automated Unity Game Template Generation from GDDs via NLP and Multi-Modal LLMs
por: Hassan, Amna
Publicado: (2025) -
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
por: Yang, Dayu, et al.
Publicado: (2025)