EditFlow: Benchmarking and Optimizing Code Edit Recommendation Systems via Reconstruction of Developer Flows
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Chenyan, Lin, Yun, Chang, Jiaxin, Liu, Jiawei, Qi, Binhang, Jiang, Bo, Huang, Zhiyong, Dong, Jin Song |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Project-wise Subsequent Code Edits via Interleaving Neural-based Induction and Tool-based Deduction
by: Liu, Chenyan, et al.
Published: (2026)
by: Liu, Chenyan, et al.
Published: (2026)
CoEdPilot: Recommending Code Edits with Learned Prior Edit Relevance, Project-wise Awareness, and Interactive Nature
by: Liu, Chenyan, et al.
Published: (2024)
by: Liu, Chenyan, et al.
Published: (2024)
EditFlow-artifacts
by: Liu, Chenyan
Published: (2026)
by: Liu, Chenyan
Published: (2026)
EfficientEdit: Accelerating Code Editing via Edit-Oriented Speculative Decoding
by: Wang, Peiding, et al.
Published: (2025)
by: Wang, Peiding, et al.
Published: (2025)
Generalizing Test Cases for Comprehensive Test Scenario Coverage
by: Qi, Binhang, et al.
Published: (2026)
by: Qi, Binhang, et al.
Published: (2026)
Generating Project-Specific Test Cases with Requirement Validation Intention
by: Qi, Binhang, et al.
Published: (2025)
by: Qi, Binhang, et al.
Published: (2025)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
by: Ebrahimi, Amir M., et al.
Published: (2026)
by: Ebrahimi, Amir M., et al.
Published: (2026)
Next Edit Prediction: Learning to Predict Code Edits from Context and Interaction History
by: Lu, Ruofan, et al.
Published: (2025)
by: Lu, Ruofan, et al.
Published: (2025)
Robust Learning of Diverse Code Edits
by: Aggarwal, Tushar, et al.
Published: (2025)
by: Aggarwal, Tushar, et al.
Published: (2025)
Let the Code LLM Edit Itself When You Edit the Code
by: He, Zhenyu, et al.
Published: (2024)
by: He, Zhenyu, et al.
Published: (2024)
QiMeng-PRepair: Precise Code Repair via Edit-Aware Reward Optimization
by: Ke, Changxin, et al.
Published: (2026)
by: Ke, Changxin, et al.
Published: (2026)
An Empirical Study of Java Code Improvements Based on Stack Overflow Answer Edits
by: Wiratsin, In-on, et al.
Published: (2025)
by: Wiratsin, In-on, et al.
Published: (2025)
EDIT-Bench: Evaluating LLM Abilities to Perform Real-World Instructed Code Edits
by: Chi, Wayne, et al.
Published: (2025)
by: Chi, Wayne, et al.
Published: (2025)
PAFT: Preservation Aware Fine-Tuning for Minimal-Edit Program Repair
by: Yang, Boyang, et al.
Published: (2026)
by: Yang, Boyang, et al.
Published: (2026)
Learning Code-Edit Embedding to Model Student Debugging Behavior
by: Heickal, Hasnain, et al.
Published: (2025)
by: Heickal, Hasnain, et al.
Published: (2025)
Learning Performance-Improving Code Edits
by: Shypula, Alexander, et al.
Published: (2023)
by: Shypula, Alexander, et al.
Published: (2023)
Revisiting Code Similarity Evaluation with Abstract Syntax Tree Edit Distance
by: Song, Yewei, et al.
Published: (2024)
by: Song, Yewei, et al.
Published: (2024)
EditLord: Learning Code Transformation Rules for Code Editing
by: Li, Weichen, et al.
Published: (2025)
by: Li, Weichen, et al.
Published: (2025)
Needle in the Repo: A Benchmark for Maintainability in AI-Generated Repository Edits
by: Zhu, Haichao, et al.
Published: (2026)
by: Zhu, Haichao, et al.
Published: (2026)
Suggesting Code Edits in Interactive Machine Learning Notebooks Using Large Language Models
by: Jin, Bihui, et al.
Published: (2025)
by: Jin, Bihui, et al.
Published: (2025)
NaturalEdit: Code Modification through Direct Interaction with Adaptive Natural Language Representation
by: Tang, Ningzhi, et al.
Published: (2025)
by: Tang, Ningzhi, et al.
Published: (2025)
CoderEval: A Benchmark of Pragmatic Code Generation with Generative Pre-trained Models
by: Yu, Hao, et al.
Published: (2023)
by: Yu, Hao, et al.
Published: (2023)
SPELL: Synthesis of Programmatic Edits using LLMs
by: Ramos, Daniel, et al.
Published: (2026)
by: Ramos, Daniel, et al.
Published: (2026)
CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation
by: Wang, Sizhe, et al.
Published: (2025)
by: Wang, Sizhe, et al.
Published: (2025)
RealBench: A Repo-Level Code Generation Benchmark Aligned with Real-World Software Development Practices
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
Automatically Recommend Code Updates: Are We There Yet?
by: Liu, Yue, et al.
Published: (2022)
by: Liu, Yue, et al.
Published: (2022)
COFFE: A Code Efficiency Benchmark for Code Generation
by: Peng, Yun, et al.
Published: (2025)
by: Peng, Yun, et al.
Published: (2025)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
by: Guo, Chengquan, et al.
Published: (2024)
by: Guo, Chengquan, et al.
Published: (2024)
Flow2Code: Evaluating Large Language Models for Flowchart-based Code Generation Capability
by: He, Mengliang, et al.
Published: (2025)
by: He, Mengliang, et al.
Published: (2025)
DFEPT: Data Flow Embedding for Enhancing Pre-Trained Model Based Vulnerability Detection
by: Jiang, Zhonghao, et al.
Published: (2024)
by: Jiang, Zhonghao, et al.
Published: (2024)
Development and Benchmarking of Multilingual Code Clone Detector
by: Zhu, Wenqing, et al.
Published: (2024)
by: Zhu, Wenqing, et al.
Published: (2024)
PEACE: Towards Efficient Project-Level Efficiency Optimization via Hybrid Code Editing
by: Ren, Xiaoxue, et al.
Published: (2025)
by: Ren, Xiaoxue, et al.
Published: (2025)
CodeScore: Evaluating Code Generation by Learning Code Execution
by: Dong, Yihong, et al.
Published: (2023)
by: Dong, Yihong, et al.
Published: (2023)
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
by: Cassano, Federico, et al.
Published: (2023)
by: Cassano, Federico, et al.
Published: (2023)
CodeFlow: Program Behavior Prediction with Dynamic Dependencies Learning
by: Le, Cuong Chi, et al.
Published: (2024)
by: Le, Cuong Chi, et al.
Published: (2024)
Primary Breadth-First Development (PBFD): An Approach to Full Stack Software Development
by: Liu, Dong
Published: (2025)
by: Liu, Dong
Published: (2025)
Learner-Tailored Program Repair: A Solution Generator with Iterative Edit-Driven Retrieval Enhancement
by: Dai, Zhenlong, et al.
Published: (2026)
by: Dai, Zhenlong, et al.
Published: (2026)
NoCode-bench: A Benchmark for Evaluating Natural Language-Driven Feature Addition
by: Deng, Le, et al.
Published: (2025)
by: Deng, Le, et al.
Published: (2025)
Self-collaboration Code Generation via ChatGPT
by: Dong, Yihong, et al.
Published: (2023)
by: Dong, Yihong, et al.
Published: (2023)
Evaluating LLM-Generated Code: A Benchmark and Developer Study
by: Szych, Joanna, et al.
Published: (2026)
by: Szych, Joanna, et al.
Published: (2026)
Similar Items
-
Learning Project-wise Subsequent Code Edits via Interleaving Neural-based Induction and Tool-based Deduction
by: Liu, Chenyan, et al.
Published: (2026) -
CoEdPilot: Recommending Code Edits with Learned Prior Edit Relevance, Project-wise Awareness, and Interactive Nature
by: Liu, Chenyan, et al.
Published: (2024) -
EditFlow-artifacts
by: Liu, Chenyan
Published: (2026) -
EfficientEdit: Accelerating Code Editing via Edit-Oriented Speculative Decoding
by: Wang, Peiding, et al.
Published: (2025) -
Generalizing Test Cases for Comprehensive Test Scenario Coverage
by: Qi, Binhang, et al.
Published: (2026)