RepoForge: Training a SOTA Fast-thinking SWE Agent with an End-to-End Data Curation Pipeline Synergizing SFT and RL at Scale
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zhilong, Zhao, Chengzong, Chen, Boyuan, Lin, Dayi, Chen, Yihao, Leung, Arthur, Rajbahadur, Gopi Krishnan, Oliva, Gustavo A., Zhang, Haoxiang, Bhatia, Aaditya, Yong, Chong Chun, Hassan, Ahmed E. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SPICE: An Automated SWE-Bench Labeling Pipeline for Issue Clarity, Test Coverage, and Effort Estimation
von: Oliva, Gustavo A., et al.
Veröffentlicht: (2025)
von: Oliva, Gustavo A., et al.
Veröffentlicht: (2025)
Data Quality Antipatterns for Software Analytics
von: Bhatia, Aaditya, et al.
Veröffentlicht: (2024)
von: Bhatia, Aaditya, et al.
Veröffentlicht: (2024)
Rethinking Software Engineering in the Foundation Model Era: A Curated Catalogue of Challenges in the Development of Trustworthy FMware
von: Hassan, Ahmed E., et al.
Veröffentlicht: (2024)
von: Hassan, Ahmed E., et al.
Veröffentlicht: (2024)
SimClone: Detecting Tabular Data Clones using Value Similarity
von: Yang, Xu, et al.
Veröffentlicht: (2024)
von: Yang, Xu, et al.
Veröffentlicht: (2024)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
von: Ebrahimi, Amir M., et al.
Veröffentlicht: (2026)
von: Ebrahimi, Amir M., et al.
Veröffentlicht: (2026)
From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap
von: Rajbahadur, Gopi Krishnan, et al.
Veröffentlicht: (2024)
von: Rajbahadur, Gopi Krishnan, et al.
Veröffentlicht: (2024)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
SWE-Effi: Re-Evaluating Software AI Agent System Effectiveness Under Resource Constraints
von: Fan, Zhiyu, et al.
Veröffentlicht: (2025)
von: Fan, Zhiyu, et al.
Veröffentlicht: (2025)
RepoGenesis: Benchmarking End-to-End Microservice Generation from Readme to Repository
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2026)
PromptExp: Multi-granularity Prompt Explanation of Large Language Models
von: Dong, Ximing, et al.
Veröffentlicht: (2024)
von: Dong, Ximing, et al.
Veröffentlicht: (2024)
Implementing AI Bill of Materials (AI BOM) with SPDX 3.0: A Comprehensive Guide to Creating AI and Dataset Bill of Materials
von: Bennet, Karen, et al.
Veröffentlicht: (2025)
von: Bennet, Karen, et al.
Veröffentlicht: (2025)
RIG: Synergizing Reasoning and Imagination in End-to-End Generalist Policy
von: Zhao, Zhonghan, et al.
Veröffentlicht: (2025)
von: Zhao, Zhonghan, et al.
Veröffentlicht: (2025)
SWE-Spot: Building Small Repo-Experts with Repository-Centric Learning
von: Peng, Jinjun, et al.
Veröffentlicht: (2026)
von: Peng, Jinjun, et al.
Veröffentlicht: (2026)
BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?
von: Chen, Guoxin, et al.
Veröffentlicht: (2026)
von: Chen, Guoxin, et al.
Veröffentlicht: (2026)
ChronoForge-RL: Chronological Forging through Reinforcement Learning for Enhanced Video Understanding
von: Chen, Kehua
Veröffentlicht: (2025)
von: Chen, Kehua
Veröffentlicht: (2025)
From Hugging Face to GitHub: Tracing License Drift in the Open-Source AI Ecosystem
von: Jewitt, James, et al.
Veröffentlicht: (2025)
von: Jewitt, James, et al.
Veröffentlicht: (2025)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
von: Jewitt, James, et al.
Veröffentlicht: (2026)
von: Jewitt, James, et al.
Veröffentlicht: (2026)
From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving
von: Ang, Sining, et al.
Veröffentlicht: (2026)
von: Ang, Sining, et al.
Veröffentlicht: (2026)
Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
von: Vasilevski, Kirill, et al.
Veröffentlicht: (2025)
von: Vasilevski, Kirill, et al.
Veröffentlicht: (2025)
QiMeng-Attention: SOTA Attention Operator is generated by SOTA Attention Algorithm
von: Zhou, Qirui, et al.
Veröffentlicht: (2025)
von: Zhou, Qirui, et al.
Veröffentlicht: (2025)
EgoPush: Learning End-to-End Egocentric Multi-Object Rearrangement for Mobile Robots
von: An, Boyuan, et al.
Veröffentlicht: (2026)
von: An, Boyuan, et al.
Veröffentlicht: (2026)
Model Context Protocol (MCP) Tool Descriptions Are Smelly! Towards Improving AI Agent Efficiency with Augmented MCP Tool Descriptions
von: Hasan, Mohammed Mehedi, et al.
Veröffentlicht: (2026)
von: Hasan, Mohammed Mehedi, et al.
Veröffentlicht: (2026)
End-to-End Reinforcement Learning of Curative Curtailment with Partial Measurement Availability
von: Wolf, Hinrikus, et al.
Veröffentlicht: (2024)
von: Wolf, Hinrikus, et al.
Veröffentlicht: (2024)
EndToEndML: An Open-Source End-to-End Pipeline for Machine Learning Applications
von: Pillai, Nisha, et al.
Veröffentlicht: (2024)
von: Pillai, Nisha, et al.
Veröffentlicht: (2024)
An End-to-End Real-World Camera Imaging Pipeline
von: Xu, Kepeng, et al.
Veröffentlicht: (2024)
von: Xu, Kepeng, et al.
Veröffentlicht: (2024)
Achieving the Safety and Security of the End-to-End AV Pipeline
von: Curran, Noah T., et al.
Veröffentlicht: (2024)
von: Curran, Noah T., et al.
Veröffentlicht: (2024)
AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
From SFT to RL: Demystifying the Post-Training Pipeline for LLM-based Vulnerability Detection
von: Li, Youpeng, et al.
Veröffentlicht: (2026)
von: Li, Youpeng, et al.
Veröffentlicht: (2026)
End-to-End Data Engineering Pipeline for E-Commerce Analytics
von: Kadavala, Priyanka Raju
Veröffentlicht: (2025)
von: Kadavala, Priyanka Raju
Veröffentlicht: (2025)
An End-to-End Depth-Based Pipeline for Selfie Image Rectification
von: Alhawwary, Ahmed, et al.
Veröffentlicht: (2024)
von: Alhawwary, Ahmed, et al.
Veröffentlicht: (2024)
A Framework for Cryptographic Verifiability of End-to-End AI Pipelines
von: Balan, Kar, et al.
Veröffentlicht: (2025)
von: Balan, Kar, et al.
Veröffentlicht: (2025)
Beyond SFT-to-RL: Pre-alignment via Black-Box On-Policy Distillation for Multimodal RL
von: Wang, Sudong, et al.
Veröffentlicht: (2026)
von: Wang, Sudong, et al.
Veröffentlicht: (2026)
Model Context Protocol (MCP) at First Glance: Studying the Security and Maintainability of MCP Servers
von: Hasan, Mohammed Mehedi, et al.
Veröffentlicht: (2025)
von: Hasan, Mohammed Mehedi, et al.
Veröffentlicht: (2025)
An Empirical Study of Testing Practices in Open Source AI Agent Frameworks and Agentic Applications
von: Hasan, Mohammed Mehedi, et al.
Veröffentlicht: (2025)
von: Hasan, Mohammed Mehedi, et al.
Veröffentlicht: (2025)
Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning
von: Chen, Liang, et al.
Veröffentlicht: (2025)
von: Chen, Liang, et al.
Veröffentlicht: (2025)
Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
StraightLine: An End-to-End Resource-Aware Scheduler for Machine Learning Application Requests
von: Ching, Cheng-Wei, et al.
Veröffentlicht: (2024)
von: Ching, Cheng-Wei, et al.
Veröffentlicht: (2024)
Understanding the Performance Behaviors of End-to-End Protein Design Pipelines on GPUs
von: Hwang, Jinwoo, et al.
Veröffentlicht: (2026)
von: Hwang, Jinwoo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SPICE: An Automated SWE-Bench Labeling Pipeline for Issue Clarity, Test Coverage, and Effort Estimation
von: Oliva, Gustavo A., et al.
Veröffentlicht: (2025) -
Data Quality Antipatterns for Software Analytics
von: Bhatia, Aaditya, et al.
Veröffentlicht: (2024) -
Rethinking Software Engineering in the Foundation Model Era: A Curated Catalogue of Challenges in the Development of Trustworthy FMware
von: Hassan, Ahmed E., et al.
Veröffentlicht: (2024) -
SimClone: Detecting Tabular Data Clones using Value Similarity
von: Yang, Xu, et al.
Veröffentlicht: (2024) -
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
von: Ebrahimi, Amir M., et al.
Veröffentlicht: (2026)