SWE-Replay: Efficient Test-Time Scaling for Software Engineering Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ding, Yifeng, Zhang, Lingming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2025)
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2025)
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
SWE-MiniSandbox: Container-Free Reinforcement Learning for Building Software Engineering Agents
von: Yuan, Danlong, et al.
Veröffentlicht: (2026)
von: Yuan, Danlong, et al.
Veröffentlicht: (2026)
Agentless: Demystifying LLM-based Software Engineering Agents
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024)
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024)
Satori-SWE: Evolutionary Test-Time Scaling for Sample-Efficient Software Engineering
von: Zeng, Guangtao, et al.
Veröffentlicht: (2025)
von: Zeng, Guangtao, et al.
Veröffentlicht: (2025)
Code Generation by Differential Test Time Scaling
von: He, Yifeng, et al.
Veröffentlicht: (2026)
von: He, Yifeng, et al.
Veröffentlicht: (2026)
GSO: Challenging Software Optimization Tasks for Evaluating SWE-Agents
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
SWE-Protégé: Learning to Selectively Collaborate With an Expert Unlocks Small Language Models as Software Engineering Agents
von: Kon, Patrick Tser Jern, et al.
Veröffentlicht: (2026)
von: Kon, Patrick Tser Jern, et al.
Veröffentlicht: (2026)
SWE-smith: Scaling Data for Software Engineering Agents
von: Yang, John, et al.
Veröffentlicht: (2025)
von: Yang, John, et al.
Veröffentlicht: (2025)
XFT: Unlocking the Power of Code Instruction Tuning by Simply Merging Upcycled Mixture-of-Experts
von: Ding, Yifeng, et al.
Veröffentlicht: (2024)
von: Ding, Yifeng, et al.
Veröffentlicht: (2024)
SWE-Bench-CL: Continual Learning for Coding Agents
von: Joshi, Thomas, et al.
Veröffentlicht: (2025)
von: Joshi, Thomas, et al.
Veröffentlicht: (2025)
SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
von: Yang, John, et al.
Veröffentlicht: (2024)
von: Yang, John, et al.
Veröffentlicht: (2024)
TOM-SWE: User Mental Modeling For Software Engineering Agents
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025)
SWE-Bench++: A Framework for the Scalable Generation of Software Engineering Benchmarks from Open-Source Repositories
von: Wang, Lilin, et al.
Veröffentlicht: (2025)
von: Wang, Lilin, et al.
Veröffentlicht: (2025)
Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents
von: Huang, Jiawei, et al.
Veröffentlicht: (2026)
von: Huang, Jiawei, et al.
Veröffentlicht: (2026)
SWE-Next: Scalable Real-World Software Engineering Tasks for Agents
von: Liang, Jiarong, et al.
Veröffentlicht: (2026)
von: Liang, Jiarong, et al.
Veröffentlicht: (2026)
EGSS: Entropy-guided Stepwise Scaling for Reliable Software Engineering
von: Mao, Chenhui, et al.
Veröffentlicht: (2026)
von: Mao, Chenhui, et al.
Veröffentlicht: (2026)
Revisiting the Plastic Surgery Hypothesis via Large Language Models
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2023)
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2023)
Challenges and Paths Towards AI for Software Engineering
von: Gu, Alex, et al.
Veröffentlicht: (2025)
von: Gu, Alex, et al.
Veröffentlicht: (2025)
SetupBench: Assessing Software Engineering Agents' Ability to Bootstrap Development Environments
von: Arora, Avi, et al.
Veröffentlicht: (2025)
von: Arora, Avi, et al.
Veröffentlicht: (2025)
Large Language Model-Based Agents for Software Engineering: A Survey
von: Liu, Junwei, et al.
Veröffentlicht: (2024)
von: Liu, Junwei, et al.
Veröffentlicht: (2024)
SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
von: Han, Tingxu, et al.
Veröffentlicht: (2026)
von: Han, Tingxu, et al.
Veröffentlicht: (2026)
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
von: Ran, Dezhi, et al.
Veröffentlicht: (2024)
von: Ran, Dezhi, et al.
Veröffentlicht: (2024)
The Rise of AI Teammates in Software Engineering (SE) 3.0: How Autonomous Coding Agents Are Reshaping Software Engineering
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
von: Trae Research Team, et al.
Veröffentlicht: (2025)
von: Trae Research Team, et al.
Veröffentlicht: (2025)
On the Replicability and Reproducibility of Deep Learning in Software Engineering
von: Liu, Chao, et al.
Veröffentlicht: (2020)
von: Liu, Chao, et al.
Veröffentlicht: (2020)
Thinking Longer, Not Larger: Enhancing Software Engineering Agents via Scaling Test-Time Compute
von: Ma, Yingwei, et al.
Veröffentlicht: (2025)
von: Ma, Yingwei, et al.
Veröffentlicht: (2025)
On Problems of Implicit Context Compression for Software Engineering Agents
von: Gelvan, Kirill, et al.
Veröffentlicht: (2026)
von: Gelvan, Kirill, et al.
Veröffentlicht: (2026)
Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
LLMs in Coding and their Impact on the Commercial Software Engineering Landscape
von: Belozerov, Vladislav, et al.
Veröffentlicht: (2025)
von: Belozerov, Vladislav, et al.
Veröffentlicht: (2025)
SWE-Spot: Building Small Repo-Experts with Repository-Centric Learning
von: Peng, Jinjun, et al.
Veröffentlicht: (2026)
von: Peng, Jinjun, et al.
Veröffentlicht: (2026)
SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
Toward Explaining Large Language Models in Software Engineering Tasks
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
The Impact of Software Testing with Quantum Optimization Meets Machine Learning
von: Bandarupalli, Gopichand
Veröffentlicht: (2025)
von: Bandarupalli, Gopichand
Veröffentlicht: (2025)
VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
von: Bai, Yifan, et al.
Veröffentlicht: (2026)
von: Bai, Yifan, et al.
Veröffentlicht: (2026)
Towards a Classification of Open-Source ML Models and Datasets for Software Engineering
von: González, Alexandra, et al.
Veröffentlicht: (2024)
von: González, Alexandra, et al.
Veröffentlicht: (2024)
SWE-Arena: An Interactive Platform for Evaluating Foundation Models in Software Engineering
von: Zhao, Zhimin
Veröffentlicht: (2025)
von: Zhao, Zhimin
Veröffentlicht: (2025)
Scaling Test-Time Compute for Agentic Coding
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
Towards Engineering Fair and Equitable Software Systems for Managing Low-Altitude Airspace Authorizations
von: Gohar, Usman, et al.
Veröffentlicht: (2024)
von: Gohar, Usman, et al.
Veröffentlicht: (2024)
SWE-Hub: A Unified Production System for Scalable, Executable Software Engineering Tasks
von: Zeng, Yucheng, et al.
Veröffentlicht: (2026)
von: Zeng, Yucheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2025) -
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025) -
SWE-MiniSandbox: Container-Free Reinforcement Learning for Building Software Engineering Agents
von: Yuan, Danlong, et al.
Veröffentlicht: (2026) -
Agentless: Demystifying LLM-based Software Engineering Agents
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024) -
Satori-SWE: Evolutionary Test-Time Scaling for Sample-Efficient Software Engineering
von: Zeng, Guangtao, et al.
Veröffentlicht: (2025)