SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Jeffrey Jian, Hashemi, Milad, Yazdanbakhsh, Amir, Swersky, Kevin, Press, Ofir, Li, Enhui, Reddi, Vijay Janapa, Ranganathan, Parthasarathy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Performance-Improving Code Edits
by: Shypula, Alexander, et al.
Published: (2023)
by: Shypula, Alexander, et al.
Published: (2023)
PerfBench: Can Agents Resolve Real-World Performance Bugs?
by: Garg, Spandan, et al.
Published: (2025)
by: Garg, Spandan, et al.
Published: (2025)
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
by: Jimenez, Carlos E., et al.
Published: (2023)
by: Jimenez, Carlos E., et al.
Published: (2023)
SWE-Perf: Can Language Models Optimize Code Performance on Real-World Repositories?
by: He, Xinyi, et al.
Published: (2025)
by: He, Xinyi, et al.
Published: (2025)
GreenMalloc: Allocator Optimisation for Industrial Workloads
by: Dakhama, Aidan, et al.
Published: (2025)
by: Dakhama, Aidan, et al.
Published: (2025)
Real-time Event Joining in Practice With Kafka and Flink
by: Saket, Srijan, et al.
Published: (2024)
by: Saket, Srijan, et al.
Published: (2024)
Do AI Models Dream of Faster Code? An Empirical Study on LLM-Proposed Performance Improvements in Real-World Software
by: Yi, Lirong, et al.
Published: (2025)
by: Yi, Lirong, et al.
Published: (2025)
SWE-Refactor: A Repository-Level Benchmark for Real-World LLM-Based Code Refactoring
by: Xu, Yisen, et al.
Published: (2026)
by: Xu, Yisen, et al.
Published: (2026)
SWE-Tester: Training Open-Source LLMs for Issue Reproduction in Real-World Repositories
by: Soni, Aditya Bharat, et al.
Published: (2026)
by: Soni, Aditya Bharat, et al.
Published: (2026)
A Magnified View into Heterogeneous-ISA Thread Migration Performance without State Transformation
by: Mavrogeorgis, Nikolaos, et al.
Published: (2025)
by: Mavrogeorgis, Nikolaos, et al.
Published: (2025)
LLMSYS-HPOBench: Hyperparameter Optimization Benchmark Suite for Real-World LLM Systems
by: Wu, Siyu, et al.
Published: (2026)
by: Wu, Siyu, et al.
Published: (2026)
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations
by: Rani, Pooja, et al.
Published: (2025)
by: Rani, Pooja, et al.
Published: (2025)
CodeRosetta: Pushing the Boundaries of Unsupervised Code Translation for Parallel Programming
by: TehraniJamsaz, Ali, et al.
Published: (2024)
by: TehraniJamsaz, Ali, et al.
Published: (2024)
Estimating the Energy Footprint of Software Systems: a Primer
by: Castor, Fernando
Published: (2024)
by: Castor, Fernando
Published: (2024)
FaaSter Troubleshooting -- Evaluating Distributed Tracing Approaches for Serverless Applications
by: Borges, Maria C., et al.
Published: (2021)
by: Borges, Maria C., et al.
Published: (2021)
gigiProfiler: Diagnosing Performance Issues by Uncovering Application Resource Bottlenecks
by: Hu, Yigong, et al.
Published: (2025)
by: Hu, Yigong, et al.
Published: (2025)
An Empirical Study on Method-Level Performance Evolution in Open-Source Java Projects
by: Shahedi, Kaveh, et al.
Published: (2025)
by: Shahedi, Kaveh, et al.
Published: (2025)
On the Role of Search Budgets in Model-Based Software Refactoring Optimization
by: Diaz-Pace, J. Andres, et al.
Published: (2023)
by: Diaz-Pace, J. Andres, et al.
Published: (2023)
What Is the Cost of Energy Monitoring? An Empirical Study on the Overhead of RAPL-Based Tools
by: Diamond, Jeremy, et al.
Published: (2026)
by: Diamond, Jeremy, et al.
Published: (2026)
Performance of Genetic Algorithms in the Context of Software Model Refactoring
by: Cortellessa, Vittorio, et al.
Published: (2023)
by: Cortellessa, Vittorio, et al.
Published: (2023)
Towards Assessing Spread in Sets of Software Architecture Designs
by: Cortellessa, Vittorio, et al.
Published: (2024)
by: Cortellessa, Vittorio, et al.
Published: (2024)
Scalable Software as a Service Architecture
by: Dedase, Ardy
Published: (2024)
by: Dedase, Ardy
Published: (2024)
Optimas: An Intelligent Analytics-Informed Generative AI Framework for Performance Optimization
by: Zaeed, Mohammad, et al.
Published: (2026)
by: Zaeed, Mohammad, et al.
Published: (2026)
RAO-SS: A Prototype of Run-time Auto-tuning Facility for Sparse Direct Solvers
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
A Test for FLOPs as a Discriminant for Linear Algebra Algorithms
by: Sankaran, Aravind, et al.
Published: (2022)
by: Sankaran, Aravind, et al.
Published: (2022)
MLKAPS: Machine Learning and Adaptive Sampling for HPC Kernel Auto-tuning
by: Jam, Mathys, et al.
Published: (2025)
by: Jam, Mathys, et al.
Published: (2025)
Towards a Higher Roofline for Matrix-Vector Multiplication in Matrix-Free HOSFEM
by: Cao, Zijian, et al.
Published: (2025)
by: Cao, Zijian, et al.
Published: (2025)
Formal Analysis of Metastable Failures in Software Systems
by: Alvaro, Peter, et al.
Published: (2025)
by: Alvaro, Peter, et al.
Published: (2025)
Employing Software Diversity in Cloud Microservices to Engineer Reliable and Performant Systems
by: Akhtarian, Nazanin, et al.
Published: (2024)
by: Akhtarian, Nazanin, et al.
Published: (2024)
Faster Base64 Encoding and Decoding Using AVX2 Instructions
by: Muła, Wojciech, et al.
Published: (2017)
by: Muła, Wojciech, et al.
Published: (2017)
Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning
by: Su, Qisheng, et al.
Published: (2026)
by: Su, Qisheng, et al.
Published: (2026)
Energy Patterns for Web: An Exploratory Study
by: Rani, Pooja, et al.
Published: (2024)
by: Rani, Pooja, et al.
Published: (2024)
SysLLMatic: Large Language Models are Software System Optimizers
by: Peng, Huiyun, et al.
Published: (2025)
by: Peng, Huiyun, et al.
Published: (2025)
Efficiently Ranking Software Variants with Minimal Benchmarks
by: Matricon, Théo, et al.
Published: (2025)
by: Matricon, Théo, et al.
Published: (2025)
Energy-Efficient Software Development: A Multi-dimensional Empirical Analysis of Stack Overflow
by: Jin, Bihui, et al.
Published: (2024)
by: Jin, Bihui, et al.
Published: (2024)
Tracing Optimization for Performance Modeling and Regression Detection
by: Shahedi, Kaveh, et al.
Published: (2024)
by: Shahedi, Kaveh, et al.
Published: (2024)
A Model-driven Approach for Continuous Performance Engineering in Microservice-based Systems
by: Cortellessa, Vittorio, et al.
Published: (2023)
by: Cortellessa, Vittorio, et al.
Published: (2023)
FlipFlop: A Static Analysis-based Energy Optimization Framework for GPU Kernels
by: Rajput, Saurabhsingh, et al.
Published: (2026)
by: Rajput, Saurabhsingh, et al.
Published: (2026)
An Empirical Study on How Architectural Topology Affects Microservice Performance and Energy Usage
by: Ristova, Irena, et al.
Published: (2026)
by: Ristova, Irena, et al.
Published: (2026)
SWE-Universe: Scale Real-World Verifiable Environments to Millions
by: Chen, Mouxiang, et al.
Published: (2026)
by: Chen, Mouxiang, et al.
Published: (2026)
Similar Items
-
Learning Performance-Improving Code Edits
by: Shypula, Alexander, et al.
Published: (2023) -
PerfBench: Can Agents Resolve Real-World Performance Bugs?
by: Garg, Spandan, et al.
Published: (2025) -
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
by: Jimenez, Carlos E., et al.
Published: (2023) -
SWE-Perf: Can Language Models Optimize Code Performance on Real-World Repositories?
by: He, Xinyi, et al.
Published: (2025) -
GreenMalloc: Allocator Optimisation for Industrial Workloads
by: Dakhama, Aidan, et al.
Published: (2025)