KernelBench: Can LLMs Write Efficient GPU Kernels?
Fuente:
arXiv
Saved in:
| Main Authors: | Ouyang, Anne, Guo, Simon, Arora, Simran, Zhang, Alex L., Hu, William, Ré, Christopher, Mirhoseini, Azalia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlipFlop: A Static Analysis-based Energy Optimization Framework for GPU Kernels
by: Rajput, Saurabhsingh, et al.
Published: (2026)
by: Rajput, Saurabhsingh, et al.
Published: (2026)
GPU Kernel Scientist: An LLM-Driven Framework for Iterative Kernel Optimization
by: Andrews, Martin, et al.
Published: (2025)
by: Andrews, Martin, et al.
Published: (2025)
MLKAPS: Machine Learning and Adaptive Sampling for HPC Kernel Auto-tuning
by: Jam, Mathys, et al.
Published: (2025)
by: Jam, Mathys, et al.
Published: (2025)
Astra: A Multi-Agent System for GPU Kernel Performance Optimization
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
by: Nichols, Daniel, et al.
Published: (2025)
by: Nichols, Daniel, et al.
Published: (2025)
MultiKernelBench: A Multi-Platform Benchmark for Kernel Generation
by: Wen, Zhongzhen, et al.
Published: (2025)
by: Wen, Zhongzhen, et al.
Published: (2025)
Kevin: Multi-Turn RL for Generating CUDA Kernels
by: Baronio, Carlo, et al.
Published: (2025)
by: Baronio, Carlo, et al.
Published: (2025)
CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe
by: Saba, Tara, et al.
Published: (2026)
by: Saba, Tara, et al.
Published: (2026)
Stencil-Lifting: Hierarchical Recursive Lifting System for Extracting Summary of Stencil Kernel in Legacy Codes
by: Li, Mingyi, et al.
Published: (2025)
by: Li, Mingyi, et al.
Published: (2025)
PerfBench: Can Agents Resolve Real-World Performance Bugs?
by: Garg, Spandan, et al.
Published: (2025)
by: Garg, Spandan, et al.
Published: (2025)
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations
by: Rani, Pooja, et al.
Published: (2025)
by: Rani, Pooja, et al.
Published: (2025)
Efficiently Ranking Software Variants with Minimal Benchmarks
by: Matricon, Théo, et al.
Published: (2025)
by: Matricon, Théo, et al.
Published: (2025)
Energy-Efficient Software Development: A Multi-dimensional Empirical Analysis of Stack Overflow
by: Jin, Bihui, et al.
Published: (2024)
by: Jin, Bihui, et al.
Published: (2024)
gigiProfiler: Diagnosing Performance Issues by Uncovering Application Resource Bottlenecks
by: Hu, Yigong, et al.
Published: (2025)
by: Hu, Yigong, et al.
Published: (2025)
AscendCraft: Automatic Ascend NPU Kernel Generation via DSL-Guided Transcompilation
by: Wen, Zhongzhen, et al.
Published: (2026)
by: Wen, Zhongzhen, et al.
Published: (2026)
Prompting for Performance: Exploring LLMs for Configuring Software
by: Spieker, Helge, et al.
Published: (2025)
by: Spieker, Helge, et al.
Published: (2025)
SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?
by: Ma, Jeffrey Jian, et al.
Published: (2025)
by: Ma, Jeffrey Jian, et al.
Published: (2025)
Estimating the Energy Footprint of Software Systems: a Primer
by: Castor, Fernando
Published: (2024)
by: Castor, Fernando
Published: (2024)
FaaSter Troubleshooting -- Evaluating Distributed Tracing Approaches for Serverless Applications
by: Borges, Maria C., et al.
Published: (2021)
by: Borges, Maria C., et al.
Published: (2021)
A Magnified View into Heterogeneous-ISA Thread Migration Performance without State Transformation
by: Mavrogeorgis, Nikolaos, et al.
Published: (2025)
by: Mavrogeorgis, Nikolaos, et al.
Published: (2025)
An Empirical Study on Method-Level Performance Evolution in Open-Source Java Projects
by: Shahedi, Kaveh, et al.
Published: (2025)
by: Shahedi, Kaveh, et al.
Published: (2025)
On the Role of Search Budgets in Model-Based Software Refactoring Optimization
by: Diaz-Pace, J. Andres, et al.
Published: (2023)
by: Diaz-Pace, J. Andres, et al.
Published: (2023)
What Is the Cost of Energy Monitoring? An Empirical Study on the Overhead of RAPL-Based Tools
by: Diamond, Jeremy, et al.
Published: (2026)
by: Diamond, Jeremy, et al.
Published: (2026)
Performance of Genetic Algorithms in the Context of Software Model Refactoring
by: Cortellessa, Vittorio, et al.
Published: (2023)
by: Cortellessa, Vittorio, et al.
Published: (2023)
Towards Assessing Spread in Sets of Software Architecture Designs
by: Cortellessa, Vittorio, et al.
Published: (2024)
by: Cortellessa, Vittorio, et al.
Published: (2024)
Scalable Software as a Service Architecture
by: Dedase, Ardy
Published: (2024)
by: Dedase, Ardy
Published: (2024)
Optimas: An Intelligent Analytics-Informed Generative AI Framework for Performance Optimization
by: Zaeed, Mohammad, et al.
Published: (2026)
by: Zaeed, Mohammad, et al.
Published: (2026)
Formal Analysis of Metastable Failures in Software Systems
by: Alvaro, Peter, et al.
Published: (2025)
by: Alvaro, Peter, et al.
Published: (2025)
Employing Software Diversity in Cloud Microservices to Engineer Reliable and Performant Systems
by: Akhtarian, Nazanin, et al.
Published: (2024)
by: Akhtarian, Nazanin, et al.
Published: (2024)
Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning
by: Su, Qisheng, et al.
Published: (2026)
by: Su, Qisheng, et al.
Published: (2026)
Energy Patterns for Web: An Exploratory Study
by: Rani, Pooja, et al.
Published: (2024)
by: Rani, Pooja, et al.
Published: (2024)
SysLLMatic: Large Language Models are Software System Optimizers
by: Peng, Huiyun, et al.
Published: (2025)
by: Peng, Huiyun, et al.
Published: (2025)
Tracing Optimization for Performance Modeling and Regression Detection
by: Shahedi, Kaveh, et al.
Published: (2024)
by: Shahedi, Kaveh, et al.
Published: (2024)
A Model-driven Approach for Continuous Performance Engineering in Microservice-based Systems
by: Cortellessa, Vittorio, et al.
Published: (2023)
by: Cortellessa, Vittorio, et al.
Published: (2023)
An Empirical Study on How Architectural Topology Affects Microservice Performance and Energy Usage
by: Ristova, Irena, et al.
Published: (2026)
by: Ristova, Irena, et al.
Published: (2026)
This Is Taking Too Long -- Investigating Time as a Proxy for Energy Consumption of LLMs
by: Krupp, Lars, et al.
Published: (2026)
by: Krupp, Lars, et al.
Published: (2026)
Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts
by: Huang, Zixuan, et al.
Published: (2026)
by: Huang, Zixuan, et al.
Published: (2026)
Input-Gen: Guided Generation of Stateful Inputs for Testing, Tuning, and Training
by: Ivanov, Ivan R., et al.
Published: (2024)
by: Ivanov, Ivan R., et al.
Published: (2024)
Real-time Event Joining in Practice With Kafka and Flink
by: Saket, Srijan, et al.
Published: (2024)
by: Saket, Srijan, et al.
Published: (2024)
AI-NativeBench: An Open-Source White-Box Agentic Benchmark Suite for AI-Native Systems
by: Wang, Zirui, et al.
Published: (2026)
by: Wang, Zirui, et al.
Published: (2026)
Similar Items
-
FlipFlop: A Static Analysis-based Energy Optimization Framework for GPU Kernels
by: Rajput, Saurabhsingh, et al.
Published: (2026) -
GPU Kernel Scientist: An LLM-Driven Framework for Iterative Kernel Optimization
by: Andrews, Martin, et al.
Published: (2025) -
MLKAPS: Machine Learning and Adaptive Sampling for HPC Kernel Auto-tuning
by: Jam, Mathys, et al.
Published: (2025) -
Astra: A Multi-Agent System for GPU Kernel Performance Optimization
by: Wei, Anjiang, et al.
Published: (2025) -
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
by: Nichols, Daniel, et al.
Published: (2025)