CppPerf: An Automated Pipeline and Dataset for Performance-Improving C++ Commits
Fuente:
arXiv
Saved in:
| Main Authors: | Ho, Tommy, Etemadi, Khashayar, Su, Zhendong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mokav: Execution-driven Differential Testing with LLMs
by: Etemadi, Khashayar, et al.
Published: (2024)
by: Etemadi, Khashayar, et al.
Published: (2024)
VeCoGen: Automating Generation of Formally Verified C Code with Large Language Models
by: Sevenhuijsen, Merlijn, et al.
Published: (2024)
by: Sevenhuijsen, Merlijn, et al.
Published: (2024)
PerfGen: Automated Performance Benchmark Generation for Big Data Analytics
by: Wang, Jiyuan, et al.
Published: (2024)
by: Wang, Jiyuan, et al.
Published: (2024)
CigaR: Cost-efficient Program Repair with LLMs
by: Hidvégi, Dávid, et al.
Published: (2024)
by: Hidvégi, Dávid, et al.
Published: (2024)
LLM-based Property-based Test Generation for Guardrailing Cyber-Physical Systems
by: Etemadi, Khashayar, et al.
Published: (2025)
by: Etemadi, Khashayar, et al.
Published: (2025)
Descriptor: C++ Self-Admitted Technical Debt Dataset (CppSATD)
by: Pham, Phuoc, et al.
Published: (2025)
by: Pham, Phuoc, et al.
Published: (2025)
FormalSpecCpp: A Dataset of C++ Formal Specifications created using LLMs
by: Chakraborty, Madhurima, et al.
Published: (2025)
by: Chakraborty, Madhurima, et al.
Published: (2025)
SWE-Perf: Can Language Models Optimize Code Performance on Real-World Repositories?
by: He, Xinyi, et al.
Published: (2025)
by: He, Xinyi, et al.
Published: (2025)
LLM4Perf: Large Language Models Are Effective Samplers for Multi-Objective Performance Modeling
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
PerfCoder: Large Language Models for Interpretable Code Performance Optimization
by: Yang, Jiuding, et al.
Published: (2025)
by: Yang, Jiuding, et al.
Published: (2025)
Finding Performance Issues in Database Systems by Exploiting Dormant Code Paths
by: Ba, Jinsheng, et al.
Published: (2026)
by: Ba, Jinsheng, et al.
Published: (2026)
SmartDoc: A Context-Aware Agentic Method Comment Generation Plugin
by: Etemadi, Vahid, et al.
Published: (2025)
by: Etemadi, Vahid, et al.
Published: (2025)
PerfBench: Can Agents Resolve Real-World Performance Bugs?
by: Garg, Spandan, et al.
Published: (2025)
by: Garg, Spandan, et al.
Published: (2025)
MLIR-Smith: A Novel Random Program Generator for Evaluating Compiler Pipelines
by: Ates, Berke, et al.
Published: (2026)
by: Ates, Berke, et al.
Published: (2026)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
by: Peng, Yun, et al.
Published: (2024)
by: Peng, Yun, et al.
Published: (2024)
Automated Generation of Commit Messages in Software Repositories
by: Palakodeti, Varun Kumar, et al.
Published: (2025)
by: Palakodeti, Varun Kumar, et al.
Published: (2025)
Mono2Sls: Automated Monolith-to-Serverless Migration via Multi-Stage Pipeline with Static Analysis
by: Chen, Xingyan, et al.
Published: (2026)
by: Chen, Xingyan, et al.
Published: (2026)
Practical, Automated Scenario-based Mobile App Testing
by: Yu, Shengcheng, et al.
Published: (2024)
by: Yu, Shengcheng, et al.
Published: (2024)
Improving the Quality of Commit Messages in Students' Projects
by: Ma, Iris, et al.
Published: (2023)
by: Ma, Iris, et al.
Published: (2023)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
by: Li, Junjie, et al.
Published: (2025)
by: Li, Junjie, et al.
Published: (2025)
API-guided Dataset Synthesis to Finetune Large Code Models
by: Li, Zongjie, et al.
Published: (2024)
by: Li, Zongjie, et al.
Published: (2024)
PerfCurator: Curating a large-scale dataset of performance bug-related commits from public repositories
by: Azad, Md Abul Kalam, et al.
Published: (2024)
by: Azad, Md Abul Kalam, et al.
Published: (2024)
PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization
by: Jing, Huihao, et al.
Published: (2026)
by: Jing, Huihao, et al.
Published: (2026)
Automated Commit Message Generation with Large Language Models: An Empirical Study and Beyond
by: Xue, Pengyu, et al.
Published: (2024)
by: Xue, Pengyu, et al.
Published: (2024)
CrossCommitVuln-Bench: A Dataset of Multi-Commit Python Vulnerabilities Invisible to Per-Commit Static Analysis
by: Majumdar, Arunabh
Published: (2026)
by: Majumdar, Arunabh
Published: (2026)
Rationale Dataset and Analysis for the Commit Messages of the Linux Kernel Out-of-Memory Killer
by: Dhaouadi, Mouna, et al.
Published: (2024)
by: Dhaouadi, Mouna, et al.
Published: (2024)
Improving Automated Secure Code Reviews: A Synthetic Dataset for Code Vulnerability Flaws
by: Centellas-Claros, Leonardo, et al.
Published: (2025)
by: Centellas-Claros, Leonardo, et al.
Published: (2025)
Brevity is the Soul of Wit: Condensing Code Changes to Improve Commit Message Generation
by: Kuang, Hongyu, et al.
Published: (2025)
by: Kuang, Hongyu, et al.
Published: (2025)
DevOps Automation Pipeline Deployment with IaC (Infrastructure as Code)
by: Saxena, Adarsh, et al.
Published: (2025)
by: Saxena, Adarsh, et al.
Published: (2025)
AI-Augmented CI/CD Pipelines: From Code Commit to Production with Autonomous Decisions
by: Baqar, Mohammad, et al.
Published: (2025)
by: Baqar, Mohammad, et al.
Published: (2025)
CommitSuite: A Comprehensive Benchmark for Commit Classification and Message Generation
by: Wan, Zirui, et al.
Published: (2026)
by: Wan, Zirui, et al.
Published: (2026)
SmartMLOps Studio: Design of an LLM-Integrated IDE with Automated MLOps Pipelines for Model Development and Monitoring
by: Jin, Jiawei, et al.
Published: (2025)
by: Jin, Jiawei, et al.
Published: (2025)
Dinkel: State-Aware and Granular Framework for Validating Graph Databases
by: Wüst, Celine, et al.
Published: (2024)
by: Wüst, Celine, et al.
Published: (2024)
Optimization is Better than Generation: Optimizing Commit Message Leveraging Human-written Commit Message
by: Li, Jiawei, et al.
Published: (2025)
by: Li, Jiawei, et al.
Published: (2025)
Improving Merge Pipeline Throughput in Continuous Integration via Pull Request Prioritization
by: Jungwirth, Maximilian, et al.
Published: (2025)
by: Jungwirth, Maximilian, et al.
Published: (2025)
RAG-Enhanced Commit Message Generation
by: Zhang, Linghao, et al.
Published: (2024)
by: Zhang, Linghao, et al.
Published: (2024)
Compilation Quotient (CQ): A Metric for the Compilation Hardness of Programming Languages
by: Szabo, Violet, et al.
Published: (2024)
by: Szabo, Violet, et al.
Published: (2024)
A Multi-Layer Testing Framework for Automated Data Quality Assurance in Cloud-Native ELT Pipelines
by: Gargouri, Ismail, et al.
Published: (2026)
by: Gargouri, Ismail, et al.
Published: (2026)
Generative AI to Generate Test Data Generators
by: Baudry, Benoit, et al.
Published: (2024)
by: Baudry, Benoit, et al.
Published: (2024)
From Natural Language to Executable Properties for Property-based Testing of Mobile Apps
by: Xiong, Yiheng, et al.
Published: (2026)
by: Xiong, Yiheng, et al.
Published: (2026)
Similar Items
-
Mokav: Execution-driven Differential Testing with LLMs
by: Etemadi, Khashayar, et al.
Published: (2024) -
VeCoGen: Automating Generation of Formally Verified C Code with Large Language Models
by: Sevenhuijsen, Merlijn, et al.
Published: (2024) -
PerfGen: Automated Performance Benchmark Generation for Big Data Analytics
by: Wang, Jiyuan, et al.
Published: (2024) -
CigaR: Cost-efficient Program Repair with LLMs
by: Hidvégi, Dávid, et al.
Published: (2024) -
LLM-based Property-based Test Generation for Guardrailing Cyber-Physical Systems
by: Etemadi, Khashayar, et al.
Published: (2025)