VecTrans: Enhancing Compiler Auto-Vectorization through LLM-Assisted Code Transformations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Zhongchun, Wu, Kan, Cheng, Long, Li, Lu, Rocha, Rodrigo C. O., Liu, Tianyi, Wei, Wei, Zeng, Jianjiang, Zhang, Xianwei, Gao, Yaoqing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Empirical Study on the Performance and Energy Usage of Compiled Python Code
von: Stoico, Vincenzo, et al.
Veröffentlicht: (2025)
von: Stoico, Vincenzo, et al.
Veröffentlicht: (2025)
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
von: Taneja, Jubi, et al.
Veröffentlicht: (2024)
von: Taneja, Jubi, et al.
Veröffentlicht: (2024)
MLKAPS: Machine Learning and Adaptive Sampling for HPC Kernel Auto-tuning
von: Jam, Mathys, et al.
Veröffentlicht: (2025)
von: Jam, Mathys, et al.
Veröffentlicht: (2025)
Should AI Optimize Your Code? A Comparative Study of Classical Optimizing Compilers Versus Current Large Language Models
von: Rosas, Miguel Romero, et al.
Veröffentlicht: (2024)
von: Rosas, Miguel Romero, et al.
Veröffentlicht: (2024)
RAO-SS: A Prototype of Run-time Auto-tuning Facility for Sparse Direct Solvers
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
Towards a Higher Roofline for Matrix-Vector Multiplication in Matrix-Free HOSFEM
von: Cao, Zijian, et al.
Veröffentlicht: (2025)
von: Cao, Zijian, et al.
Veröffentlicht: (2025)
QOPS: A Compiler Framework for Quantum Circuit Simulation Acceleration with Profile Guided Optimizations
von: Wu, Yu-Tsung, et al.
Veröffentlicht: (2024)
von: Wu, Yu-Tsung, et al.
Veröffentlicht: (2024)
A Magnified View into Heterogeneous-ISA Thread Migration Performance without State Transformation
von: Mavrogeorgis, Nikolaos, et al.
Veröffentlicht: (2025)
von: Mavrogeorgis, Nikolaos, et al.
Veröffentlicht: (2025)
Do AI Models Dream of Faster Code? An Empirical Study on LLM-Proposed Performance Improvements in Real-World Software
von: Yi, Lirong, et al.
Veröffentlicht: (2025)
von: Yi, Lirong, et al.
Veröffentlicht: (2025)
On the Compression of Language Models for Code: An Empirical Study on CodeBERT
von: d'Aloisio, Giordano, et al.
Veröffentlicht: (2024)
von: d'Aloisio, Giordano, et al.
Veröffentlicht: (2024)
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations
von: Rani, Pooja, et al.
Veröffentlicht: (2025)
von: Rani, Pooja, et al.
Veröffentlicht: (2025)
Library Liberation: Competitive Performance Matmul Through Compiler-composed Nanokernels
von: Thangamani, Arun, et al.
Veröffentlicht: (2025)
von: Thangamani, Arun, et al.
Veröffentlicht: (2025)
Learning Performance-Improving Code Edits
von: Shypula, Alexander, et al.
Veröffentlicht: (2023)
von: Shypula, Alexander, et al.
Veröffentlicht: (2023)
Xabclib:A Fully Auto-tuned Sparse Iterative Solver
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
Stencil-Lifting: Hierarchical Recursive Lifting System for Extracting Summary of Stencil Kernel in Legacy Codes
von: Li, Mingyi, et al.
Veröffentlicht: (2025)
von: Li, Mingyi, et al.
Veröffentlicht: (2025)
Enhancing Energy-Awareness in Deep Learning through Fine-Grained Energy Measurement
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2023)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2023)
Estimating the Energy Footprint of Software Systems: a Primer
von: Castor, Fernando
Veröffentlicht: (2024)
von: Castor, Fernando
Veröffentlicht: (2024)
FaaSter Troubleshooting -- Evaluating Distributed Tracing Approaches for Serverless Applications
von: Borges, Maria C., et al.
Veröffentlicht: (2021)
von: Borges, Maria C., et al.
Veröffentlicht: (2021)
gigiProfiler: Diagnosing Performance Issues by Uncovering Application Resource Bottlenecks
von: Hu, Yigong, et al.
Veröffentlicht: (2025)
von: Hu, Yigong, et al.
Veröffentlicht: (2025)
An Empirical Study on Method-Level Performance Evolution in Open-Source Java Projects
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
On the Role of Search Budgets in Model-Based Software Refactoring Optimization
von: Diaz-Pace, J. Andres, et al.
Veröffentlicht: (2023)
von: Diaz-Pace, J. Andres, et al.
Veröffentlicht: (2023)
What Is the Cost of Energy Monitoring? An Empirical Study on the Overhead of RAPL-Based Tools
von: Diamond, Jeremy, et al.
Veröffentlicht: (2026)
von: Diamond, Jeremy, et al.
Veröffentlicht: (2026)
Performance of Genetic Algorithms in the Context of Software Model Refactoring
von: Cortellessa, Vittorio, et al.
Veröffentlicht: (2023)
von: Cortellessa, Vittorio, et al.
Veröffentlicht: (2023)
Towards Assessing Spread in Sets of Software Architecture Designs
von: Cortellessa, Vittorio, et al.
Veröffentlicht: (2024)
von: Cortellessa, Vittorio, et al.
Veröffentlicht: (2024)
Scalable Software as a Service Architecture
von: Dedase, Ardy
Veröffentlicht: (2024)
von: Dedase, Ardy
Veröffentlicht: (2024)
Optimas: An Intelligent Analytics-Informed Generative AI Framework for Performance Optimization
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2026)
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2026)
A Test for FLOPs as a Discriminant for Linear Algebra Algorithms
von: Sankaran, Aravind, et al.
Veröffentlicht: (2022)
von: Sankaran, Aravind, et al.
Veröffentlicht: (2022)
Formal Analysis of Metastable Failures in Software Systems
von: Alvaro, Peter, et al.
Veröffentlicht: (2025)
von: Alvaro, Peter, et al.
Veröffentlicht: (2025)
Employing Software Diversity in Cloud Microservices to Engineer Reliable and Performant Systems
von: Akhtarian, Nazanin, et al.
Veröffentlicht: (2024)
von: Akhtarian, Nazanin, et al.
Veröffentlicht: (2024)
Faster Base64 Encoding and Decoding Using AVX2 Instructions
von: Muła, Wojciech, et al.
Veröffentlicht: (2017)
von: Muła, Wojciech, et al.
Veröffentlicht: (2017)
Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning
von: Su, Qisheng, et al.
Veröffentlicht: (2026)
von: Su, Qisheng, et al.
Veröffentlicht: (2026)
Energy Patterns for Web: An Exploratory Study
von: Rani, Pooja, et al.
Veröffentlicht: (2024)
von: Rani, Pooja, et al.
Veröffentlicht: (2024)
SysLLMatic: Large Language Models are Software System Optimizers
von: Peng, Huiyun, et al.
Veröffentlicht: (2025)
von: Peng, Huiyun, et al.
Veröffentlicht: (2025)
Efficiently Ranking Software Variants with Minimal Benchmarks
von: Matricon, Théo, et al.
Veröffentlicht: (2025)
von: Matricon, Théo, et al.
Veröffentlicht: (2025)
Energy-Efficient Software Development: A Multi-dimensional Empirical Analysis of Stack Overflow
von: Jin, Bihui, et al.
Veröffentlicht: (2024)
von: Jin, Bihui, et al.
Veröffentlicht: (2024)
Tracing Optimization for Performance Modeling and Regression Detection
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
A Model-driven Approach for Continuous Performance Engineering in Microservice-based Systems
von: Cortellessa, Vittorio, et al.
Veröffentlicht: (2023)
von: Cortellessa, Vittorio, et al.
Veröffentlicht: (2023)
FlipFlop: A Static Analysis-based Energy Optimization Framework for GPU Kernels
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
An Empirical Study on How Architectural Topology Affects Microservice Performance and Energy Usage
von: Ristova, Irena, et al.
Veröffentlicht: (2026)
von: Ristova, Irena, et al.
Veröffentlicht: (2026)
Enabling Performant and Flexible Model-Internal Observability for LLM Inference
von: Yu, Nengneng, et al.
Veröffentlicht: (2026)
von: Yu, Nengneng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
An Empirical Study on the Performance and Energy Usage of Compiled Python Code
von: Stoico, Vincenzo, et al.
Veröffentlicht: (2025) -
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
von: Taneja, Jubi, et al.
Veröffentlicht: (2024) -
MLKAPS: Machine Learning and Adaptive Sampling for HPC Kernel Auto-tuning
von: Jam, Mathys, et al.
Veröffentlicht: (2025) -
Should AI Optimize Your Code? A Comparative Study of Classical Optimizing Compilers Versus Current Large Language Models
von: Rosas, Miguel Romero, et al.
Veröffentlicht: (2024) -
RAO-SS: A Prototype of Run-time Auto-tuning Facility for Sparse Direct Solvers
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)