Counting Without Running: Evaluating LLMs' Reasoning About Code Complexity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bolet, Gregory, Georgakoudis, Giorgis, Parasyris, Konstantinos, Menon, Harshitha, Hasabnis, Niranjan, Cameron, Kirk W., Oren, Gal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Large Language Models Predict Parallel Code Performance?
von: Bolet, Gregory, et al.
Veröffentlicht: (2025)
von: Bolet, Gregory, et al.
Veröffentlicht: (2025)
Taking GPU Programming Models to Task for Performance Portability
von: Davis, Joshua H., et al.
Veröffentlicht: (2024)
von: Davis, Joshua H., et al.
Veröffentlicht: (2024)
Record-Remix-Replay: Hierarchical GPU Kernel Optimization using Evolutionary Search
von: Nichols, Daniel, et al.
Veröffentlicht: (2026)
von: Nichols, Daniel, et al.
Veröffentlicht: (2026)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
von: Nichols, Daniel, et al.
Veröffentlicht: (2025)
von: Nichols, Daniel, et al.
Veröffentlicht: (2025)
HPAC-ML: A Programming Model for Embedding ML Surrogates in Scientific Applications
von: Fink, Zane, et al.
Veröffentlicht: (2024)
von: Fink, Zane, et al.
Veröffentlicht: (2024)
LLMs as Packagers of HPC Software
von: Melone, Caetano, et al.
Veröffentlicht: (2025)
von: Melone, Caetano, et al.
Veröffentlicht: (2025)
Leveraging AI for Productive and Trustworthy HPC Software: Challenges and Research Directions
von: Teranishi, Keita, et al.
Veröffentlicht: (2025)
von: Teranishi, Keita, et al.
Veröffentlicht: (2025)
ParaCodex: A Profiling-Guided Autonomous Coding Agent for Reliable Parallel Code Generation and Translation
von: Kaplan, Erel, et al.
Veröffentlicht: (2026)
von: Kaplan, Erel, et al.
Veröffentlicht: (2026)
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC
von: Bitan, Tomer, et al.
Veröffentlicht: (2025)
von: Bitan, Tomer, et al.
Veröffentlicht: (2025)
An Auto-tuning Method for Run-time Data Transformation for Sparse Matrix-Vector Multiplication
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
von: Katagiri, Takahiro, et al.
Veröffentlicht: (2024)
Fine-Grained Energy Prediction For Parallellized LLM Inference With PIE-P
von: Dutt, Anurag, et al.
Veröffentlicht: (2025)
von: Dutt, Anurag, et al.
Veröffentlicht: (2025)
OMPILOT: Harnessing Transformer Models for Auto Parallelization to Shared Memory Computing Paradigms
von: Bhattacharjee, Arijit, et al.
Veröffentlicht: (2025)
von: Bhattacharjee, Arijit, et al.
Veröffentlicht: (2025)
Optimizing Agentic Language Model Inference via Speculative Tool Calls
von: Nichols, Daniel, et al.
Veröffentlicht: (2025)
von: Nichols, Daniel, et al.
Veröffentlicht: (2025)
Experimental Assessment of Containers Running on Top of Virtual Machines
von: Aqasizade, Hossein, et al.
Veröffentlicht: (2024)
von: Aqasizade, Hossein, et al.
Veröffentlicht: (2024)
Understanding Inference Scaling for LLMs: Bottlenecks, Trade-offs, and Performance Principles
von: Arif, Moiz, et al.
Veröffentlicht: (2026)
von: Arif, Moiz, et al.
Veröffentlicht: (2026)
Should I Run My Cloud Benchmark on Black Friday?
von: Henning, Sören, et al.
Veröffentlicht: (2025)
von: Henning, Sören, et al.
Veröffentlicht: (2025)
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
von: Andersson, Måns I., et al.
Veröffentlicht: (2025)
von: Andersson, Måns I., et al.
Veröffentlicht: (2025)
Architecture Specific Generation of Large Scale Lattice Boltzmann Methods for Sparse Complex Geometries
von: Suffa, Philipp, et al.
Veröffentlicht: (2024)
von: Suffa, Philipp, et al.
Veröffentlicht: (2024)
Fake Runs, Real Fixes -- Analyzing xPU Performance Through Simulation
von: Zarkadas, Ioannis, et al.
Veröffentlicht: (2025)
von: Zarkadas, Ioannis, et al.
Veröffentlicht: (2025)
When Should I Run My Application Benchmark?: Studying Cloud Performance Variability for the Case of Stream Processing Applications
von: Henning, Sören, et al.
Veröffentlicht: (2025)
von: Henning, Sören, et al.
Veröffentlicht: (2025)
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
von: Chen, Le, et al.
Veröffentlicht: (2024)
von: Chen, Le, et al.
Veröffentlicht: (2024)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
Testing the Unknown: A Framework for OpenMP Testing via Random Program Generation
von: Laguna, Ignacio, et al.
Veröffentlicht: (2024)
von: Laguna, Ignacio, et al.
Veröffentlicht: (2024)
On Optimal Batch Size in Coded Computing
von: Saha, Swapnil, et al.
Veröffentlicht: (2025)
von: Saha, Swapnil, et al.
Veröffentlicht: (2025)
CPMA: An Efficient Batch-Parallel Compressed Set Without Pointers
von: Wheatman, Brian, et al.
Veröffentlicht: (2023)
von: Wheatman, Brian, et al.
Veröffentlicht: (2023)
CoNST: Code Generator for Sparse Tensor Networks
von: Raje, Saurabh, et al.
Veröffentlicht: (2024)
von: Raje, Saurabh, et al.
Veröffentlicht: (2024)
LEGO: A Layout Expression Language for Code Generation of Hierarchical Mapping
von: Tavakkoli, Amir Mohammad, et al.
Veröffentlicht: (2025)
von: Tavakkoli, Amir Mohammad, et al.
Veröffentlicht: (2025)
Extracting Practical, Actionable Energy Insights from Supercomputer Telemetry and Logs
von: Cornelius, Melanie, et al.
Veröffentlicht: (2025)
von: Cornelius, Melanie, et al.
Veröffentlicht: (2025)
Scaling Large-scale GNN Training to Thousands of Processors on CPU-based Supercomputers
von: Zhuang, Chen, et al.
Veröffentlicht: (2024)
von: Zhuang, Chen, et al.
Veröffentlicht: (2024)
Profiling and optimization of multi-card GPU machine learning jobs
von: Lawenda, Marcin, et al.
Veröffentlicht: (2025)
von: Lawenda, Marcin, et al.
Veröffentlicht: (2025)
Optimal Parallel Scheduling under Concave Speedup Functions
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)
von: Li, Chengzhang, et al.
Veröffentlicht: (2025)
WebAssembly and Unikernels: A Comparative Study for Serverless at the Edge
von: Besozzi, Valerio, et al.
Veröffentlicht: (2025)
von: Besozzi, Valerio, et al.
Veröffentlicht: (2025)
Efficient GPU-Centered Singular Value Decomposition Using the Divide-and-Conquer Method
von: Liu, Shifang, et al.
Veröffentlicht: (2025)
von: Liu, Shifang, et al.
Veröffentlicht: (2025)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
von: Mao, Ying, et al.
Veröffentlicht: (2020)
von: Mao, Ying, et al.
Veröffentlicht: (2020)
Staging Blocked Evaluation over Structured Sparse Matrices
von: Das, Pratyush, et al.
Veröffentlicht: (2024)
von: Das, Pratyush, et al.
Veröffentlicht: (2024)
Cloud Performance Decomposition for Long-Term Performance Engineering: A Case Study
von: Debnath, Shimul, et al.
Veröffentlicht: (2026)
von: Debnath, Shimul, et al.
Veröffentlicht: (2026)
Collaborative Processing for Multi-Tenant Inference on Memory-Constrained Edge TPUs
von: Ng, Nathan, et al.
Veröffentlicht: (2026)
von: Ng, Nathan, et al.
Veröffentlicht: (2026)
Preliminary report: Initial evaluation of StdPar implementations on AMD GPUs for HPC
von: Lin, Wei-Chen, et al.
Veröffentlicht: (2024)
von: Lin, Wei-Chen, et al.
Veröffentlicht: (2024)
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
von: Sun, Tingyang, et al.
Veröffentlicht: (2026)
von: Sun, Tingyang, et al.
Veröffentlicht: (2026)
Reducing Tail Latencies Through Environment- and Neighbour-aware Thread Management
von: Jeffery, Andrew, et al.
Veröffentlicht: (2024)
von: Jeffery, Andrew, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Can Large Language Models Predict Parallel Code Performance?
von: Bolet, Gregory, et al.
Veröffentlicht: (2025) -
Taking GPU Programming Models to Task for Performance Portability
von: Davis, Joshua H., et al.
Veröffentlicht: (2024) -
Record-Remix-Replay: Hierarchical GPU Kernel Optimization using Evolutionary Search
von: Nichols, Daniel, et al.
Veröffentlicht: (2026) -
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
von: Nichols, Daniel, et al.
Veröffentlicht: (2025) -
HPAC-ML: A Programming Model for Embedding ML Surrogates in Scientific Applications
von: Fink, Zane, et al.
Veröffentlicht: (2024)