Teaching An Old Dog New Tricks: Porting Legacy Code to Heterogeneous Compute Architectures With Automated Code Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Nytko, Nicolas, Reisner, Andrew, Moulton, J. David, Olson, Luke N., West, Matthew |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Code once, Run Green: Automated Green Code Translation in Serverless Computing
by: Werner, Sebastian, et al.
Published: (2025)
by: Werner, Sebastian, et al.
Published: (2025)
Fully-Automated Code Generation for Efficient Computation of Sparse Matrix Permanents on GPUs
by: Elbek, Deniz, et al.
Published: (2025)
by: Elbek, Deniz, et al.
Published: (2025)
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
by: Rudi, Johann, et al.
Published: (2024)
by: Rudi, Johann, et al.
Published: (2024)
Distributed Matrix-Vector Multiplication: A Convolutional Coding Approach
by: Das, Anindya Bijoy, et al.
Published: (2019)
by: Das, Anindya Bijoy, et al.
Published: (2019)
Two-Stage Block Orthogonalization to Improve Performance of $s$-step GMRES
by: Yamazaki, Ichitaro, et al.
Published: (2024)
by: Yamazaki, Ichitaro, et al.
Published: (2024)
Random-sketching Techniques to Enhance the Numerical Stability of Block Orthogonalization Algorithms for s-step GMRES
by: Yamazaki, Ichitaro, et al.
Published: (2025)
by: Yamazaki, Ichitaro, et al.
Published: (2025)
Computing the Saturation Throughput for Heterogeneous p-CSMA in a General Wireless Network
by: Tarzjani, Faezeh Dehghan, et al.
Published: (2025)
by: Tarzjani, Faezeh Dehghan, et al.
Published: (2025)
On Similarity of Computational Kernels in our Codes and Proxies
by: McKinsey, Michael, et al.
Published: (2026)
by: McKinsey, Michael, et al.
Published: (2026)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
by: Higgins, Andrew J., et al.
Published: (2025)
by: Higgins, Andrew J., et al.
Published: (2025)
Barycentric Coded Distributed Computing with Flexible Recovery Threshold for Collaborative Mobile Edge Computing
by: Qiu, Houming, et al.
Published: (2025)
by: Qiu, Houming, et al.
Published: (2025)
Error Analysis of Matrix Multiplication Emulation Using Ozaki-II Scheme
by: Uchino, Yuki, et al.
Published: (2026)
by: Uchino, Yuki, et al.
Published: (2026)
Towards a GPU-Parallelization of the neXtSIM-DG Dynamical Core
by: Jendersie, Robert, et al.
Published: (2024)
by: Jendersie, Robert, et al.
Published: (2024)
Parallel simulation and adaptive mesh refinement for 3D elastostatic contact mechanics problems between deformable bodies
by: Epalle, Alexandre, et al.
Published: (2025)
by: Epalle, Alexandre, et al.
Published: (2025)
SUNDIALS Time Integrators for Exascale Applications with Many Independent ODE Systems
by: Balos, Cody J., et al.
Published: (2024)
by: Balos, Cody J., et al.
Published: (2024)
Neural Acceleration of Incomplete Cholesky Preconditioners
by: Booth, Joshua Dennis, et al.
Published: (2024)
by: Booth, Joshua Dennis, et al.
Published: (2024)
A simple GPU implementation of spectral-element methods for solving 3D Poisson type equations on rectangular domains and its applications
by: Liu, Xinyu, et al.
Published: (2023)
by: Liu, Xinyu, et al.
Published: (2023)
On some orthogonalization schemes in Tensor Train format
by: Coulaud, Olivier, et al.
Published: (2022)
by: Coulaud, Olivier, et al.
Published: (2022)
Asymptotic Analysis of a Leader Election Algorithm
by: Lavault, Christian, et al.
Published: (2006)
by: Lavault, Christian, et al.
Published: (2006)
A Parallel in Time Algorithm Based on ParaExp for Optimal Control Problems
by: Kwok, Felix, et al.
Published: (2024)
by: Kwok, Felix, et al.
Published: (2024)
Cucheb: A GPU implementation of the filtered Lanczos procedure
by: Aurentz, Jared L., et al.
Published: (2024)
by: Aurentz, Jared L., et al.
Published: (2024)
RAPTOR: Practical Numerical Profiling of Scientific Applications
by: Hoerold, Faveo, et al.
Published: (2025)
by: Hoerold, Faveo, et al.
Published: (2025)
GPU Accelerated Implicit Kinetic Meshfree Method based on Modified LU-SGS
by: Verma, Mayuri, et al.
Published: (2024)
by: Verma, Mayuri, et al.
Published: (2024)
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension
by: Coleman, Evan
Published: (2026)
by: Coleman, Evan
Published: (2026)
Algebraic Temporal Blocking for Sparse Iterative Solvers on Multi-Core CPUs
by: Alappat, Christie, et al.
Published: (2023)
by: Alappat, Christie, et al.
Published: (2023)
Modifying the Asynchronous Jacobi Method for Data Corruption Resilience
by: Vogl, Christopher J., et al.
Published: (2022)
by: Vogl, Christopher J., et al.
Published: (2022)
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC
by: Bitan, Tomer, et al.
Published: (2025)
by: Bitan, Tomer, et al.
Published: (2025)
Privacy-Preserving Coding Schemes for Multi-Access Distributed Computing Models
by: Sasi, Shanuja
Published: (2026)
by: Sasi, Shanuja
Published: (2026)
SUperman: Efficient Permanent Computation on GPUs
by: Elbek, Deniz, et al.
Published: (2025)
by: Elbek, Deniz, et al.
Published: (2025)
Approximated Coded Computing: Towards Fast, Private and Secure Distributed Machine Learning
by: Qiu, Houming, et al.
Published: (2024)
by: Qiu, Houming, et al.
Published: (2024)
DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration
by: Lai, Shih-Yu, et al.
Published: (2026)
by: Lai, Shih-Yu, et al.
Published: (2026)
Cooperative Gradient Coding
by: Weng, Shudi, et al.
Published: (2025)
by: Weng, Shudi, et al.
Published: (2025)
SIMT/GPU Data Race Verification using ISCC and Intermediary Code Representations: A Case Study
by: Osterhout, Andrew, et al.
Published: (2025)
by: Osterhout, Andrew, et al.
Published: (2025)
A Novel Approach to Translate Structural Aggregation Queries to MapReduce Code
by: Abdelmoniem, Ahmed M., et al.
Published: (2025)
by: Abdelmoniem, Ahmed M., et al.
Published: (2025)
Design and Implementation of Code Completion System Based on LLM and CodeBERT Hybrid Subsystem
by: Zhang, Bingbing, et al.
Published: (2025)
by: Zhang, Bingbing, et al.
Published: (2025)
Augur: Pre-Execution Energy Prediction for Workflow Tasks in Heterogeneous Clusters
by: West, Kathleen, et al.
Published: (2026)
by: West, Kathleen, et al.
Published: (2026)
Dynamic Detection of Inefficient Data Mapping Patterns in Heterogeneous OpenMP Applications
by: Marzen, Luke, et al.
Published: (2026)
by: Marzen, Luke, et al.
Published: (2026)
ParaCodex: A Profiling-Guided Autonomous Coding Agent for Reliable Parallel Code Generation and Translation
by: Kaplan, Erel, et al.
Published: (2026)
by: Kaplan, Erel, et al.
Published: (2026)
Memory-aware Adaptive Scheduling of Scientific Workflows on Heterogeneous Architectures
by: Kulagina, Svetlana, et al.
Published: (2025)
by: Kulagina, Svetlana, et al.
Published: (2025)
Benchmarking Machine Learning Applications on Heterogeneous Architecture using Reframe
by: Rae, Christopher, et al.
Published: (2024)
by: Rae, Christopher, et al.
Published: (2024)
General Coded Computing: Adversarial Settings
by: Moradi, Parsa, et al.
Published: (2025)
by: Moradi, Parsa, et al.
Published: (2025)
Similar Items
-
Code once, Run Green: Automated Green Code Translation in Serverless Computing
by: Werner, Sebastian, et al.
Published: (2025) -
Fully-Automated Code Generation for Efficient Computation of Sparse Matrix Permanents on GPUs
by: Elbek, Deniz, et al.
Published: (2025) -
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
by: Rudi, Johann, et al.
Published: (2024) -
Distributed Matrix-Vector Multiplication: A Convolutional Coding Approach
by: Das, Anindya Bijoy, et al.
Published: (2019) -
Two-Stage Block Orthogonalization to Improve Performance of $s$-step GMRES
by: Yamazaki, Ichitaro, et al.
Published: (2024)