LAPIS: A Performance Portable, High Productivity Compiler Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Kelley, Brian, Rajamanickam, Sivasankaran |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CELLO: Co-designing Schedule and Hybrid Implicit/Explicit Buffer for Complex Tensor Reuse
by: Garg, Raveesh, et al.
Published: (2023)
by: Garg, Raveesh, et al.
Published: (2023)
HPDR: High-Performance Portable Scientific Data Reduction Framework
by: Chen, Jieyang, et al.
Published: (2025)
by: Chen, Jieyang, et al.
Published: (2025)
Jet: Multilevel Graph Partitioning on Graphics Processing Units
by: Gilbert, Michael S., et al.
Published: (2023)
by: Gilbert, Michael S., et al.
Published: (2023)
Portability Efficiency Approach for Calculating Performance Portability
by: Marowka, Ami
Published: (2024)
by: Marowka, Ami
Published: (2024)
Enhancing Real-Time Master Data Management with Complex Match and Merge Algorithms
by: Rajamanickam, Durai
Published: (2024)
by: Rajamanickam, Durai
Published: (2024)
miniLB: A Performance Portability Study of Lattice-Boltzmann Simulations
by: Crisci, Luigi, et al.
Published: (2024)
by: Crisci, Luigi, et al.
Published: (2024)
XaaS Containers: Performance-Portable Representation With Source and IR Containers
by: Copik, Marcin, et al.
Published: (2025)
by: Copik, Marcin, et al.
Published: (2025)
A Portable Framework for Accelerating Stencil Computations on Modern Node Architectures
by: Sai, Ryuichi, et al.
Published: (2023)
by: Sai, Ryuichi, et al.
Published: (2023)
Towards High-Performance and Portable Molecular Docking on CPUs through Vectorization
by: Accordi, Gianmarco, et al.
Published: (2025)
by: Accordi, Gianmarco, et al.
Published: (2025)
High-Performance Portable GPU Primitives for Arbitrary Types and Operators in Julia
by: Pilliat, Emmanuel
Published: (2026)
by: Pilliat, Emmanuel
Published: (2026)
Closer in the Gap: Towards Portable Performance on RISC-V Vector Processors
by: Shi, Ruimin, et al.
Published: (2026)
by: Shi, Ruimin, et al.
Published: (2026)
Beyond Exascale: Dataflow Domain Translation on a Cerebras Cluster
by: Oppelstrup, Tomas, et al.
Published: (2025)
by: Oppelstrup, Tomas, et al.
Published: (2025)
Performance Portable Monte Carlo Particle Transport on Intel, NVIDIA, and AMD GPUs
by: Tramm, John, et al.
Published: (2024)
by: Tramm, John, et al.
Published: (2024)
A Parallel and Highly-Portable HPC Poisson Solver: Preconditioned Bi-CGSTAB with alpaka
by: Pennati, Luca, et al.
Published: (2025)
by: Pennati, Luca, et al.
Published: (2025)
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
by: Andersson, Måns I., et al.
Published: (2025)
by: Andersson, Måns I., et al.
Published: (2025)
Analyzing the Performance Portability of SYCL across CPUs, GPUs, and Hybrid Systems with SW Sequence Alignment
by: Costanzo, Manuel, et al.
Published: (2024)
by: Costanzo, Manuel, et al.
Published: (2024)
Taking GPU Programming Models to Task for Performance Portability
by: Davis, Joshua H., et al.
Published: (2024)
by: Davis, Joshua H., et al.
Published: (2024)
HP-MDR: High-performance and Portable Data Refactoring and Progressive Retrieval with Advanced GPUs
by: Li, Yanliang, et al.
Published: (2025)
by: Li, Yanliang, et al.
Published: (2025)
Zen-Attention: A Compiler Framework for Dynamic Attention Folding on AMD NPUs
by: Deshmukh, Aadesh, et al.
Published: (2025)
by: Deshmukh, Aadesh, et al.
Published: (2025)
Syndeo: Portable Ray Clusters with Secure Containerization
by: Li, William, et al.
Published: (2024)
by: Li, William, et al.
Published: (2024)
DeepCompile: A Compiler-Driven Approach to Optimizing Distributed Deep Learning Training
by: Tanaka, Masahiro, et al.
Published: (2025)
by: Tanaka, Masahiro, et al.
Published: (2025)
Implementing Multi-GPU Scientific Computing Miniapps Across Performance Portable Frameworks
by: Villalobos, Johansell, et al.
Published: (2025)
by: Villalobos, Johansell, et al.
Published: (2025)
PSI/J: A Portable Interface for Submitting, Monitoring, and Managing Jobs
by: Hategan-Marandiuc, Mihael, et al.
Published: (2023)
by: Hategan-Marandiuc, Mihael, et al.
Published: (2023)
HPCTransCompile: An AI Compiler Generated Dataset for High-Performance CUDA Transpilation and LLM Preliminary Exploration
by: Lv, Jiaqi, et al.
Published: (2025)
by: Lv, Jiaqi, et al.
Published: (2025)
Maple: A Multi-agent System for Portable Deep Learning across Clusters
by: Wu, Molang, et al.
Published: (2025)
by: Wu, Molang, et al.
Published: (2025)
Efficient and Portable Support for Overdecomposition on Distributed Memory GPGPU Platforms
by: Bhosale, Aditya, et al.
Published: (2026)
by: Bhosale, Aditya, et al.
Published: (2026)
Understanding Layered Portability from HPC to Cloud in Containerized Environments
by: Medeiros, Daniel, et al.
Published: (2024)
by: Medeiros, Daniel, et al.
Published: (2024)
Portable, heterogeneous ensemble workflows at scale using libEnsemble
by: Hudson, Stephen, et al.
Published: (2024)
by: Hudson, Stephen, et al.
Published: (2024)
A Massively Parallel Performance Portable Free-space Spectral Poisson Solver
by: Mayani, Sonali, et al.
Published: (2024)
by: Mayani, Sonali, et al.
Published: (2024)
DiOMP-Offloading: Toward Portable Distributed Heterogeneous OpenMP
by: Shan, Baodi, et al.
Published: (2025)
by: Shan, Baodi, et al.
Published: (2025)
HPC Containers for EBRAINS: Towards Portable Cross-Domain Software Environment
by: Singh, Krishna Kant, et al.
Published: (2026)
by: Singh, Krishna Kant, et al.
Published: (2026)
XaaS: Acceleration as a Service to Enable Productive High-Performance Cloud Computing
by: Hoefler, Torsten, et al.
Published: (2024)
by: Hoefler, Torsten, et al.
Published: (2024)
Workflow Mini-Apps: Portable, Scalable, Tunable & Faithful Representations of Scientific Workflows
by: Kilic, Ozgur Ozan, et al.
Published: (2024)
by: Kilic, Ozgur Ozan, et al.
Published: (2024)
Towards Portability at Scale: A Cross-Architecture Performance Evaluation of a GPU-enabled Shallow Water Solver
by: Villalobos, Johansell, et al.
Published: (2025)
by: Villalobos, Johansell, et al.
Published: (2025)
Gensor: A Graph-based Construction Tensor Compilation Method for Deep Learning
by: Liu, Hangda, et al.
Published: (2025)
by: Liu, Hangda, et al.
Published: (2025)
EvoSort: A Genetic-Algorithm-Based Adaptive Parallel Sorting Framework for Large-Scale High Performance Computing
by: Raj, Shashank, et al.
Published: (2025)
by: Raj, Shashank, et al.
Published: (2025)
Efficient Parallel Compilation and Profiling of Quantum Circuits at Large Scales
by: Moore, Jane, et al.
Published: (2026)
by: Moore, Jane, et al.
Published: (2026)
Toward Portable GPU Performance: Julia Recursive Implementation of TRMM and TRSM
by: Carrica, Vicki, et al.
Published: (2025)
by: Carrica, Vicki, et al.
Published: (2025)
LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU
by: Liao, Changyue, et al.
Published: (2024)
by: Liao, Changyue, et al.
Published: (2024)
Triton-distributed: Programming Overlapping Kernels on Distributed AI Systems with the Triton Compiler
by: Zheng, Size, et al.
Published: (2025)
by: Zheng, Size, et al.
Published: (2025)
Similar Items
-
CELLO: Co-designing Schedule and Hybrid Implicit/Explicit Buffer for Complex Tensor Reuse
by: Garg, Raveesh, et al.
Published: (2023) -
HPDR: High-Performance Portable Scientific Data Reduction Framework
by: Chen, Jieyang, et al.
Published: (2025) -
Jet: Multilevel Graph Partitioning on Graphics Processing Units
by: Gilbert, Michael S., et al.
Published: (2023) -
Portability Efficiency Approach for Calculating Performance Portability
by: Marowka, Ami
Published: (2024) -
Enhancing Real-Time Master Data Management with Complex Match and Merge Algorithms
by: Rajamanickam, Durai
Published: (2024)