A Parallel and Highly-Portable HPC Poisson Solver: Preconditioned Bi-CGSTAB with alpaka
Fuente:
arXiv
Salvato in:
| Autori principali: | Pennati, Luca, Andersson, Måns I., Steiniger, Klaus, Widera, Rene, Narwal, Tapish, Bussmann, Michael, Markidis, Stefano |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
di: Andersson, Måns I., et al.
Pubblicazione: (2025)
di: Andersson, Måns I., et al.
Pubblicazione: (2025)
Enabling High-Throughput Parallel I/O in Particle-in-Cell Monte Carlo Simulations with openPMD and Darshan I/O Monitoring
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
Mass Matrix Assembly on Tensor Cores for Implicit Particle-In-Cell Methods
di: Pennati, Luca, et al.
Pubblicazione: (2026)
di: Pennati, Luca, et al.
Pubblicazione: (2026)
What is Quantum Parallelism, Anyhow?
di: Markidis, Stefano
Pubblicazione: (2024)
di: Markidis, Stefano
Pubblicazione: (2024)
Enabling AI Deep Potentials for Ab Initio-quality Molecular Dynamics Simulations in GROMACS
di: Hu, Andong, et al.
Pubblicazione: (2026)
di: Hu, Andong, et al.
Pubblicazione: (2026)
A Massively Parallel Performance Portable Free-space Spectral Poisson Solver
di: Mayani, Sonali, et al.
Pubblicazione: (2024)
di: Mayani, Sonali, et al.
Pubblicazione: (2024)
Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS
di: Pennati, Luca, et al.
Pubblicazione: (2026)
di: Pennati, Luca, et al.
Pubblicazione: (2026)
QPU Micro-Kernels for Stencil Computation
di: Markidis, Stefano, et al.
Pubblicazione: (2025)
di: Markidis, Stefano, et al.
Pubblicazione: (2025)
Optimizations on Graph-Level for Domain Specific Computations in Julia and Application to QED
di: Reinhard, Anton, et al.
Pubblicazione: (2025)
di: Reinhard, Anton, et al.
Pubblicazione: (2025)
Krylov Solvers for Interior Point Methods with Applications in Radiation Therapy and Support Vector Machines
di: Liu, Felix, et al.
Pubblicazione: (2023)
di: Liu, Felix, et al.
Pubblicazione: (2023)
Leveraging HPC Profiling & Tracing Tools to Understand the Performance of Particle-in-Cell Monte Carlo Simulations
di: Williams, Jeremy J., et al.
Pubblicazione: (2023)
di: Williams, Jeremy J., et al.
Pubblicazione: (2023)
Understanding Layered Portability from HPC to Cloud in Containerized Environments
di: Medeiros, Daniel, et al.
Pubblicazione: (2024)
di: Medeiros, Daniel, et al.
Pubblicazione: (2024)
HPC Containers for EBRAINS: Towards Portable Cross-Domain Software Environment
di: Singh, Krishna Kant, et al.
Pubblicazione: (2026)
di: Singh, Krishna Kant, et al.
Pubblicazione: (2026)
Boosting Performance of Iterative Applications on GPUs: Kernel Batching with CUDA Graphs
di: Ekelund, Jonah, et al.
Pubblicazione: (2025)
di: Ekelund, Jonah, et al.
Pubblicazione: (2025)
FlashMP: Fast Discrete Transform-Based Solver for Preconditioning Maxwell's Equations on GPUs
di: Zhang, Haoyuan, et al.
Pubblicazione: (2025)
di: Zhang, Haoyuan, et al.
Pubblicazione: (2025)
ParaLog: Consistent Host-side Logging for Parallel Checkpoints
di: Chien, Steven W. D., et al.
Pubblicazione: (2024)
di: Chien, Steven W. D., et al.
Pubblicazione: (2024)
Parallelizing Drug Discovery: HPC Pipelines for Alzheimer's Molecular Docking and Simulation
di: Alliata, Paul Ruiz, et al.
Pubblicazione: (2025)
di: Alliata, Paul Ruiz, et al.
Pubblicazione: (2025)
Parallel Paradigms in Modern HPC: A Comparative Analysis of MPI, OpenMP, and CUDA
di: ALHafez, Nizar, et al.
Pubblicazione: (2025)
di: ALHafez, Nizar, et al.
Pubblicazione: (2025)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
di: Ather, Hammad, et al.
Pubblicazione: (2024)
di: Ather, Hammad, et al.
Pubblicazione: (2024)
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC
di: Bitan, Tomer, et al.
Pubblicazione: (2025)
di: Bitan, Tomer, et al.
Pubblicazione: (2025)
Towards Exascale Computing for Astrophysical Simulation Leveraging the Leonardo EuroHPC System
di: Shukla, Nitin, et al.
Pubblicazione: (2025)
di: Shukla, Nitin, et al.
Pubblicazione: (2025)
Harnessing CUDA-Q's MPS for Tensor Network Simulations of Large-Scale Quantum Circuits
di: Schieffer, Gabin, et al.
Pubblicazione: (2025)
di: Schieffer, Gabin, et al.
Pubblicazione: (2025)
Portability Efficiency Approach for Calculating Performance Portability
di: Marowka, Ami
Pubblicazione: (2024)
di: Marowka, Ami
Pubblicazione: (2024)
Characterizing the Performance of the Implicit Massively Parallel Particle-in-Cell iPIC3D Code
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
Multi-GPU Acceleration of PALABOS Fluid Solver using C++ Standard Parallelism
di: Latt, Jonas, et al.
Pubblicazione: (2025)
di: Latt, Jonas, et al.
Pubblicazione: (2025)
Understanding Data Movement in AMD Multi-GPU Systems with Infinity Fabric
di: Schieffer, Gabin, et al.
Pubblicazione: (2024)
di: Schieffer, Gabin, et al.
Pubblicazione: (2024)
Linear Complexity $\mathcal{H}^2$ Direct Solver for Fine-Grained Parallel Architectures
di: Boukaram, Wajih, et al.
Pubblicazione: (2025)
di: Boukaram, Wajih, et al.
Pubblicazione: (2025)
Integrating and Characterizing HPC Task Runtime Systems for hybrid AI-HPC workloads
di: Merzky, Andre, et al.
Pubblicazione: (2025)
di: Merzky, Andre, et al.
Pubblicazione: (2025)
HPC-based Solvers of Minimisation Problems for Signal Processing
di: Cammarasana, Simone, et al.
Pubblicazione: (2023)
di: Cammarasana, Simone, et al.
Pubblicazione: (2023)
Beyond Pre-Training: The Full Lifecycle of Foundation Models on HPC Systems
di: Conciatore, Dino, et al.
Pubblicazione: (2026)
di: Conciatore, Dino, et al.
Pubblicazione: (2026)
Energy-aware operation of HPC systems in Germany
di: Suarez, Estela, et al.
Pubblicazione: (2024)
di: Suarez, Estela, et al.
Pubblicazione: (2024)
Application Experiences on a GPU-Accelerated Arm-based HPC Testbed
di: Elwasif, Wael, et al.
Pubblicazione: (2022)
di: Elwasif, Wael, et al.
Pubblicazione: (2022)
HPC with Enhanced User Separation
di: Prout, Andrew, et al.
Pubblicazione: (2024)
di: Prout, Andrew, et al.
Pubblicazione: (2024)
Analysis of the carbon footprint of HPC
di: Benhari, Abdessalam, et al.
Pubblicazione: (2025)
di: Benhari, Abdessalam, et al.
Pubblicazione: (2025)
IOAgent: Democratizing Trustworthy HPC I/O Performance Diagnosis Capability via LLMs
di: Egersdoerfer, Chris, et al.
Pubblicazione: (2026)
di: Egersdoerfer, Chris, et al.
Pubblicazione: (2026)
Towards Portability at Scale: A Cross-Architecture Performance Evaluation of a GPU-enabled Shallow Water Solver
di: Villalobos, Johansell, et al.
Pubblicazione: (2025)
di: Villalobos, Johansell, et al.
Pubblicazione: (2025)
On the Convergence of Malleability and the HPC PowerStack: Exploiting Dynamism in Over-Provisioned and Power-Constrained HPC Systems
di: Arima, Eishi, et al.
Pubblicazione: (2024)
di: Arima, Eishi, et al.
Pubblicazione: (2024)
MRSch: Multi-Resource Scheduling for HPC
di: Li, Boyang, et al.
Pubblicazione: (2024)
di: Li, Boyang, et al.
Pubblicazione: (2024)
HPC-Coder: Modeling Parallel Programs using Large Language Models
di: Nichols, Daniel, et al.
Pubblicazione: (2023)
di: Nichols, Daniel, et al.
Pubblicazione: (2023)
UNR: Unified Notifiable RMA Library for HPC
di: Feng, Guangnan, et al.
Pubblicazione: (2024)
di: Feng, Guangnan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
di: Andersson, Måns I., et al.
Pubblicazione: (2025) -
Enabling High-Throughput Parallel I/O in Particle-in-Cell Monte Carlo Simulations with openPMD and Darshan I/O Monitoring
di: Williams, Jeremy J., et al.
Pubblicazione: (2024) -
Mass Matrix Assembly on Tensor Cores for Implicit Particle-In-Cell Methods
di: Pennati, Luca, et al.
Pubblicazione: (2026) -
What is Quantum Parallelism, Anyhow?
di: Markidis, Stefano
Pubblicazione: (2024) -
Enabling AI Deep Potentials for Ab Initio-quality Molecular Dynamics Simulations in GROMACS
di: Hu, Andong, et al.
Pubblicazione: (2026)