An extension of C++ with memory-centric specifications for HPC to reduce memory footprints and streamline MPI development
Fuente:
arXiv
Salvato in:
| Autori principali: | Radtke, Pawel K., Barrera-Hinojosa, Cristian G., Ivkovic, Mladen, Weinzierl, Tobias |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Compiler-supported reduced precision and AoS-SoA transformations for heterogeneous hardware
di: Radtke, Pawel K., et al.
Pubblicazione: (2025)
di: Radtke, Pawel K., et al.
Pubblicazione: (2025)
Annotation-guided AoS-to-SoA conversions and GPU offloading with data views in C++
di: Radtke, Pawel K., et al.
Pubblicazione: (2025)
di: Radtke, Pawel K., et al.
Pubblicazione: (2025)
SYCL compute kernels for ExaHyPE
di: Loi, Chung Ming, et al.
Pubblicazione: (2023)
di: Loi, Chung Ming, et al.
Pubblicazione: (2023)
Annotation‐Guided AoS‐to‐SoA Conversions and GPU Offloading With Data Views in C++
di: Pawel K. Radtke, et al.
Pubblicazione: (2025)
di: Pawel K. Radtke, et al.
Pubblicazione: (2025)
A shared compilation stack for distributed-memory parallelism in stencil DSLs
di: Bisbas, George, et al.
Pubblicazione: (2024)
di: Bisbas, George, et al.
Pubblicazione: (2024)
Fast Higher-Order Interpolation and Restriction in ExaHyPE Avoiding Non-physical Reflections
di: Stokes, Timothy, et al.
Pubblicazione: (2025)
di: Stokes, Timothy, et al.
Pubblicazione: (2025)
Compiler support for semi-manual AoS-to-SoA conversions with data views
di: Radtke, Pawel K., et al.
Pubblicazione: (2024)
di: Radtke, Pawel K., et al.
Pubblicazione: (2024)
Distributed-memory Algorithms for Sparse Matrix Permutation, Extraction, and Assignment
di: Hassani, Elaheh, et al.
Pubblicazione: (2025)
di: Hassani, Elaheh, et al.
Pubblicazione: (2025)
Performance measurements of modern Fortran MPI applications with Score-P
di: Corbin, Gregor
Pubblicazione: (2025)
di: Corbin, Gregor
Pubblicazione: (2025)
Automated MPI-X code generation for scalable finite-difference solvers
di: Bisbas, George, et al.
Pubblicazione: (2023)
di: Bisbas, George, et al.
Pubblicazione: (2023)
On the Challenges of Energy-Efficiency Analysis in HPC Systems: Evaluating Synthetic Benchmarks and Gromacs
di: Machado, Rafael Ravedutti Lucio, et al.
Pubblicazione: (2025)
di: Machado, Rafael Ravedutti Lucio, et al.
Pubblicazione: (2025)
Experience converting a large mathematical software package written in C++ to C++20 modules
di: Bangerth, Wolfgang
Pubblicazione: (2025)
di: Bangerth, Wolfgang
Pubblicazione: (2025)
Enabling MPI communication within Numba/LLVM JIT-compiled Python code using numba-mpi v1.0
di: Derlatka, Kacper, et al.
Pubblicazione: (2024)
di: Derlatka, Kacper, et al.
Pubblicazione: (2024)
Towards a user-centric HPC-QC environment
di: Wennersteen, Aleksander, et al.
Pubblicazione: (2025)
di: Wennersteen, Aleksander, et al.
Pubblicazione: (2025)
svds-C: A Multi-Thread C Code for Computing Truncated Singular Value Decomposition
di: Feng, Xu, et al.
Pubblicazione: (2024)
di: Feng, Xu, et al.
Pubblicazione: (2024)
TTK is Getting MPI-Ready
di: Guillou, Eve Le, et al.
Pubblicazione: (2023)
di: Guillou, Eve Le, et al.
Pubblicazione: (2023)
TumorTwin: A python framework for patient-specific digital twins in oncology
di: Kapteyn, Michael, et al.
Pubblicazione: (2025)
di: Kapteyn, Michael, et al.
Pubblicazione: (2025)
SySTeC: A Symmetric Sparse Tensor Compiler
di: Patel, Radha, et al.
Pubblicazione: (2024)
di: Patel, Radha, et al.
Pubblicazione: (2024)
PennyLane-Lightning MPI: A massively scalable quantum circuit simulator based on distributed computing in CPU clusters
di: Kang, Ji-Hoon, et al.
Pubblicazione: (2025)
di: Kang, Ji-Hoon, et al.
Pubblicazione: (2025)
QMCPy: A Python Software for Randomized Low-Discrepancy Sequences, Quasi-Monte Carlo, and Fast Kernel Methods
di: Sorokin, Aleksei G
Pubblicazione: (2025)
di: Sorokin, Aleksei G
Pubblicazione: (2025)
Efficient space-time reduced order model for linear dynamical systems in Python using less than 120 lines of code
di: Kim, Youngkyu, et al.
Pubblicazione: (2020)
di: Kim, Youngkyu, et al.
Pubblicazione: (2020)
A comparison of two effective methods for reordering columns within supernodes
di: Karsavuran, M. Ozan, et al.
Pubblicazione: (2025)
di: Karsavuran, M. Ozan, et al.
Pubblicazione: (2025)
A modular and extensible library for parameterized terrain generation
di: Wallin, Erik
Pubblicazione: (2025)
di: Wallin, Erik
Pubblicazione: (2025)
Deriving Algorithms for Triangular Tridiagonalization a Skew-Symmetric Matrix
di: van de Geijn, Robert, et al.
Pubblicazione: (2023)
di: van de Geijn, Robert, et al.
Pubblicazione: (2023)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
di: Diehl, Patrick, et al.
Pubblicazione: (2025)
di: Diehl, Patrick, et al.
Pubblicazione: (2025)
Enhancing non-Perl bioinformatic applications with Perl: Building novel, component based applications using Object Orientation, PDL, Alien, FFI, Inline and OpenMP
di: Argyropoulos, Christos
Pubblicazione: (2024)
di: Argyropoulos, Christos
Pubblicazione: (2024)
Towards Richer Challenge Problems for Scientific Computing Correctness
di: Sottile, Matthew, et al.
Pubblicazione: (2025)
di: Sottile, Matthew, et al.
Pubblicazione: (2025)
Performant Tridiagonal Factorization of Skew-Symmetric Matrices
di: Satyarth, Ishna, et al.
Pubblicazione: (2024)
di: Satyarth, Ishna, et al.
Pubblicazione: (2024)
Porting the Nonlinear Optimization Library HiOp to Accelerator-Based Hardware Architectures
di: Peles, Slaven, et al.
Pubblicazione: (2026)
di: Peles, Slaven, et al.
Pubblicazione: (2026)
Interface for Sparse Linear Algebra Operations
di: Abdelfattah, Ahmad, et al.
Pubblicazione: (2024)
di: Abdelfattah, Ahmad, et al.
Pubblicazione: (2024)
Minimization of Nonlinear Energies in Python Using FEM and Automatic Differentiation Tools
di: Béreš, Michal, et al.
Pubblicazione: (2024)
di: Béreš, Michal, et al.
Pubblicazione: (2024)
On a vectorized basic linear algebra package for prototyping codes in MATLAB
di: Moskovka, Alexej, et al.
Pubblicazione: (2024)
di: Moskovka, Alexej, et al.
Pubblicazione: (2024)
QHyper: an integration library for hybrid quantum-classical optimization
di: Lamża, Tomasz, et al.
Pubblicazione: (2024)
di: Lamża, Tomasz, et al.
Pubblicazione: (2024)
Finch: Sparse and Structured Tensor Programming with Control Flow
di: Ahrens, Willow, et al.
Pubblicazione: (2024)
di: Ahrens, Willow, et al.
Pubblicazione: (2024)
PyOptInterface: Design and implementation of an efficient modeling language for mathematical optimization
di: Yang, Yue, et al.
Pubblicazione: (2024)
di: Yang, Yue, et al.
Pubblicazione: (2024)
Efficient Calculations for Inverse of $k$-diagonal Circulant Matrices and Cyclic Banded Matrices
di: Wang, Chen, et al.
Pubblicazione: (2024)
di: Wang, Chen, et al.
Pubblicazione: (2024)
Toward a comprehensive simulation framework for hypergraphs: a Python-base approach
di: Nguyen, Quoc Chuong, et al.
Pubblicazione: (2024)
di: Nguyen, Quoc Chuong, et al.
Pubblicazione: (2024)
MPAT: Modular Petri Net Assembly Toolkit
di: Chiaradonna, Stefano, et al.
Pubblicazione: (2024)
di: Chiaradonna, Stefano, et al.
Pubblicazione: (2024)
GridapTopOpt.jl: A scalable Julia toolbox for level set-based topology optimisation
di: Wegert, Zachary J., et al.
Pubblicazione: (2024)
di: Wegert, Zachary J., et al.
Pubblicazione: (2024)
Predefined Software Environment Runtimes As A Measure For Reproducibility
di: Kaushik, Aaruni
Pubblicazione: (2024)
di: Kaushik, Aaruni
Pubblicazione: (2024)
Documenti analoghi
-
Compiler-supported reduced precision and AoS-SoA transformations for heterogeneous hardware
di: Radtke, Pawel K., et al.
Pubblicazione: (2025) -
Annotation-guided AoS-to-SoA conversions and GPU offloading with data views in C++
di: Radtke, Pawel K., et al.
Pubblicazione: (2025) -
SYCL compute kernels for ExaHyPE
di: Loi, Chung Ming, et al.
Pubblicazione: (2023) -
Annotation‐Guided AoS‐to‐SoA Conversions and GPU Offloading With Data Views in C++
di: Pawel K. Radtke, et al.
Pubblicazione: (2025) -
A shared compilation stack for distributed-memory parallelism in stencil DSLs
di: Bisbas, George, et al.
Pubblicazione: (2024)