Exploring Performance-Productivity Trade-offs in AMT Runtimes: A Task Bench Study of Itoyori, ItoyoriFBC, HPX, and MPI
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lahnor, Torben R., Reitz, Mia, Posner, Jonas, Diehl, Patrick |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Closing a Source Complexity Gap between Chapel and HPX
par: Atre, Shreyas, et autres
Publié: (2025)
par: Atre, Shreyas, et autres
Publié: (2025)
HPX -- An open source C++ Standard Library for Parallelism and Concurrency
par: Heller, Thomas, et autres
Publié: (2023)
par: Heller, Thomas, et autres
Publié: (2023)
HPX with Spack and Singularity Containers: Evaluating Overheads for HPX/Kokkos using an astrophysics application
par: Diehl, Patrick, et autres
Publié: (2024)
par: Diehl, Patrick, et autres
Publié: (2024)
Preparing for HPC on RISC-V: Examining Vectorization and Distributed Performance of an Astrophyiscs Application with HPX and Kokkos
par: Diehl, Patrick, et autres
Publié: (2024)
par: Diehl, Patrick, et autres
Publié: (2024)
A New Execution Model and Executor for Adaptively Optimizing the Performance of Parallel Algorithms Using HPX Runtime System
par: Mohammadiporshokooh, Karame, et autres
Publié: (2025)
par: Mohammadiporshokooh, Karame, et autres
Publié: (2025)
Overcoming Latency-bound Limitations of Distributed Graph Algorithms using the HPX Runtime System
par: Mohammadiporshokooh, Karame, et autres
Publié: (2026)
par: Mohammadiporshokooh, Karame, et autres
Publié: (2026)
Radiation Hydrodynamics at Scale: Comparing MPI and Asynchronous Many-Task Runtimes with FleCSI
par: Strack, Alexander, et autres
Publié: (2026)
par: Strack, Alexander, et autres
Publié: (2026)
Asynchronous-Many-Task Systems: Challenges and Opportunities -- Scaling an AMR Astrophysics Code on Exascale machines using Kokkos and HPX
par: Daiß, Gregor, et autres
Publié: (2024)
par: Daiß, Gregor, et autres
Publié: (2024)
Benchmarking the Parallel 1D Heat Equation Solver in Chapel, Charm++, C++, HPX, Go, Julia, Python, Rust, Swift, and Java
par: Diehl, Patrick, et autres
Publié: (2023)
par: Diehl, Patrick, et autres
Publié: (2023)
Evaluating Malleable Job Scheduling in HPC Clusters using Real-World Workloads
par: Zojer, Patrick, et autres
Publié: (2026)
par: Zojer, Patrick, et autres
Publié: (2026)
GPU-Resident Gaussian Process Regression Leveraging Asynchronous Tasks with HPX
par: Möllmann, Henrik, et autres
Publié: (2026)
par: Möllmann, Henrik, et autres
Publié: (2026)
Parallel FFTW on RISC-V: A Comparative Study including OpenMP, MPI, and HPX
par: Strack, Alexander, et autres
Publié: (2025)
par: Strack, Alexander, et autres
Publié: (2025)
Understanding the Communication Needs of Asynchronous Many-Task Systems -- A Case Study of HPX+LCI
par: Yan, Jiakun, et autres
Publié: (2025)
par: Yan, Jiakun, et autres
Publié: (2025)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
par: Diehl, Patrick, et autres
Publié: (2025)
par: Diehl, Patrick, et autres
Publié: (2025)
A HPX Communication Benchmark: Distributed FFT using Collectives
par: Strack, Alexander, et autres
Publié: (2025)
par: Strack, Alexander, et autres
Publié: (2025)
An Initial Evaluation of Distributed Graph Algorithms using NWGraph and HPX
par: Mohammadiporshokooh, Karame, et autres
Publié: (2026)
par: Mohammadiporshokooh, Karame, et autres
Publié: (2026)
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks
par: Nader, Noujoud, et autres
Publié: (2025)
par: Nader, Noujoud, et autres
Publié: (2025)
Implementing True MPI Sessions and Evaluating MPI Initialization Scalability
par: Zhou, Hui, et autres
Publié: (2026)
par: Zhou, Hui, et autres
Publié: (2026)
Experiences Porting Distributed Applications to Asynchronous Tasks: A Multidimensional FFT Case-study
par: Strack, Alexander, et autres
Publié: (2024)
par: Strack, Alexander, et autres
Publié: (2024)
High-Performance Parallelization of Dijkstra's Algorithm Using MPI and CUDA
par: Song, Boyang
Publié: (2025)
par: Song, Boyang
Publié: (2025)
Understanding GPU Triggering APIs for MPI+X Communication
par: Bridges, Patrick G., et autres
Publié: (2024)
par: Bridges, Patrick G., et autres
Publié: (2024)
Automatic Tracing in Task-Based Runtime Systems
par: Yadav, Rohan, et autres
Publié: (2024)
par: Yadav, Rohan, et autres
Publié: (2024)
MPI Progress For All
par: Zhou, Hui, et autres
Publié: (2024)
par: Zhou, Hui, et autres
Publié: (2024)
Parallel DNA Sequence Alignment on High-Performance Systems with CUDA and MPI
par: Zwaka, Linus
Publié: (2024)
par: Zwaka, Linus
Publié: (2024)
Performance of a high-order MPI-Kokkos accelerated fluid solver
par: Sporykhin, Filipp, et autres
Publié: (2025)
par: Sporykhin, Filipp, et autres
Publié: (2025)
Analyzing Persistent Alltoallv RMA Implementations for High-Performance MPI Communication
par: Namugwanya, Evelyn
Publié: (2026)
par: Namugwanya, Evelyn
Publié: (2026)
Scaling MPI Applications on Aurora
par: Ibeid, Huda, et autres
Publié: (2025)
par: Ibeid, Huda, et autres
Publié: (2025)
On the performance of two-sided MPI, MPI-3 RMA and SHMEM in a Lagrangian particle cluster algorithm
par: Frey, Matthias, et autres
Publié: (2024)
par: Frey, Matthias, et autres
Publié: (2024)
Performance Trade-offs of High Order Meshless Approximation on Distributed Memory Systems
par: Vehovar, Jon, et autres
Publié: (2025)
par: Vehovar, Jon, et autres
Publié: (2025)
Persistent and Partitioned MPI for Stencil Communication
par: Collom, Gerald, et autres
Publié: (2025)
par: Collom, Gerald, et autres
Publié: (2025)
Some New Approaches to MPI Implementations
par: Xiong, Yuqing
Publié: (2024)
par: Xiong, Yuqing
Publié: (2024)
Designing and Prototyping Extensions to MPI in MPICH
par: Zhou, Hui, et autres
Publié: (2024)
par: Zhou, Hui, et autres
Publié: (2024)
INSPIRIT: Optimizing Heterogeneous Task Scheduling through Adaptive Priority in Task-based Runtime Systems
par: Wang, Yiqing, et autres
Publié: (2024)
par: Wang, Yiqing, et autres
Publié: (2024)
Concepts for designing modern C++ interfaces for MPI
par: Avans, C. Nicole, et autres
Publié: (2025)
par: Avans, C. Nicole, et autres
Publié: (2025)
Frustrated with MPI+Threads? Try MPIxThreads!
par: Zhou, Hui, et autres
Publié: (2024)
par: Zhou, Hui, et autres
Publié: (2024)
Integrating and Characterizing HPC Task Runtime Systems for hybrid AI-HPC workloads
par: Merzky, Andre, et autres
Publié: (2025)
par: Merzky, Andre, et autres
Publié: (2025)
Is RISC-V Ready for Machine Learning? Portable Gaussian Processes Using Asynchronous Tasks
par: Strack, Alexander, et autres
Publié: (2026)
par: Strack, Alexander, et autres
Publié: (2026)
Co-Design and Evaluation of a CPU-Free MPI GPU Communication Abstraction and Implementation
par: Bridges, Patrick G., et autres
Publié: (2026)
par: Bridges, Patrick G., et autres
Publié: (2026)
Examining MPI and its Extensions for Asynchronous Multithreaded Communication
par: Yan, Jiakun, et autres
Publié: (2025)
par: Yan, Jiakun, et autres
Publié: (2025)
The Case for ABI Interoperability in a Fault Tolerant MPI
par: Xu, Yao, et autres
Publié: (2025)
par: Xu, Yao, et autres
Publié: (2025)
Documents similaires
-
Closing a Source Complexity Gap between Chapel and HPX
par: Atre, Shreyas, et autres
Publié: (2025) -
HPX -- An open source C++ Standard Library for Parallelism and Concurrency
par: Heller, Thomas, et autres
Publié: (2023) -
HPX with Spack and Singularity Containers: Evaluating Overheads for HPX/Kokkos using an astrophysics application
par: Diehl, Patrick, et autres
Publié: (2024) -
Preparing for HPC on RISC-V: Examining Vectorization and Distributed Performance of an Astrophyiscs Application with HPX and Kokkos
par: Diehl, Patrick, et autres
Publié: (2024) -
A New Execution Model and Executor for Adaptively Optimizing the Performance of Parallel Algorithms Using HPX Runtime System
par: Mohammadiporshokooh, Karame, et autres
Publié: (2025)