Examining MPI and its Extensions for Asynchronous Multithreaded Communication
Fuente:
arXiv
Saved in:
| Main Authors: | Yan, Jiakun, Snir, Marc, Guo, Yanfei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LCI: a Lightweight Communication Interface for Efficient Asynchronous Multithreaded Communication
by: Yan, Jiakun, et al.
Published: (2025)
by: Yan, Jiakun, et al.
Published: (2025)
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
by: Cheng, Minyu, et al.
Published: (2026)
by: Cheng, Minyu, et al.
Published: (2026)
Contemplating a Lightweight Communication Interface for Asynchronous Many-Task Systems
by: Yan, Jiakun, et al.
Published: (2025)
by: Yan, Jiakun, et al.
Published: (2025)
Understanding the Communication Needs of Asynchronous Many-Task Systems -- A Case Study of HPX+LCI
by: Yan, Jiakun, et al.
Published: (2025)
by: Yan, Jiakun, et al.
Published: (2025)
Designing and Prototyping Extensions to MPI in MPICH
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
Implementing True MPI Sessions and Evaluating MPI Initialization Scalability
by: Zhou, Hui, et al.
Published: (2026)
by: Zhou, Hui, et al.
Published: (2026)
MPI Progress For All
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
Frustrated with MPI+Threads? Try MPIxThreads!
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
Radiation Hydrodynamics at Scale: Comparing MPI and Asynchronous Many-Task Runtimes with FleCSI
by: Strack, Alexander, et al.
Published: (2026)
by: Strack, Alexander, et al.
Published: (2026)
Exploring Fine-grained Task Parallelism on Simultaneous Multithreading Cores
by: Los, Denis, et al.
Published: (2024)
by: Los, Denis, et al.
Published: (2024)
Persistent and Partitioned MPI for Stencil Communication
by: Collom, Gerald, et al.
Published: (2025)
by: Collom, Gerald, et al.
Published: (2025)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
by: Fan, Jiakun, et al.
Published: (2025)
by: Fan, Jiakun, et al.
Published: (2025)
Dynamic Simultaneous Multithreaded Architecture
by: Ortiz-Arroyo, Daniel, et al.
Published: (2024)
by: Ortiz-Arroyo, Daniel, et al.
Published: (2024)
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
by: Korndörfer, Jonas H. Müller, et al.
Published: (2021)
by: Korndörfer, Jonas H. Müller, et al.
Published: (2021)
An Optimized Error-controlled MPI Collective Framework Integrated with Lossy Compression
by: Huang, Jiajun, et al.
Published: (2023)
by: Huang, Jiajun, et al.
Published: (2023)
Understanding GPU Triggering APIs for MPI+X Communication
by: Bridges, Patrick G., et al.
Published: (2024)
by: Bridges, Patrick G., et al.
Published: (2024)
Multithreaded parallelism for heterogeneous clusters of QPUs
by: Seitz, Philipp, et al.
Published: (2023)
by: Seitz, Philipp, et al.
Published: (2023)
MPI-over-CXL: Enhancing Communication Efficiency in Distributed HPC Systems
by: Kwon, Miryeong, et al.
Published: (2025)
by: Kwon, Miryeong, et al.
Published: (2025)
Analyzing Persistent Alltoallv RMA Implementations for High-Performance MPI Communication
by: Namugwanya, Evelyn
Published: (2026)
by: Namugwanya, Evelyn
Published: (2026)
Communication Round and Computation Efficient Exclusive Prefix-Sums Algorithms (for MPI_Exscan)
by: Träff, Jesper Larsson
Published: (2025)
by: Träff, Jesper Larsson
Published: (2025)
Scaling MPI Applications on Aurora
by: Ibeid, Huda, et al.
Published: (2025)
by: Ibeid, Huda, et al.
Published: (2025)
Co-Design and Evaluation of a CPU-Free MPI GPU Communication Abstraction and Implementation
by: Bridges, Patrick G., et al.
Published: (2026)
by: Bridges, Patrick G., et al.
Published: (2026)
On the performance of two-sided MPI, MPI-3 RMA and SHMEM in a Lagrangian particle cluster algorithm
by: Frey, Matthias, et al.
Published: (2024)
by: Frey, Matthias, et al.
Published: (2024)
Some New Approaches to MPI Implementations
by: Xiong, Yuqing
Published: (2024)
by: Xiong, Yuqing
Published: (2024)
Synthesizing Proxy Applications for MPI Programs
by: Luo, Jiyu, et al.
Published: (2023)
by: Luo, Jiyu, et al.
Published: (2023)
Asynchronous-Many-Task Systems: Challenges and Opportunities -- Scaling an AMR Astrophysics Code on Exascale machines using Kokkos and HPX
by: Daiß, Gregor, et al.
Published: (2024)
by: Daiß, Gregor, et al.
Published: (2024)
Concepts for designing modern C++ interfaces for MPI
by: Avans, C. Nicole, et al.
Published: (2025)
by: Avans, C. Nicole, et al.
Published: (2025)
Leveraging Caliper and Benchpark to Analyze MPI Communication Patterns: Insights from AMG2023, Kripke, and Laghos
by: Nansamba, Grace, et al.
Published: (2025)
by: Nansamba, Grace, et al.
Published: (2025)
MPI-Q: A Message Communication Library for Large-Scale Classical-Quantum Heterogeneous Hybrid Distributed Computing
by: Wang, Feng, et al.
Published: (2026)
by: Wang, Feng, et al.
Published: (2026)
Formal Definitions and Performance Comparison of Consistency Models for Parallel File Systems
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
The Case for ABI Interoperability in a Fault Tolerant MPI
by: Xu, Yao, et al.
Published: (2025)
by: Xu, Yao, et al.
Published: (2025)
Parallel Spawning Strategies for Dynamic-Aware MPI Applications
by: Martín-Álvarez, Iker, et al.
Published: (2025)
by: Martín-Álvarez, Iker, et al.
Published: (2025)
Towards the Democratization and Standardization of Dynamic Resources with MPI Spawning
by: Iserte, Sergio, et al.
Published: (2026)
by: Iserte, Sergio, et al.
Published: (2026)
Do MPI Derived Datatypes Actually Help? A Single-Node Cross-Implementation Study on Shared-Memory Communication
by: Adefemi, Temitayo
Published: (2025)
by: Adefemi, Temitayo
Published: (2025)
Layout-Agnostic MPI Abstraction for Distributed Computing in Modern C++
by: Klepl, Jiří, et al.
Published: (2025)
by: Klepl, Jiří, et al.
Published: (2025)
High-Performance Parallelization of Dijkstra's Algorithm Using MPI and CUDA
by: Song, Boyang
Published: (2025)
by: Song, Boyang
Published: (2025)
To Repair or Not to Repair: Assessing Fault Resilience in MPI Stencil Applications
by: Rocco, Roberto, et al.
Published: (2024)
by: Rocco, Roberto, et al.
Published: (2024)
Exploring the Efficiency of Renewable Energy-based Modular Data Centers at Scale
by: Sun, Jinghan, et al.
Published: (2024)
by: Sun, Jinghan, et al.
Published: (2024)
Resource Optimization with MPI Process Malleability for Dynamic Workloads in HPC Clusters
by: Iserte, Sergio, et al.
Published: (2025)
by: Iserte, Sergio, et al.
Published: (2025)
Performance of a high-order MPI-Kokkos accelerated fluid solver
by: Sporykhin, Filipp, et al.
Published: (2025)
by: Sporykhin, Filipp, et al.
Published: (2025)
Similar Items
-
LCI: a Lightweight Communication Interface for Efficient Asynchronous Multithreaded Communication
by: Yan, Jiakun, et al.
Published: (2025) -
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
by: Cheng, Minyu, et al.
Published: (2026) -
Contemplating a Lightweight Communication Interface for Asynchronous Many-Task Systems
by: Yan, Jiakun, et al.
Published: (2025) -
Understanding the Communication Needs of Asynchronous Many-Task Systems -- A Case Study of HPX+LCI
by: Yan, Jiakun, et al.
Published: (2025) -
Designing and Prototyping Extensions to MPI in MPICH
by: Zhou, Hui, et al.
Published: (2024)