To Repair or Not to Repair: Assessing Fault Resilience in MPI Stencil Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Rocco, Roberto, Boella, Elisabetta, Gregori, Daniele, Palermo, Gianluca |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing Performance and Porting Strategies for Gravitational $N$-Body Simulations on the RISC-V-Based Tenstorrent Wormhole\textsuperscript{\texttrademark}
by: Almerol, Jenny Lynn, et al.
Published: (2026)
by: Almerol, Jenny Lynn, et al.
Published: (2026)
Persistent and Partitioned MPI for Stencil Communication
by: Collom, Gerald, et al.
Published: (2025)
by: Collom, Gerald, et al.
Published: (2025)
Efficient Parameter Tuning for a Structure-Based Virtual Screening HPC Application
by: Guindani, Bruno, et al.
Published: (2024)
by: Guindani, Bruno, et al.
Published: (2024)
GPU Acceleration of Learning With Errors KEMs Using OpenACC for Post-Quantum Cryptography
by: Liberati, Tiziana, et al.
Published: (2026)
by: Liberati, Tiziana, et al.
Published: (2026)
Stencil Matrixization
by: Zhao, Wenxuan, et al.
Published: (2023)
by: Zhao, Wenxuan, et al.
Published: (2023)
The Case for ABI Interoperability in a Fault Tolerant MPI
by: Xu, Yao, et al.
Published: (2025)
by: Xu, Yao, et al.
Published: (2025)
Stencil Computations on Tenstorrent Wormhole
by: Piarulli, Lorenzo, et al.
Published: (2026)
by: Piarulli, Lorenzo, et al.
Published: (2026)
Accelerating Gravitational $N$-Body Simulations Using the RISC-V-Based Tenstorrent Wormhole
by: Almerol, Jenny Lynn, et al.
Published: (2025)
by: Almerol, Jenny Lynn, et al.
Published: (2025)
Scaling MPI Applications on Aurora
by: Ibeid, Huda, et al.
Published: (2025)
by: Ibeid, Huda, et al.
Published: (2025)
A Logic for Repair and State Recovery in Byzantine Fault-tolerant Multi-agent Systems
by: van Ditmarsch, Hans, et al.
Published: (2024)
by: van Ditmarsch, Hans, et al.
Published: (2024)
An Adaptive Distributed Stencil Abstraction for GPUs
by: Bhosale, Aditya, et al.
Published: (2025)
by: Bhosale, Aditya, et al.
Published: (2025)
Assessing the Elephant in the Room in Scheduling for Current Hybrid HPC-QC Clusters
by: Viviani, Paolo, et al.
Published: (2025)
by: Viviani, Paolo, et al.
Published: (2025)
Parallel Spawning Strategies for Dynamic-Aware MPI Applications
by: Martín-Álvarez, Iker, et al.
Published: (2025)
by: Martín-Álvarez, Iker, et al.
Published: (2025)
User Experiences with MPI RMA and ULFM in a Resilient Key-Value Store Implementation
by: Fohry, Claudia, et al.
Published: (2026)
by: Fohry, Claudia, et al.
Published: (2026)
Do We Need Tensor Cores for Stencil Computations?
by: Gu, Qiqi, et al.
Published: (2026)
by: Gu, Qiqi, et al.
Published: (2026)
Implementing True MPI Sessions and Evaluating MPI Initialization Scalability
by: Zhou, Hui, et al.
Published: (2026)
by: Zhou, Hui, et al.
Published: (2026)
Synthesizing Proxy Applications for MPI Programs
by: Luo, Jiyu, et al.
Published: (2023)
by: Luo, Jiyu, et al.
Published: (2023)
MPI Progress For All
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
Evaluation of Programming Models and Performance for Stencil Computation on Current GPU Architectures
by: Shan, Baodi, et al.
Published: (2024)
by: Shan, Baodi, et al.
Published: (2024)
A Portable Framework for Accelerating Stencil Computations on Modern Node Architectures
by: Sai, Ryuichi, et al.
Published: (2023)
by: Sai, Ryuichi, et al.
Published: (2023)
Dynamic Solutions for Hybrid Quantum-HPC Resource Allocation
by: Rocco, Roberto, et al.
Published: (2025)
by: Rocco, Roberto, et al.
Published: (2025)
MMStencil: Optimizing High-order Stencils on Multicore CPU using Matrix Unit
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
SPIDER: Unleashing Sparse Tensor Cores for Stencil Computation via Strided Swapping
by: GU, Qiqi, et al.
Published: (2025)
by: GU, Qiqi, et al.
Published: (2025)
On the performance of two-sided MPI, MPI-3 RMA and SHMEM in a Lagrangian particle cluster algorithm
by: Frey, Matthias, et al.
Published: (2024)
by: Frey, Matthias, et al.
Published: (2024)
Some New Approaches to MPI Implementations
by: Xiong, Yuqing
Published: (2024)
by: Xiong, Yuqing
Published: (2024)
Designing and Prototyping Extensions to MPI in MPICH
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
Characterization-Guided GPU Fault Resilience in NVIDIA MPS
by: Liu, Rixin, et al.
Published: (2026)
by: Liu, Rixin, et al.
Published: (2026)
Towards Exascale Computing for Astrophysical Simulation Leveraging the Leonardo EuroHPC System
by: Shukla, Nitin, et al.
Published: (2025)
by: Shukla, Nitin, et al.
Published: (2025)
Frustrated with MPI+Threads? Try MPIxThreads!
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
Concepts for designing modern C++ interfaces for MPI
by: Avans, C. Nicole, et al.
Published: (2025)
by: Avans, C. Nicole, et al.
Published: (2025)
Understanding GPU Triggering APIs for MPI+X Communication
by: Bridges, Patrick G., et al.
Published: (2024)
by: Bridges, Patrick G., et al.
Published: (2024)
Examining MPI and its Extensions for Asynchronous Multithreaded Communication
by: Yan, Jiakun, et al.
Published: (2025)
by: Yan, Jiakun, et al.
Published: (2025)
Towards the Democratization and Standardization of Dynamic Resources with MPI Spawning
by: Iserte, Sergio, et al.
Published: (2026)
by: Iserte, Sergio, et al.
Published: (2026)
Accelerating the Particle-In-Cell code ECsim with OpenACC
by: Boella, Elisabetta, et al.
Published: (2026)
by: Boella, Elisabetta, et al.
Published: (2026)
Layout-Agnostic MPI Abstraction for Distributed Computing in Modern C++
by: Klepl, Jiří, et al.
Published: (2025)
by: Klepl, Jiří, et al.
Published: (2025)
High-Performance Parallelization of Dijkstra's Algorithm Using MPI and CUDA
by: Song, Boyang
Published: (2025)
by: Song, Boyang
Published: (2025)
Cyclic Data Streaming on GPUs for Short Range Stencils Applied to Molecular Dynamics
by: Rose, Martin, et al.
Published: (2025)
by: Rose, Martin, et al.
Published: (2025)
Making Wide Stripes Practical: Cascaded Parity LRCs for Efficient Repair and High Reliability
by: Yu, Fan, et al.
Published: (2025)
by: Yu, Fan, et al.
Published: (2025)
Parallel DNA Sequence Alignment on High-Performance Systems with CUDA and MPI
by: Zwaka, Linus
Published: (2024)
by: Zwaka, Linus
Published: (2024)
KaMPIng: Flexible and (Near) Zero-Overhead C++ Bindings for MPI
by: Uhl, Tim Niklas, et al.
Published: (2024)
by: Uhl, Tim Niklas, et al.
Published: (2024)
Similar Items
-
Assessing Performance and Porting Strategies for Gravitational $N$-Body Simulations on the RISC-V-Based Tenstorrent Wormhole\textsuperscript{\texttrademark}
by: Almerol, Jenny Lynn, et al.
Published: (2026) -
Persistent and Partitioned MPI for Stencil Communication
by: Collom, Gerald, et al.
Published: (2025) -
Efficient Parameter Tuning for a Structure-Based Virtual Screening HPC Application
by: Guindani, Bruno, et al.
Published: (2024) -
GPU Acceleration of Learning With Errors KEMs Using OpenACC for Post-Quantum Cryptography
by: Liberati, Tiziana, et al.
Published: (2026) -
Stencil Matrixization
by: Zhao, Wenxuan, et al.
Published: (2023)