VSS Challenge Problem: Verifying the Correctness of AllReduce Algorithms in the MPICH Implementation of MPI
Fuente:
arXiv
Saved in:
| Main Author: | Hovland, Paul D. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Evaluation of Massively Parallel Algorithms for DFA Minimization
by: Martens, Jan, et al.
Published: (2024)
by: Martens, Jan, et al.
Published: (2024)
Enabling Practical Transparent Checkpointing for MPI: A Topological Sort Approach
by: Xu, Yao, et al.
Published: (2024)
by: Xu, Yao, et al.
Published: (2024)
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
by: Xie, Tianfang
Published: (2026)
by: Xie, Tianfang
Published: (2026)
Construction of a Byzantine Linearizable SWMR Atomic Register from SWSR Atomic Registers
by: Kshemkalyani, Ajay D., et al.
Published: (2024)
by: Kshemkalyani, Ajay D., et al.
Published: (2024)
Data Race Satisfiability on Array Elements
by: Shim, Junhyung, et al.
Published: (2025)
by: Shim, Junhyung, et al.
Published: (2025)
Categorical Message Passing Language (CaMPL) for programmers
by: Hashimoto, Daniel Kiyoshi, et al.
Published: (2026)
by: Hashimoto, Daniel Kiyoshi, et al.
Published: (2026)
Shipwright: Proving liveness of distributed systems with Byzantine participants
by: Leung, Derek, et al.
Published: (2025)
by: Leung, Derek, et al.
Published: (2025)
DNA sequence alignment: An assignment for OpenMP, MPI, and CUDA/OpenCL
by: Gonzalez-Escribano, Arturo, et al.
Published: (2024)
by: Gonzalez-Escribano, Arturo, et al.
Published: (2024)
Verifying In-Network Computing Systems for Design Risks
by: Bai, Tianyu, et al.
Published: (2026)
by: Bai, Tianyu, et al.
Published: (2026)
Challenging Portability Paradigms: FPGA Acceleration Using SYCL and OpenCL
by: de Castro, Manuel, et al.
Published: (2024)
by: de Castro, Manuel, et al.
Published: (2024)
Optimizing Fine-Grained Parallelism Through Dynamic Load Balancing on Multi-Socket Many-Core Systems
by: Wang, Wenyi, et al.
Published: (2025)
by: Wang, Wenyi, et al.
Published: (2025)
AutoTSMM: An Auto-tuning Framework for Building High-Performance Tall-and-Skinny Matrix-Matrix Multiplication on CPUs
by: Li, Chendi, et al.
Published: (2022)
by: Li, Chendi, et al.
Published: (2022)
Design and Implementation of an Analysis Pipeline for Heterogeneous Data
by: Sarker, Arup Kumar, et al.
Published: (2024)
by: Sarker, Arup Kumar, et al.
Published: (2024)
Designing and Prototyping Extensions to MPI in MPICH
by: Zhou, Hui, et al.
Published: (2024)
by: Zhou, Hui, et al.
Published: (2024)
DDS: DPU-optimized Disaggregated Storage [Extended Report]
by: Zhang, Qizhen, et al.
Published: (2024)
by: Zhang, Qizhen, et al.
Published: (2024)
DPDPU: Data Processing with DPUs
by: Hu, Jiasheng, et al.
Published: (2024)
by: Hu, Jiasheng, et al.
Published: (2024)
Alea-BFT: Practical Asynchronous Byzantine Fault Tolerance
by: Antunes, Diogo S., et al.
Published: (2024)
by: Antunes, Diogo S., et al.
Published: (2024)
Laminar: A Probe-First Scheduling Paradigm with Deterministic Runtime Survival
by: Chu, Zhengyan
Published: (2026)
by: Chu, Zhengyan
Published: (2026)
Characterizing and Fixing Silent Data Loss in Spark-on-AWS-Lambda with Open Table Formats
by: Gandla, Srujan Kumar
Published: (2026)
by: Gandla, Srujan Kumar
Published: (2026)
Trident: Adaptive Scheduling for Heterogeneous Multimodal Data Pipelines
by: Pan, Ding, et al.
Published: (2026)
by: Pan, Ding, et al.
Published: (2026)
VeriFx: Correct Replicated Data Types for the Masses
by: De Porre, Kevin, et al.
Published: (2022)
by: De Porre, Kevin, et al.
Published: (2022)
Scheduler-Driven Job Atomization
by: Konopa, Michal, et al.
Published: (2025)
by: Konopa, Michal, et al.
Published: (2025)
JASDA: Introducing Job-Aware Scheduling in Scheduler-Driven Job Atomization
by: Konopa, Michal, et al.
Published: (2025)
by: Konopa, Michal, et al.
Published: (2025)
A C++17 Thread Pool for High-Performance Scientific Computing
by: Shoshany, Barak
Published: (2021)
by: Shoshany, Barak
Published: (2021)
Stream parallel skeleton optimization
by: Aldinucci, Marco, et al.
Published: (2024)
by: Aldinucci, Marco, et al.
Published: (2024)
StreamFlow: cross-breeding cloud with HPC
by: Colonnelli, Iacopo, et al.
Published: (2020)
by: Colonnelli, Iacopo, et al.
Published: (2020)
SpaDA: A Spatial Dataflow Architecture Programming Language
by: Gianinazzi, Lukas, et al.
Published: (2025)
by: Gianinazzi, Lukas, et al.
Published: (2025)
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
by: Homola, Jakub, et al.
Published: (2025)
by: Homola, Jakub, et al.
Published: (2025)
A Sufficient Epistemic Condition for Solving Stabilizing Agreement
by: Cignarale, Giorgio, et al.
Published: (2024)
by: Cignarale, Giorgio, et al.
Published: (2024)
Technical Report: Exploring Automatic Model-Checking of the Ethereum specification
by: Konnov, Igor, et al.
Published: (2025)
by: Konnov, Igor, et al.
Published: (2025)
Combining Serverless and High-Performance Computing Paradigms to support ML Data-Intensive Applications
by: Staylor, Mills, et al.
Published: (2025)
by: Staylor, Mills, et al.
Published: (2025)
Deep RC: A Scalable Data Engineering and Deep Learning Pipeline
by: Sarker, Arup Kumar, et al.
Published: (2025)
by: Sarker, Arup Kumar, et al.
Published: (2025)
Sharded Elimination and Combining for Highly-Efficient Concurrent Stacks
by: Singh, Ajay, et al.
Published: (2026)
by: Singh, Ajay, et al.
Published: (2026)
A Treasure Trove of Performance: Analyzing the IO500 Submission Data
by: Kunkel, Julian, et al.
Published: (2026)
by: Kunkel, Julian, et al.
Published: (2026)
Revisiting the Time Cost Model of AllReduce
by: Xiong, Dian, et al.
Published: (2024)
by: Xiong, Dian, et al.
Published: (2024)
push0: Scalable and Fault-Tolerant Orchestration for Zero-Knowledge Proof Generation
by: Ahmadvand, Mohsen, et al.
Published: (2026)
by: Ahmadvand, Mohsen, et al.
Published: (2026)
Privacy-Aware Split Inference with Speculative Decoding for Large Language Models over Wide-Area Networks
by: Cunningham, Michael
Published: (2026)
by: Cunningham, Michael
Published: (2026)
Advocate -- Trustworthy Evidence in Cloud Systems
by: Werner, Sebastian, et al.
Published: (2024)
by: Werner, Sebastian, et al.
Published: (2024)
Operational Memory Architecture for Kubernetes:Preserving Causal Context Across the Evidence Horizon
by: Khan, Shamsher
Published: (2026)
by: Khan, Shamsher
Published: (2026)
PoCL-R: An Open Standard Based Offloading Layer for Heterogeneous Multi-Access Edge Computing with Server Side Scalability
by: Solanti, Jan, et al.
Published: (2023)
by: Solanti, Jan, et al.
Published: (2023)
Similar Items
-
An Evaluation of Massively Parallel Algorithms for DFA Minimization
by: Martens, Jan, et al.
Published: (2024) -
Enabling Practical Transparent Checkpointing for MPI: A Topological Sort Approach
by: Xu, Yao, et al.
Published: (2024) -
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
by: Xie, Tianfang
Published: (2026) -
Construction of a Byzantine Linearizable SWMR Atomic Register from SWSR Atomic Registers
by: Kshemkalyani, Ajay D., et al.
Published: (2024) -
Data Race Satisfiability on Array Elements
by: Shim, Junhyung, et al.
Published: (2025)