A More Scalable Sparse Dynamic Data Exchange
Fuente:
arXiv
Saved in:
| Main Authors: | Geyko, Andrew, Collom, Gerald, Schafer, Derek, Bridges, Patrick, Bienz, Amanda |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Persistent and Partitioned MPI for Stencil Communication
by: Collom, Gerald, et al.
Published: (2025)
by: Collom, Gerald, et al.
Published: (2025)
Optimizing Allreduce Operations for Modern Heterogeneous Architectures with Multiple Processes per GPU
by: Adams, Michael, et al.
Published: (2025)
by: Adams, Michael, et al.
Published: (2025)
Understanding GPU Triggering APIs for MPI+X Communication
by: Bridges, Patrick G., et al.
Published: (2024)
by: Bridges, Patrick G., et al.
Published: (2024)
Beatnik: A Novel Global Communication Mini-Application
by: Stewart, Jason A., et al.
Published: (2024)
by: Stewart, Jason A., et al.
Published: (2024)
Scaling All-to-all Operations Across Emerging Many-Core Supercomputers
by: Kinkead, Shannon, et al.
Published: (2026)
by: Kinkead, Shannon, et al.
Published: (2026)
Co-Design and Evaluation of a CPU-Free MPI GPU Communication Abstraction and Implementation
by: Bridges, Patrick G., et al.
Published: (2026)
by: Bridges, Patrick G., et al.
Published: (2026)
Scalable and Performant Data Loading
by: Hira, Moto, et al.
Published: (2025)
by: Hira, Moto, et al.
Published: (2025)
Is Sparse Matrix Reordering Effective for Sparse Matrix-Vector Multiplication?
by: Asudeh, Omid, et al.
Published: (2025)
by: Asudeh, Omid, et al.
Published: (2025)
Scalable Maxflow Processing for Dynamic Graphs
by: Kannappan, Shruthi, et al.
Published: (2025)
by: Kannappan, Shruthi, et al.
Published: (2025)
SparseServe: Unlocking Parallelism for Dynamic Sparse Attention in Long-Context LLM Serving
by: Zhou, Qihui, et al.
Published: (2025)
by: Zhou, Qihui, et al.
Published: (2025)
SDSL-Solver: Scalable Distributed Sparse Linear Solvers for Large-Scale Interior Point Methods
by: Yang, Shaofeng, et al.
Published: (2026)
by: Yang, Shaofeng, et al.
Published: (2026)
Performance Portable Monte Carlo Particle Transport on Intel, NVIDIA, and AMD GPUs
by: Tramm, John, et al.
Published: (2024)
by: Tramm, John, et al.
Published: (2024)
Hecate: Unlocking Efficient Sparse Model Training via Fully Sharded Sparse Data Parallelism
by: Qing, Yuhao, et al.
Published: (2025)
by: Qing, Yuhao, et al.
Published: (2025)
Open Science Data Federation -- operation and monitoring
by: Andrijauskas, Fabio, et al.
Published: (2026)
by: Andrijauskas, Fabio, et al.
Published: (2026)
LEISA: A Scalable Microservice-based System for Efficient Livestock Data Sharing
by: Habib, Mahir, et al.
Published: (2025)
by: Habib, Mahir, et al.
Published: (2025)
SchEdge: A Dynamic, Multi-agent, and Scalable Scheduling Simulator for IoT Edge
by: Hamedi, Ali, et al.
Published: (2025)
by: Hamedi, Ali, et al.
Published: (2025)
BigSUMO: A Scalable Framework for Big Data Traffic Analytics and Parallel Simulation
by: Sengupta, Rahul, et al.
Published: (2026)
by: Sengupta, Rahul, et al.
Published: (2026)
Radar DataTree: A FAIR and Cloud-Native Framework for Scalable Weather Radar Archives
by: Ladino-Rincon, Alfonso, et al.
Published: (2025)
by: Ladino-Rincon, Alfonso, et al.
Published: (2025)
MAGNUS: Generating Data Locality to Accelerate Sparse Matrix-Matrix Multiplication on CPUs
by: Wolfson-Pou, Jordi, et al.
Published: (2025)
by: Wolfson-Pou, Jordi, et al.
Published: (2025)
Belief Propagation Converges to Gaussian Distributions in Sparsely-Connected Factor Graphs
by: Yates, Tom, et al.
Published: (2026)
by: Yates, Tom, et al.
Published: (2026)
Leveraging Caliper and Benchpark to Analyze MPI Communication Patterns: Insights from AMG2023, Kripke, and Laghos
by: Nansamba, Grace, et al.
Published: (2025)
by: Nansamba, Grace, et al.
Published: (2025)
FedCDC: A Collaborative Framework for Data Consumers in Federated Learning Market
by: Shi, Zhuan, et al.
Published: (2025)
by: Shi, Zhuan, et al.
Published: (2025)
The Carnot Bound: Limits and Possibilities for Bandwidth-Efficient Consensus
by: Lewis-Pye, Andrew, et al.
Published: (2026)
by: Lewis-Pye, Andrew, et al.
Published: (2026)
SpComm3D: A Framework for Enabling Sparse Communication in 3D Sparse Kernels
by: Abubaker, Nabil, et al.
Published: (2024)
by: Abubaker, Nabil, et al.
Published: (2024)
SWARM+: Scalable and Resilient Multi-Agent Consensus for Fully-Decentralized Data-Aware Workload Management
by: Thareja, Komal, et al.
Published: (2026)
by: Thareja, Komal, et al.
Published: (2026)
AsyncSparse: Accelerating Sparse Matrix-Matrix Multiplication on Asynchronous GPU Architectures
by: Liu, Jie, et al.
Published: (2026)
by: Liu, Jie, et al.
Published: (2026)
ML-based Adaptive Prefetching and Data Placement for US HEP Systems
by: Karanam, Venkat Sai Suman Lamba, et al.
Published: (2025)
by: Karanam, Venkat Sai Suman Lamba, et al.
Published: (2025)
Parallel Data Object Creation: Towards Scalable Metadata Management in High-Performance I/O Library
by: Li, Youjia, et al.
Published: (2025)
by: Li, Youjia, et al.
Published: (2025)
Hetu v2: A General and Scalable Deep Learning System with Hierarchical and Heterogeneous Single Program Multiple Data Annotations
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
A Hybrid Communication Approach for Metadata Exchange in Geo-Distributed Fog Environments
by: Kruber, Marvin, et al.
Published: (2023)
by: Kruber, Marvin, et al.
Published: (2023)
Minimize Your Critical Path with Combine-and-Exchange Locks
by: König, Simon, et al.
Published: (2025)
by: König, Simon, et al.
Published: (2025)
SparseMap: Loop Mapping for Sparse CNNs on Streaming Coarse-grained Reconfigurable Array
by: Ni, Xiaobing, et al.
Published: (2024)
by: Ni, Xiaobing, et al.
Published: (2024)
Distributed-Memory Parallel Algorithms for Sparse Matrix and Sparse Tall-and-Skinny Matrix Multiplication
by: Ranawaka, Isuru, et al.
Published: (2024)
by: Ranawaka, Isuru, et al.
Published: (2024)
SIMT/GPU Data Race Verification using ISCC and Intermediary Code Representations: A Case Study
by: Osterhout, Andrew, et al.
Published: (2025)
by: Osterhout, Andrew, et al.
Published: (2025)
Federated Learning Optimization: A Comparative Study of Data and Model Exchange Strategies in Dynamic Networks
by: Luqman, Alka, et al.
Published: (2024)
by: Luqman, Alka, et al.
Published: (2024)
AB-Sparse: Sparse Attention with Adaptive Block Size for Accurate and Efficient Long-Context Inference
by: Liu, Di, et al.
Published: (2026)
by: Liu, Di, et al.
Published: (2026)
Optimality of Simultaneous Consensus with Limited Information Exchange (Extended Abstract)
by: Alpturer, Kaya, et al.
Published: (2025)
by: Alpturer, Kaya, et al.
Published: (2025)
Model Input Verification of Large Scale Simulations
by: Neykova, Rumyana, et al.
Published: (2024)
by: Neykova, Rumyana, et al.
Published: (2024)
LSKV: A Confidential Distributed Datastore to Protect Critical Data in the Cloud
by: Jeffery, Andrew, et al.
Published: (2024)
by: Jeffery, Andrew, et al.
Published: (2024)
Highly Dynamic and Fully Distributed Data Structures
by: Augustine, John, et al.
Published: (2024)
by: Augustine, John, et al.
Published: (2024)
Similar Items
-
Persistent and Partitioned MPI for Stencil Communication
by: Collom, Gerald, et al.
Published: (2025) -
Optimizing Allreduce Operations for Modern Heterogeneous Architectures with Multiple Processes per GPU
by: Adams, Michael, et al.
Published: (2025) -
Understanding GPU Triggering APIs for MPI+X Communication
by: Bridges, Patrick G., et al.
Published: (2024) -
Beatnik: A Novel Global Communication Mini-Application
by: Stewart, Jason A., et al.
Published: (2024) -
Scaling All-to-all Operations Across Emerging Many-Core Supercomputers
by: Kinkead, Shannon, et al.
Published: (2026)