Comparing Cross-Platform Performance via Node-to-Node Scaling Studies
Fuente:
arXiv
Salvato in:
| Autori principali: | Weiss, Kenneth, Stitt, Thomas M., Hawkins, Daryl, Pearce, Olga, Brink, Stephanie, Rieben, Robert N. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On Similarity of Computational Kernels in our Codes and Proxies
di: McKinsey, Michael, et al.
Pubblicazione: (2026)
di: McKinsey, Michael, et al.
Pubblicazione: (2026)
Understanding Power and Energy Utilization in Large Scale Production Physics Simulation Codes
di: Bertsch, Adam, et al.
Pubblicazione: (2022)
di: Bertsch, Adam, et al.
Pubblicazione: (2022)
Automating Multi-Tenancy Performance Evaluation on Edge Compute Nodes
di: Georgiou, Joanna, et al.
Pubblicazione: (2025)
di: Georgiou, Joanna, et al.
Pubblicazione: (2025)
Do MPI Derived Datatypes Actually Help? A Single-Node Cross-Implementation Study on Shared-Memory Communication
di: Adefemi, Temitayo
Pubblicazione: (2025)
di: Adefemi, Temitayo
Pubblicazione: (2025)
Recognizing Hereditary Properties in the Presence of Byzantine Nodes
di: Cifuentes-Núñez, David, et al.
Pubblicazione: (2023)
di: Cifuentes-Núñez, David, et al.
Pubblicazione: (2023)
Self-healing Nodes with Adaptive Data-Sharding
di: Thakur, Ayush, et al.
Pubblicazione: (2024)
di: Thakur, Ayush, et al.
Pubblicazione: (2024)
Eliminating Hidden Serialization in Multi-Node Megakernel Communication
di: Oh, Byungsoo, et al.
Pubblicazione: (2026)
di: Oh, Byungsoo, et al.
Pubblicazione: (2026)
Completing the Node-Averaged Complexity Landscape of LCLs on Trees
di: Balliu, Alkida, et al.
Pubblicazione: (2024)
di: Balliu, Alkida, et al.
Pubblicazione: (2024)
Balancing Fixed Number of Nodes Among Multiple Fixed Clusters
di: Ranjan, Paritosh, et al.
Pubblicazione: (2025)
di: Ranjan, Paritosh, et al.
Pubblicazione: (2025)
Load Balanced Parallel Node Generation for Meshless Numerical Methods
di: Vehovar, Jon, et al.
Pubblicazione: (2026)
di: Vehovar, Jon, et al.
Pubblicazione: (2026)
Simulations between Strongly Sublinear MPC and Node-Capacitated Clique
di: Schneider, Philipp, et al.
Pubblicazione: (2025)
di: Schneider, Philipp, et al.
Pubblicazione: (2025)
Learning Process Energy Profiles from Node-Level Power Data
di: Bader, Jonathan, et al.
Pubblicazione: (2025)
di: Bader, Jonathan, et al.
Pubblicazione: (2025)
A Portable Framework for Accelerating Stencil Computations on Modern Node Architectures
di: Sai, Ryuichi, et al.
Pubblicazione: (2023)
di: Sai, Ryuichi, et al.
Pubblicazione: (2023)
MalleTrain: Deep Neural Network Training on Unfillable Supercomputer Nodes
di: Ma, Xiaolong, et al.
Pubblicazione: (2024)
di: Ma, Xiaolong, et al.
Pubblicazione: (2024)
CHIRON: Accelerating Node Synchronization without Security Trade-offs in Distributed Ledgers
di: Neiheiser, Ray, et al.
Pubblicazione: (2024)
di: Neiheiser, Ray, et al.
Pubblicazione: (2024)
Generic and ML Workloads in an HPC Datacenter: Node Energy, Job Failures, and Node-Job Analysis
di: Chu, Xiaoyu, et al.
Pubblicazione: (2024)
di: Chu, Xiaoyu, et al.
Pubblicazione: (2024)
BlockRaFT: A Distributed Framework for Fault-Tolerant and Scalable Blockchain Nodes
di: Piduguralla, Manaswini, et al.
Pubblicazione: (2026)
di: Piduguralla, Manaswini, et al.
Pubblicazione: (2026)
Byzantine Fault Tolerant Protocols with Near-Constant Work per Node without Signatures
di: Schneider, Philipp
Pubblicazione: (2025)
di: Schneider, Philipp
Pubblicazione: (2025)
A Scalable State Sharing Protocol for Low-Resource Validator Nodes in Blockchain Networks
di: Hias, Ruben, et al.
Pubblicazione: (2024)
di: Hias, Ruben, et al.
Pubblicazione: (2024)
Advancing Blockchain Scalability: A Linear Optimization Framework for Diversified Node Allocation in Shards
di: Assmann, Björn, et al.
Pubblicazione: (2024)
di: Assmann, Björn, et al.
Pubblicazione: (2024)
Accelerating Intra-Node GPU-to-GPU Communication Through Multi-Path Transfers with CUDA Graphs
di: Sojoodi, Amirhossein, et al.
Pubblicazione: (2026)
di: Sojoodi, Amirhossein, et al.
Pubblicazione: (2026)
SpotVista: Availability-Aware Recommendation System for Reliable and Cost-Efficient Multi-Node Spot Instances
di: Kim, Taeyoon, et al.
Pubblicazione: (2026)
di: Kim, Taeyoon, et al.
Pubblicazione: (2026)
Exploring Distributed Vector Databases Performance on HPC Platforms: A Study with Qdrant
di: Ockerman, Seth, et al.
Pubblicazione: (2025)
di: Ockerman, Seth, et al.
Pubblicazione: (2025)
Persistent and Partitioned MPI for Stencil Communication
di: Collom, Gerald, et al.
Pubblicazione: (2025)
di: Collom, Gerald, et al.
Pubblicazione: (2025)
CCCL: Node-Spanning GPU Collectives with CXL Memory Pooling
di: Xu, Dong, et al.
Pubblicazione: (2026)
di: Xu, Dong, et al.
Pubblicazione: (2026)
Deal: Distributed End-to-End GNN Inference for All Nodes
di: Chen, Shiyang, et al.
Pubblicazione: (2025)
di: Chen, Shiyang, et al.
Pubblicazione: (2025)
Node Compass: Multilevel Tracing and Debugging of Request Executions in JavaScript-Based Web-Servers
di: Kabamba, Herve Mbikayi, et al.
Pubblicazione: (2023)
di: Kabamba, Herve Mbikayi, et al.
Pubblicazione: (2023)
BCL: A Cross-Platform Distributed Container Library
di: Brock, Benjamin, et al.
Pubblicazione: (2018)
di: Brock, Benjamin, et al.
Pubblicazione: (2018)
CIR: Lightweight Container Image for Cross-Platform Deployment
di: Li, Fengzhi, et al.
Pubblicazione: (2026)
di: Li, Fengzhi, et al.
Pubblicazione: (2026)
RAFI -- A Ray/Work Forwarding Infrastructure for Data Parallel Multi-Node/Multi-GPU Computing
di: Wald, Ingo, et al.
Pubblicazione: (2026)
di: Wald, Ingo, et al.
Pubblicazione: (2026)
Case Study: Performance Analysis of a Virtualized XRootD Frontend in Large-Scale WAN Transfers
di: da Silva, J M, et al.
Pubblicazione: (2026)
di: da Silva, J M, et al.
Pubblicazione: (2026)
Large-Scale Metric Computation in Online Controlled Experiment Platform
di: Xiong, Tao, et al.
Pubblicazione: (2024)
di: Xiong, Tao, et al.
Pubblicazione: (2024)
Predictive Autoscaling for Node.js on Kubernetes: Lower Latency, Right-Sized Capacity
di: Tymoshenko, Ivan, et al.
Pubblicazione: (2026)
di: Tymoshenko, Ivan, et al.
Pubblicazione: (2026)
Enabling Time-Aware Priority Traffic Management over Distributed FPGA Nodes
di: Scionti, Alberto, et al.
Pubblicazione: (2025)
di: Scionti, Alberto, et al.
Pubblicazione: (2025)
Optimizing Sensor Node Localization for Achieving Sustainable Smart Agriculture System Connectivity
di: Naeem, Mohamed
Pubblicazione: (2025)
di: Naeem, Mohamed
Pubblicazione: (2025)
Virtual Nodes Can Help: Tackling Distribution Shifts in Federated Graph Learning
di: Fu, Xingbo, et al.
Pubblicazione: (2024)
di: Fu, Xingbo, et al.
Pubblicazione: (2024)
Guard: Scalable Straggler Detection and Node Health Management for Large-Scale Training
di: Liu, Guanliang, et al.
Pubblicazione: (2026)
di: Liu, Guanliang, et al.
Pubblicazione: (2026)
Driving Computational Efficiency in Large-Scale Platforms using HPC Technologies
di: Mendez, Alexander Martinez, et al.
Pubblicazione: (2026)
di: Mendez, Alexander Martinez, et al.
Pubblicazione: (2026)
The HEAL Data Platform
di: Larrick, Brienna M., et al.
Pubblicazione: (2025)
di: Larrick, Brienna M., et al.
Pubblicazione: (2025)
AntDT: A Self-Adaptive Distributed Training Framework for Leader and Straggler Nodes
di: Xiao, Youshao, et al.
Pubblicazione: (2024)
di: Xiao, Youshao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
On Similarity of Computational Kernels in our Codes and Proxies
di: McKinsey, Michael, et al.
Pubblicazione: (2026) -
Understanding Power and Energy Utilization in Large Scale Production Physics Simulation Codes
di: Bertsch, Adam, et al.
Pubblicazione: (2022) -
Automating Multi-Tenancy Performance Evaluation on Edge Compute Nodes
di: Georgiou, Joanna, et al.
Pubblicazione: (2025) -
Do MPI Derived Datatypes Actually Help? A Single-Node Cross-Implementation Study on Shared-Memory Communication
di: Adefemi, Temitayo
Pubblicazione: (2025) -
Recognizing Hereditary Properties in the Presence of Byzantine Nodes
di: Cifuentes-Núñez, David, et al.
Pubblicazione: (2023)