ParaLog: Consistent Host-side Logging for Parallel Checkpoints
Fuente:
arXiv
Salvato in:
| Autori principali: | Chien, Steven W. D., Sato, Kento, Podobas, Artur, Jansson, Niclas, Markidis, Stefano, Honda, Michio |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
di: Andersson, Måns I., et al.
Pubblicazione: (2025)
di: Andersson, Måns I., et al.
Pubblicazione: (2025)
Understanding Power Consumption Metric on Heterogeneous Memory Systems
di: Proaño, Andrès Rubio, et al.
Pubblicazione: (2024)
di: Proaño, Andrès Rubio, et al.
Pubblicazione: (2024)
Extracting Practical, Actionable Energy Insights from Supercomputer Telemetry and Logs
di: Cornelius, Melanie, et al.
Pubblicazione: (2025)
di: Cornelius, Melanie, et al.
Pubblicazione: (2025)
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
Characterizing the Performance of the Implicit Massively Parallel Particle-in-Cell iPIC3D Code
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
DIAL: Decentralized I/O AutoTuning via Learned Client-side Local Metrics for Parallel File System
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers
di: Maurya, Avinash, et al.
Pubblicazione: (2026)
di: Maurya, Avinash, et al.
Pubblicazione: (2026)
On Orchestrating Parallel Broadcasts for Distributed Ledgers
di: Sheng, Peiyao, et al.
Pubblicazione: (2024)
di: Sheng, Peiyao, et al.
Pubblicazione: (2024)
Automated Programmatic Performance Analysis of Parallel Programs
di: Cankur, Onur, et al.
Pubblicazione: (2024)
di: Cankur, Onur, et al.
Pubblicazione: (2024)
An Auto-tuning Method for Run-time Data Transformation for Sparse Matrix-Vector Multiplication
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
di: Katagiri, Takahiro, et al.
Pubblicazione: (2024)
Optimal Parallel Scheduling under Concave Speedup Functions
di: Li, Chengzhang, et al.
Pubblicazione: (2025)
di: Li, Chengzhang, et al.
Pubblicazione: (2025)
Recorder: Comprehensive Parallel I/O Tracing and Analysis
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
di: Lacey, Dane C., et al.
Pubblicazione: (2024)
di: Lacey, Dane C., et al.
Pubblicazione: (2024)
Fine-Grained Energy Prediction For Parallellized LLM Inference With PIE-P
di: Dutt, Anurag, et al.
Pubblicazione: (2025)
di: Dutt, Anurag, et al.
Pubblicazione: (2025)
Automated Calibration of Parallel and Distributed Computing Simulators: A Case Study
di: McDonald, Jesse, et al.
Pubblicazione: (2024)
di: McDonald, Jesse, et al.
Pubblicazione: (2024)
What is Quantum Parallelism, Anyhow?
di: Markidis, Stefano
Pubblicazione: (2024)
di: Markidis, Stefano
Pubblicazione: (2024)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
di: Putra, Jody Almaida
Pubblicazione: (2026)
di: Putra, Jody Almaida
Pubblicazione: (2026)
Matryoshka: Optimization of Dynamic Diverse Quantum Chemistry Systems via Elastic Parallelism Transformation
di: Wang, Tuowei, et al.
Pubblicazione: (2024)
di: Wang, Tuowei, et al.
Pubblicazione: (2024)
CARAT: Client-Side Adaptive RPC and Cache Co-Tuning for Parallel File Systems
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
di: Rashid, Md Hasanur, et al.
Pubblicazione: (2026)
HeteGen: Heterogeneous Parallel Inference for Large Language Models on Resource-Constrained Devices
di: Zhao, Xuanlei, et al.
Pubblicazione: (2024)
di: Zhao, Xuanlei, et al.
Pubblicazione: (2024)
Enabling High-Throughput Parallel I/O in Particle-in-Cell Monte Carlo Simulations with openPMD and Darshan I/O Monitoring
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
AcceleratedKernels.jl: Cross-Architecture Parallel Algorithms from a Unified, Transpiled Codebase
di: Nicusan, Andrei-Leonard, et al.
Pubblicazione: (2025)
di: Nicusan, Andrei-Leonard, et al.
Pubblicazione: (2025)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
di: Ather, Hammad, et al.
Pubblicazione: (2024)
di: Ather, Hammad, et al.
Pubblicazione: (2024)
EfiMon: A Process Analyser for Granular Power Consumption Prediction
di: León-Vega, Luis G., et al.
Pubblicazione: (2024)
di: León-Vega, Luis G., et al.
Pubblicazione: (2024)
A Comprehensive Analysis of Process Energy Consumption on Multi-Socket Systems with GPUs
di: León-Vega, Luis G., et al.
Pubblicazione: (2024)
di: León-Vega, Luis G., et al.
Pubblicazione: (2024)
Kino-PAX: Highly Parallel Kinodynamic Sampling-based Planner
di: Perrault, Nicolas, et al.
Pubblicazione: (2024)
di: Perrault, Nicolas, et al.
Pubblicazione: (2024)
GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving
di: Jayakody, Shakya, et al.
Pubblicazione: (2026)
di: Jayakody, Shakya, et al.
Pubblicazione: (2026)
Can Large Language Models Predict Parallel Code Performance?
di: Bolet, Gregory, et al.
Pubblicazione: (2025)
di: Bolet, Gregory, et al.
Pubblicazione: (2025)
GigaAPI for GPU Parallelization
di: Suvarna, M., et al.
Pubblicazione: (2025)
di: Suvarna, M., et al.
Pubblicazione: (2025)
Supercomputers as a Continous Medium
di: Karp, Martin, et al.
Pubblicazione: (2024)
di: Karp, Martin, et al.
Pubblicazione: (2024)
Accelerating Particle-in-Cell Monte Carlo Simulations with MPI, OpenMP/OpenACC and Asynchronous Multi-GPU Programming
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)
Comparing Parallel Functional Array Languages: Programming and Performance
di: van Balen, David, et al.
Pubblicazione: (2025)
di: van Balen, David, et al.
Pubblicazione: (2025)
Integrating High Performance In-Memory Data Streaming and In-Situ Visualization in Hybrid MPI+OpenMP PIC MC Simulations Towards Exascale
di: Williams, Jeremy J., et al.
Pubblicazione: (2025)
di: Williams, Jeremy J., et al.
Pubblicazione: (2025)
Parallelizing a modern GPU simulator
di: Huerta, Rodrigo, et al.
Pubblicazione: (2025)
di: Huerta, Rodrigo, et al.
Pubblicazione: (2025)
Multi-GPU Hybrid Particle-in-Cell Monte Carlo Simulations for Exascale Computing Systems
di: Williams, Jeremy J., et al.
Pubblicazione: (2026)
di: Williams, Jeremy J., et al.
Pubblicazione: (2026)
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
di: Barron, Ryan, et al.
Pubblicazione: (2024)
di: Barron, Ryan, et al.
Pubblicazione: (2024)
Scrutinizing Variables for Checkpoint Using Automatic Differentiation
di: Huang, Xin, et al.
Pubblicazione: (2026)
di: Huang, Xin, et al.
Pubblicazione: (2026)
Scaling Large-scale GNN Training to Thousands of Processors on CPU-based Supercomputers
di: Zhuang, Chen, et al.
Pubblicazione: (2024)
di: Zhuang, Chen, et al.
Pubblicazione: (2024)
Profiling and optimization of multi-card GPU machine learning jobs
di: Lawenda, Marcin, et al.
Pubblicazione: (2025)
di: Lawenda, Marcin, et al.
Pubblicazione: (2025)
WebAssembly and Unikernels: A Comparative Study for Serverless at the Edge
di: Besozzi, Valerio, et al.
Pubblicazione: (2025)
di: Besozzi, Valerio, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
di: Andersson, Måns I., et al.
Pubblicazione: (2025) -
Understanding Power Consumption Metric on Heterogeneous Memory Systems
di: Proaño, Andrès Rubio, et al.
Pubblicazione: (2024) -
Extracting Practical, Actionable Energy Insights from Supercomputer Telemetry and Logs
di: Cornelius, Melanie, et al.
Pubblicazione: (2025) -
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
di: Wang, Yuxin, et al.
Pubblicazione: (2023) -
Characterizing the Performance of the Implicit Massively Parallel Particle-in-Cell iPIC3D Code
di: Williams, Jeremy J., et al.
Pubblicazione: (2024)