Profiling and Modeling of Power Characteristics of Leadership-Scale HPC System Workloads
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Karimi, Ahmad Maroof, Sattar, Naw Safrin, Shin, Woong, Wang, Feiyi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring the Frontiers of Energy Efficiency using Power Management at System Scale
von: Karimi, Ahmad Maroof, et al.
Veröffentlicht: (2024)
von: Karimi, Ahmad Maroof, et al.
Veröffentlicht: (2024)
Minos: Systematically Classifying Performance and Power Characteristics of GPU Workloads on HPC Clusters
von: Jain, Rutwik, et al.
Veröffentlicht: (2026)
von: Jain, Rutwik, et al.
Veröffentlicht: (2026)
AI-coupled HPC Workflow Applications, Middleware and Performance
von: Brewer, Wes, et al.
Veröffentlicht: (2024)
von: Brewer, Wes, et al.
Veröffentlicht: (2024)
Kub: Enabling Elastic HPC Workloads on Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
A Review of Tools and Techniques for Optimization of Workload Mapping and Scheduling in Heterogeneous HPC System
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)
GreenFaaS: Maximizing Energy Efficiency of HPC Workloads with FaaS
von: Kamatar, Alok, et al.
Veröffentlicht: (2024)
von: Kamatar, Alok, et al.
Veröffentlicht: (2024)
Running Cloud-native Workloads on HPC with High-Performance Kubernetes
von: Chazapis, Antony, et al.
Veröffentlicht: (2024)
von: Chazapis, Antony, et al.
Veröffentlicht: (2024)
On the Convergence of Malleability and the HPC PowerStack: Exploiting Dynamism in Over-Provisioned and Power-Constrained HPC Systems
von: Arima, Eishi, et al.
Veröffentlicht: (2024)
von: Arima, Eishi, et al.
Veröffentlicht: (2024)
Resource Optimization with MPI Process Malleability for Dynamic Workloads in HPC Clusters
von: Iserte, Sergio, et al.
Veröffentlicht: (2025)
von: Iserte, Sergio, et al.
Veröffentlicht: (2025)
ARC-V: Vertical Resource Adaptivity for HPC Workloads in Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2025)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2025)
Evaluating Malleable Job Scheduling in HPC Clusters using Real-World Workloads
von: Zojer, Patrick, et al.
Veröffentlicht: (2026)
von: Zojer, Patrick, et al.
Veröffentlicht: (2026)
Agentic AI Workload Characteristics
von: Yuan, Yichao, et al.
Veröffentlicht: (2026)
von: Yuan, Yichao, et al.
Veröffentlicht: (2026)
Hydra: Brokering Cloud and HPC Resources to Support the Execution of Heterogeneous Workloads at Scale
von: Alsaadi, Aymen, et al.
Veröffentlicht: (2024)
von: Alsaadi, Aymen, et al.
Veröffentlicht: (2024)
An Incremental Multi-Level, Multi-Scale Approach to Assessment of Multifidelity HPC Systems
von: Shilpika, Shilpika, et al.
Veröffentlicht: (2025)
von: Shilpika, Shilpika, et al.
Veröffentlicht: (2025)
Integrating and Characterizing HPC Task Runtime Systems for hybrid AI-HPC workloads
von: Merzky, Andre, et al.
Veröffentlicht: (2025)
von: Merzky, Andre, et al.
Veröffentlicht: (2025)
RHAPSODY: Execution of Hybrid AI-HPC Workflows at Scale
von: Alsaadi, Aymen, et al.
Veröffentlicht: (2025)
von: Alsaadi, Aymen, et al.
Veröffentlicht: (2025)
Understanding Large-Scale HPC System Behavior Through Cluster-Based Visual Analytics
von: Austin, Allison, et al.
Veröffentlicht: (2026)
von: Austin, Allison, et al.
Veröffentlicht: (2026)
Scheduling Data-Intensive Workloads in Large-Scale Distributed Systems: Trends and Challenges
von: Stavrinides, Georgios L., et al.
Veröffentlicht: (2025)
von: Stavrinides, Georgios L., et al.
Veröffentlicht: (2025)
MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production
von: Xue, Chunyu, et al.
Veröffentlicht: (2026)
von: Xue, Chunyu, et al.
Veröffentlicht: (2026)
Parallel Paradigms in Modern HPC: A Comparative Analysis of MPI, OpenMP, and CUDA
von: ALHafez, Nizar, et al.
Veröffentlicht: (2025)
von: ALHafez, Nizar, et al.
Veröffentlicht: (2025)
FLYING SERVING: On-the-Fly Parallelism Switching for Large Language Model Serving
von: Gao, Shouwei, et al.
Veröffentlicht: (2026)
von: Gao, Shouwei, et al.
Veröffentlicht: (2026)
Beyond Pre-Training: The Full Lifecycle of Foundation Models on HPC Systems
von: Conciatore, Dino, et al.
Veröffentlicht: (2026)
von: Conciatore, Dino, et al.
Veröffentlicht: (2026)
SIREN: Software Identification and Recognition in HPC Systems
von: Jakobsche, Thomas, et al.
Veröffentlicht: (2025)
von: Jakobsche, Thomas, et al.
Veröffentlicht: (2025)
Power-Aware Scheduling for Multi-Center HPC Electricity Cost Optimization
von: Hossain, Abrar, et al.
Veröffentlicht: (2025)
von: Hossain, Abrar, et al.
Veröffentlicht: (2025)
Extrae.jl: Julia bindings for the Extrae HPC Profiler
von: Sanchez-Ramirez, Sergio, et al.
Veröffentlicht: (2025)
von: Sanchez-Ramirez, Sergio, et al.
Veröffentlicht: (2025)
ThirstyFLOPS: Water Footprint Modeling and Analysis Toward Sustainable HPC Systems
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
Driving Computational Efficiency in Large-Scale Platforms using HPC Technologies
von: Mendez, Alexander Martinez, et al.
Veröffentlicht: (2026)
von: Mendez, Alexander Martinez, et al.
Veröffentlicht: (2026)
A Heavy-Load-Enhanced and Changeable-Periodicity-Perceived Workload Prediction Network
von: Chen, Feiyi, et al.
Veröffentlicht: (2023)
von: Chen, Feiyi, et al.
Veröffentlicht: (2023)
Leveraging HPC Profiling & Tracing Tools to Understand the Performance of Particle-in-Cell Monte Carlo Simulations
von: Williams, Jeremy J., et al.
Veröffentlicht: (2023)
von: Williams, Jeremy J., et al.
Veröffentlicht: (2023)
Towards an Adaptive Runtime System for Cloud-Native HPC
von: Bhosale, Aditya, et al.
Veröffentlicht: (2026)
von: Bhosale, Aditya, et al.
Veröffentlicht: (2026)
SPARS: A Reinforcement Learning-Enabled Simulator for Power Management in HPC Job Scheduling
von: Amrizal, Muhammad Alfian, et al.
Veröffentlicht: (2025)
von: Amrizal, Muhammad Alfian, et al.
Veröffentlicht: (2025)
Report on Challenges of Practical Reproducibility for Systems and HPC Computer Science
von: Keahey, Kate, et al.
Veröffentlicht: (2025)
von: Keahey, Kate, et al.
Veröffentlicht: (2025)
Data Management System Analysis for Distributed Computing Workloads
von: Hsu, Kuan-Chieh, et al.
Veröffentlicht: (2025)
von: Hsu, Kuan-Chieh, et al.
Veröffentlicht: (2025)
The (R)evolution of Scientific Workflows in the Agentic AI Era: Towards Autonomous Science
von: Shin, Woong, et al.
Veröffentlicht: (2025)
von: Shin, Woong, et al.
Veröffentlicht: (2025)
A Deep Dive into the Google Cluster Workload Traces: Analyzing the Application Failure Characteristics and User Behaviors
von: Bappy, Faisal Haque, et al.
Veröffentlicht: (2023)
von: Bappy, Faisal Haque, et al.
Veröffentlicht: (2023)
A Contention-Free Model for Converged Kubernetes on HPC
von: Sochat, Vanessa, et al.
Veröffentlicht: (2024)
von: Sochat, Vanessa, et al.
Veröffentlicht: (2024)
MIDAS: Adaptive Proxy Middleware for Mitigating Metadata Hotspots in HPC I/O at Scale
von: Ghimire, Sangam, et al.
Veröffentlicht: (2025)
von: Ghimire, Sangam, et al.
Veröffentlicht: (2025)
Workflow-Driven Modeling for the Compute Continuum: An Optimization Approach to Automated System and Workload Scheduling
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)
PRISM: Dynamic Primitive-Based Forecasting for Large-Scale GPU Cluster Workloads
von: Wu, Xin, et al.
Veröffentlicht: (2026)
von: Wu, Xin, et al.
Veröffentlicht: (2026)
AI Surrogate Model for Distributed Computing Workloads
von: Park, David K., et al.
Veröffentlicht: (2024)
von: Park, David K., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Exploring the Frontiers of Energy Efficiency using Power Management at System Scale
von: Karimi, Ahmad Maroof, et al.
Veröffentlicht: (2024) -
Minos: Systematically Classifying Performance and Power Characteristics of GPU Workloads on HPC Clusters
von: Jain, Rutwik, et al.
Veröffentlicht: (2026) -
AI-coupled HPC Workflow Applications, Middleware and Performance
von: Brewer, Wes, et al.
Veröffentlicht: (2024) -
Kub: Enabling Elastic HPC Workloads on Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024) -
A Review of Tools and Techniques for Optimization of Workload Mapping and Scheduling in Heterogeneous HPC System
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)