Interactive and Urgent HPC: Challenges and Opportunities
Fuente:
arXiv
Saved in:
| Main Authors: | Reuther, Albert, Brown, Nick, Arndt, William, Blaschke, Johannes, Boehme, Christian, Chazapis, Antony, Enders, Bjoern, Henschel, Robert, Kunkel, Julian, Martinasso, Maxime |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Interactive and Urgent HPC: State of the Research
by: Reuther, Albert, et al.
Published: (2026)
by: Reuther, Albert, et al.
Published: (2026)
Running Cloud-native Workloads on HPC with High-Performance Kubernetes
by: Chazapis, Antony, et al.
Published: (2024)
by: Chazapis, Antony, et al.
Published: (2024)
Optimizing the Longhorn Cloud-native Software Defined Storage Engine for High Performance
by: Kampadais, Konstantinos, et al.
Published: (2025)
by: Kampadais, Konstantinos, et al.
Published: (2025)
A Review of Tools and Techniques for Optimization of Workload Mapping and Scheduling in Heterogeneous HPC System
by: Sharma, Aasish Kumar, et al.
Published: (2025)
by: Sharma, Aasish Kumar, et al.
Published: (2025)
Beyond Pre-Training: The Full Lifecycle of Foundation Models on HPC Systems
by: Conciatore, Dino, et al.
Published: (2026)
by: Conciatore, Dino, et al.
Published: (2026)
RISC-V for HPC: Where we are and where we need to go
by: Brown, Nick
Published: (2024)
by: Brown, Nick
Published: (2024)
RISC-V for HPC: An update of where we are and main action points
by: Brown, Nick
Published: (2025)
by: Brown, Nick
Published: (2025)
Alps, a versatile research infrastructure
by: Martinasso, Maxime, et al.
Published: (2025)
by: Martinasso, Maxime, et al.
Published: (2025)
Workflow-Driven Modeling for the Compute Continuum: An Optimization Approach to Automated System and Workload Scheduling
by: Sharma, Aasish Kumar, et al.
Published: (2025)
by: Sharma, Aasish Kumar, et al.
Published: (2025)
Optimizing Checkpoint-Restart Mechanisms for HPC with DMTCP in Containers at NERSC
by: Timalsina, Madan, et al.
Published: (2024)
by: Timalsina, Madan, et al.
Published: (2024)
Performance characterisation of the 64-core SG2042 RISC-V CPU for HPC
by: Brown, Nick, et al.
Published: (2024)
by: Brown, Nick, et al.
Published: (2024)
Implementing OpenMP for Zig to enable its use in HPC context
by: Kacs, David, et al.
Published: (2024)
by: Kacs, David, et al.
Published: (2024)
GrapheonRL: A Graph Neural Network and Reinforcement Learning Framework for Constraint and Data-Aware Workflow Mapping and Scheduling in Heterogeneous HPC Systems
by: Sharma, Aasish Kumar, et al.
Published: (2025)
by: Sharma, Aasish Kumar, et al.
Published: (2025)
Guardian: Safe GPU Sharing in Multi-Tenant Environments
by: Pavlidakis, Manos, et al.
Published: (2024)
by: Pavlidakis, Manos, et al.
Published: (2024)
Chat AI: A Seamless Slurm-Native Solution for HPC-Based Services
by: Doosthosseini, Ali, et al.
Published: (2024)
by: Doosthosseini, Ali, et al.
Published: (2024)
HPC with Enhanced User Separation
by: Prout, Andrew, et al.
Published: (2024)
by: Prout, Andrew, et al.
Published: (2024)
Investigations of multi-socket high core count RISC-V for HPC workloads
by: Brown, Nick, et al.
Published: (2025)
by: Brown, Nick, et al.
Published: (2025)
Evolving HPC services to enable ML workloads on HPE Cray EX
by: Schuppli, Stefano, et al.
Published: (2025)
by: Schuppli, Stefano, et al.
Published: (2025)
Addressing Reproducibility Challenges in HPC with Continuous Integration
by: Hayot-Sasson, Valérie, et al.
Published: (2025)
by: Hayot-Sasson, Valérie, et al.
Published: (2025)
DECICE: AI-Driven Scheduling and Digital Twin Integration for the Cloud-HPC-Edge Compute Continuum
by: Sharma, Aasish Kumar, et al.
Published: (2026)
by: Sharma, Aasish Kumar, et al.
Published: (2026)
The First OpenFOAM HPC Challenge (OHC-1)
by: Lesnik, Sergey, et al.
Published: (2026)
by: Lesnik, Sergey, et al.
Published: (2026)
Is RISC-V ready for High Performance Computing? An evaluation of the Sophon SG2044
by: Brown, Nick
Published: (2025)
by: Brown, Nick
Published: (2025)
Report on Challenges of Practical Reproducibility for Systems and HPC Computer Science
by: Keahey, Kate, et al.
Published: (2025)
by: Keahey, Kate, et al.
Published: (2025)
Scalable HPC Job Scheduling and Resource Management in SST
by: Abdurahman, Abubeker, et al.
Published: (2025)
by: Abdurahman, Abubeker, et al.
Published: (2025)
LLload: Simplifying Real-Time Job Monitoring for HPC Users
by: Byun, Chansup, et al.
Published: (2024)
by: Byun, Chansup, et al.
Published: (2024)
Core Hours and Carbon Credits: Incentivizing Sustainability in HPC
by: Kamatar, Alok, et al.
Published: (2025)
by: Kamatar, Alok, et al.
Published: (2025)
RHAPSODY: Execution of Hybrid AI-HPC Workflows at Scale
by: Alsaadi, Aymen, et al.
Published: (2025)
by: Alsaadi, Aymen, et al.
Published: (2025)
Integrating and Characterizing HPC Task Runtime Systems for hybrid AI-HPC workloads
by: Merzky, Andre, et al.
Published: (2025)
by: Merzky, Andre, et al.
Published: (2025)
Accelerating stencils on the Tenstorrent Grayskull RISC-V accelerator
by: Brown, Nick, et al.
Published: (2024)
by: Brown, Nick, et al.
Published: (2024)
Accelerating Time-to-Science by Streaming Detector Data Directly into Perlmutter Compute Nodes
by: Welborn, Samuel S., et al.
Published: (2024)
by: Welborn, Samuel S., et al.
Published: (2024)
Real-Time XFEL Data Analysis at SLAC and NERSC: a Trial Run of Nascent Exascale Experimental Data Analysis
by: Blaschke, Johannes P., et al.
Published: (2021)
by: Blaschke, Johannes P., et al.
Published: (2021)
Analysis of the carbon footprint of HPC
by: Benhari, Abdessalam, et al.
Published: (2025)
by: Benhari, Abdessalam, et al.
Published: (2025)
Lifting to tensors when compiling scientific computing workloads for AI Engines
by: Brown, Nick, et al.
Published: (2026)
by: Brown, Nick, et al.
Published: (2026)
On the Convergence of Malleability and the HPC PowerStack: Exploiting Dynamism in Over-Provisioned and Power-Constrained HPC Systems
by: Arima, Eishi, et al.
Published: (2024)
by: Arima, Eishi, et al.
Published: (2024)
MRSch: Multi-Resource Scheduling for HPC
by: Li, Boyang, et al.
Published: (2024)
by: Li, Boyang, et al.
Published: (2024)
When GPUs Fail Quietly: Observability-Aware Early Warning Beyond Numeric Telemetry
by: Bidollahkhani, Michael, et al.
Published: (2026)
by: Bidollahkhani, Michael, et al.
Published: (2026)
UNR: Unified Notifiable RMA Library for HPC
by: Feng, Guangnan, et al.
Published: (2024)
by: Feng, Guangnan, et al.
Published: (2024)
An Elastic Job Scheduler for HPC Applications on the Cloud
by: Bhosale, Aditya, et al.
Published: (2025)
by: Bhosale, Aditya, et al.
Published: (2025)
Sarus Suite: Cloud-native Containers for HPC
by: Madonna, Alberto, et al.
Published: (2026)
by: Madonna, Alberto, et al.
Published: (2026)
Wilkins: HPC In Situ Workflows Made Easy
by: Yildiz, Orcun, et al.
Published: (2024)
by: Yildiz, Orcun, et al.
Published: (2024)
Similar Items
-
Interactive and Urgent HPC: State of the Research
by: Reuther, Albert, et al.
Published: (2026) -
Running Cloud-native Workloads on HPC with High-Performance Kubernetes
by: Chazapis, Antony, et al.
Published: (2024) -
Optimizing the Longhorn Cloud-native Software Defined Storage Engine for High Performance
by: Kampadais, Konstantinos, et al.
Published: (2025) -
A Review of Tools and Techniques for Optimization of Workload Mapping and Scheduling in Heterogeneous HPC System
by: Sharma, Aasish Kumar, et al.
Published: (2025) -
Beyond Pre-Training: The Full Lifecycle of Foundation Models on HPC Systems
by: Conciatore, Dino, et al.
Published: (2026)