Adventures with Grace Hopper AI Super Chip and the National Research Platform
Fuente:
arXiv
Salvato in:
| Autori principali: | Hurt, J. Alex, Scott, Grant J., Weitzel, Derek, Zhu, Huijun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The National Research Platform: Stretched, Multi-Tenant, Scientific Kubernetes Cluster
di: Weitzel, Derek, et al.
Pubblicazione: (2025)
di: Weitzel, Derek, et al.
Pubblicazione: (2025)
Cross-Layer Energy Analysis of Multimodal Training on Grace Hopper Superchips
di: Ahmed, Mahmoud, et al.
Pubblicazione: (2026)
di: Ahmed, Mahmoud, et al.
Pubblicazione: (2026)
Automatic BLAS Offloading on Unified Memory Architecture: A Study on NVIDIA Grace-Hopper
di: Li, Junjie, et al.
Pubblicazione: (2024)
di: Li, Junjie, et al.
Pubblicazione: (2024)
Harnessing Integrated CPU-GPU System Memory for HPC: a first look into Grace Hopper
di: Schieffer, Gabin, et al.
Pubblicazione: (2024)
di: Schieffer, Gabin, et al.
Pubblicazione: (2024)
Understanding Data Movement in Tightly Coupled Heterogeneous Systems: A Case Study with the Grace Hopper Superchip
di: Fusco, Luigi, et al.
Pubblicazione: (2024)
di: Fusco, Luigi, et al.
Pubblicazione: (2024)
Scaling Deep Learning Research with Kubernetes on the NRP Nautilus HyperCluster
di: Hurt, J. Alex, et al.
Pubblicazione: (2024)
di: Hurt, J. Alex, et al.
Pubblicazione: (2024)
Open Science Data Federation -- operation and monitoring
di: Andrijauskas, Fabio, et al.
Pubblicazione: (2026)
di: Andrijauskas, Fabio, et al.
Pubblicazione: (2026)
ML-based Adaptive Prefetching and Data Placement for US HEP Systems
di: Karanam, Venkat Sai Suman Lamba, et al.
Pubblicazione: (2025)
di: Karanam, Venkat Sai Suman Lamba, et al.
Pubblicazione: (2025)
Resolving Conflicts with Grace: Dynamically Concurrent Universality
di: Kuznetsov, Petr, et al.
Pubblicazione: (2025)
di: Kuznetsov, Petr, et al.
Pubblicazione: (2025)
ARM SVE Unleashed: Performance and Insights Across HPC Applications on Nvidia Grace
di: Shi, Ruimin, et al.
Pubblicazione: (2025)
di: Shi, Ruimin, et al.
Pubblicazione: (2025)
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
di: Zhu, Jianwei, et al.
Pubblicazione: (2024)
di: Zhu, Jianwei, et al.
Pubblicazione: (2024)
SuperBench: Improving Cloud AI Infrastructure Reliability with Proactive Validation
di: Xiong, Yifan, et al.
Pubblicazione: (2024)
di: Xiong, Yifan, et al.
Pubblicazione: (2024)
Microarchitectural comparison and in-core modeling of state-of-the-art CPUs: Grace, Sapphire Rapids, and Genoa
di: Laukemann, Jan, et al.
Pubblicazione: (2024)
di: Laukemann, Jan, et al.
Pubblicazione: (2024)
Story of Two GPUs: Characterizing the Resilience of Hopper H100 and Ampere A100 GPUs
di: Cui, Shengkun, et al.
Pubblicazione: (2025)
di: Cui, Shengkun, et al.
Pubblicazione: (2025)
Driving Computational Efficiency in Large-Scale Platforms using HPC Technologies
di: Mendez, Alexander Martinez, et al.
Pubblicazione: (2026)
di: Mendez, Alexander Martinez, et al.
Pubblicazione: (2026)
Dissecting the NVIDIA Hopper Architecture through Microbenchmarking and Multiple Level Analysis
di: Luo, Weile, et al.
Pubblicazione: (2025)
di: Luo, Weile, et al.
Pubblicazione: (2025)
Cooperative Graceful Degradation In Containerized Clouds
di: Agrawal, Kapil, et al.
Pubblicazione: (2023)
di: Agrawal, Kapil, et al.
Pubblicazione: (2023)
SuperSFL: Resource-Heterogeneous Federated Split Learning with Weight-Sharing Super-Networks
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
The HEAL Data Platform
di: Larrick, Brienna M., et al.
Pubblicazione: (2025)
di: Larrick, Brienna M., et al.
Pubblicazione: (2025)
MoEntwine: Unleashing the Potential of Wafer-scale Chips for Large-scale Expert Parallel Inference
di: Tang, Xinru, et al.
Pubblicazione: (2025)
di: Tang, Xinru, et al.
Pubblicazione: (2025)
Toward Self-Healing Networks-on-Chip: RL-Driven Routing in 2D Torus Architectures
di: Charrwi, Mohammad Walid, et al.
Pubblicazione: (2025)
di: Charrwi, Mohammad Walid, et al.
Pubblicazione: (2025)
Characterizing Adaptive Mesh Refinement on Heterogeneous Platforms with Parthenon-VIBE
di: Poptani, Akash, et al.
Pubblicazione: (2025)
di: Poptani, Akash, et al.
Pubblicazione: (2025)
Model Input Verification of Large Scale Simulations
di: Neykova, Rumyana, et al.
Pubblicazione: (2024)
di: Neykova, Rumyana, et al.
Pubblicazione: (2024)
6G EdgeAI: Performance Evaluation and Analysis
di: Yang, Chien-Sheng, et al.
Pubblicazione: (2025)
di: Yang, Chien-Sheng, et al.
Pubblicazione: (2025)
GeoFaaS: An Edge-to-Cloud FaaS Platform
di: Malekabbasi, Mohammadreza, et al.
Pubblicazione: (2024)
di: Malekabbasi, Mohammadreza, et al.
Pubblicazione: (2024)
Towards a Testbed for Scalable FaaS Platforms
di: Schirmer, Trever, et al.
Pubblicazione: (2025)
di: Schirmer, Trever, et al.
Pubblicazione: (2025)
AI4EOSC: a Federated Cloud Platform for Artificial Intelligence in Scientific Research
di: Heredia, Ignacio, et al.
Pubblicazione: (2025)
di: Heredia, Ignacio, et al.
Pubblicazione: (2025)
Huawei Cloud Model-as-a-Service on the CloudMatrix384 SuperPod
di: Xiao, Ao, et al.
Pubblicazione: (2025)
di: Xiao, Ao, et al.
Pubblicazione: (2025)
GreenWhisk: Emission-Aware Computing for Serverless Platform
di: Serenari, Jayden, et al.
Pubblicazione: (2024)
di: Serenari, Jayden, et al.
Pubblicazione: (2024)
BCL: A Cross-Platform Distributed Container Library
di: Brock, Benjamin, et al.
Pubblicazione: (2018)
di: Brock, Benjamin, et al.
Pubblicazione: (2018)
Optimizing FaaS Platforms for MCP-enabled Agentic Workflows
di: Kulkarni, Varad, et al.
Pubblicazione: (2026)
di: Kulkarni, Varad, et al.
Pubblicazione: (2026)
CIR: Lightweight Container Image for Cross-Platform Deployment
di: Li, Fengzhi, et al.
Pubblicazione: (2026)
di: Li, Fengzhi, et al.
Pubblicazione: (2026)
Holoscope: Open and Lightweight Distributed Telescope & Honeypot Platform
di: Sordello, Andrea, et al.
Pubblicazione: (2025)
di: Sordello, Andrea, et al.
Pubblicazione: (2025)
H2:Towards Efficient Large-Scale LLM Training on Hyper-Heterogeneous Cluster over 1,000 Chips
di: Tang, Ding, et al.
Pubblicazione: (2025)
di: Tang, Ding, et al.
Pubblicazione: (2025)
Mapping Large Memory-constrained Workflows onto Heterogeneous Platforms
di: Kulagina, Svetlana, et al.
Pubblicazione: (2024)
di: Kulagina, Svetlana, et al.
Pubblicazione: (2024)
PLASMA -- Platform for Service Management in Digital Remote Maintenance Applications
di: Stumpp, Natascha, et al.
Pubblicazione: (2024)
di: Stumpp, Natascha, et al.
Pubblicazione: (2024)
Large-Scale Metric Computation in Online Controlled Experiment Platform
di: Xiong, Tao, et al.
Pubblicazione: (2024)
di: Xiong, Tao, et al.
Pubblicazione: (2024)
ElastiBench: Scalable Continuous Benchmarking on Cloud FaaS Platforms
di: Schirmer, Trever, et al.
Pubblicazione: (2024)
di: Schirmer, Trever, et al.
Pubblicazione: (2024)
Sustainable Grid through Distributed Data Centers: Spinning AI Demand for Grid Stabilization and Optimization
di: Evans, Scott C, et al.
Pubblicazione: (2025)
di: Evans, Scott C, et al.
Pubblicazione: (2025)
Efficient and Portable Support for Overdecomposition on Distributed Memory GPGPU Platforms
di: Bhosale, Aditya, et al.
Pubblicazione: (2026)
di: Bhosale, Aditya, et al.
Pubblicazione: (2026)
Documenti analoghi
-
The National Research Platform: Stretched, Multi-Tenant, Scientific Kubernetes Cluster
di: Weitzel, Derek, et al.
Pubblicazione: (2025) -
Cross-Layer Energy Analysis of Multimodal Training on Grace Hopper Superchips
di: Ahmed, Mahmoud, et al.
Pubblicazione: (2026) -
Automatic BLAS Offloading on Unified Memory Architecture: A Study on NVIDIA Grace-Hopper
di: Li, Junjie, et al.
Pubblicazione: (2024) -
Harnessing Integrated CPU-GPU System Memory for HPC: a first look into Grace Hopper
di: Schieffer, Gabin, et al.
Pubblicazione: (2024) -
Understanding Data Movement in Tightly Coupled Heterogeneous Systems: A Case Study with the Grace Hopper Superchip
di: Fusco, Luigi, et al.
Pubblicazione: (2024)