Characterizing GPU Energy Usage in Exascale-Ready Portable Science Applications
Fuente:
arXiv
Salvato in:
| Autori principali: | Godoy, William F., Hernandez, Oscar, Kent, Paul R. C., Patrou, Maria, Asifuzzaman, Kazi, Miniskar, Narasinga Rao, Valero-Lara, Pedro, Vetter, Jeffrey S., Sinclair, Matthew D., Lowe-Power, Jason, Bruce, Bobby R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mapping Spiking Neural Networks to Heterogeneous Crossbar Architectures using Integer Linear Programming
di: Pohl, Devin, et al.
Pubblicazione: (2025)
di: Pohl, Devin, et al.
Pubblicazione: (2025)
Classic and Quantum Task-Based Intelligent Runtime for QIRs Running on Multiple QPUs
di: Miniskar, Narasinga Rao, et al.
Pubblicazione: (2026)
di: Miniskar, Narasinga Rao, et al.
Pubblicazione: (2026)
Q-IRIS: The Evolution of the IRIS Task-Based Runtime to Enable Classical-Quantum Workflows
di: Miniskar, Narasinga Rao, et al.
Pubblicazione: (2025)
di: Miniskar, Narasinga Rao, et al.
Pubblicazione: (2025)
Mojo: MLIR-Based Performance-Portable HPC Science Kernels on GPUs for the Python Ecosystem
di: Godoy, William F., et al.
Pubblicazione: (2025)
di: Godoy, William F., et al.
Pubblicazione: (2025)
Large language model evaluation for high‐performance computing software development
di: William F. Godoy, et al.
Pubblicazione: (2024)
di: William F. Godoy, et al.
Pubblicazione: (2024)
LAMMPS-KOKKOS: Performance Portable Molecular Dynamics Across Exascale Architectures
di: Johansson, Anders, et al.
Pubblicazione: (2025)
di: Johansson, Anders, et al.
Pubblicazione: (2025)
MARUT: An Exascale-Ready, GPU-Accelerated High-Order CFD Framework with AMR for High-Speed Flows and Finite-Rate Chemistry
di: Mondal, Trishit, et al.
Pubblicazione: (2026)
di: Mondal, Trishit, et al.
Pubblicazione: (2026)
exa-AMD: An Exascale-Ready Framework for Accelerating the Discovery and Design of Functional Materials
di: Xia, Weiyi, et al.
Pubblicazione: (2025)
di: Xia, Weiyi, et al.
Pubblicazione: (2025)
Portable Targeted Sampling Framework Using LLVM
di: Qiu, Zhantong, et al.
Pubblicazione: (2025)
di: Qiu, Zhantong, et al.
Pubblicazione: (2025)
PETSc/TAO Developments for GPU-Based Early Exascale Systems
di: Mills, Richard Tran, et al.
Pubblicazione: (2024)
di: Mills, Richard Tran, et al.
Pubblicazione: (2024)
Monitoring and Characterizing GPU Usage
di: Le Mai Weakley, et al.
Pubblicazione: (2025)
di: Le Mai Weakley, et al.
Pubblicazione: (2025)
Power-Capping Metric Evaluation for Improving Energy Efficiency in HPC Applications
di: Patrou, Maria, et al.
Pubblicazione: (2025)
di: Patrou, Maria, et al.
Pubblicazione: (2025)
Fine-Grained Power and Energy Attribution on AMD GPU/APU-Based Exascale Nodes
di: McDaniel, Adam, et al.
Pubblicazione: (2026)
di: McDaniel, Adam, et al.
Pubblicazione: (2026)
Portable GPU implementation of the WP-CCC ion-atom collisions code
di: Abdurakhmanov, I. B., et al.
Pubblicazione: (2024)
di: Abdurakhmanov, I. B., et al.
Pubblicazione: (2024)
Exploring the Readiness of Prominent Small Language Models for the Democratization of Financial Literacy
di: Kosireddy, Tagore Rao, et al.
Pubblicazione: (2024)
di: Kosireddy, Tagore Rao, et al.
Pubblicazione: (2024)
ELEQTRONeX: A GPU-Accelerated Exascale Framework for Non-Equilibrium Quantum Transport in Nanomaterials
di: Sawant, Saurabh, et al.
Pubblicazione: (2024)
di: Sawant, Saurabh, et al.
Pubblicazione: (2024)
Multi-GPU Hybrid Particle-in-Cell Monte Carlo Simulations for Exascale Computing Systems
di: Williams, Jeremy J., et al.
Pubblicazione: (2026)
di: Williams, Jeremy J., et al.
Pubblicazione: (2026)
GROMACS on AMD GPU-Based HPC Platforms: Using SYCL for Performance and Portability
di: Alekseenko, Andrey, et al.
Pubblicazione: (2024)
di: Alekseenko, Andrey, et al.
Pubblicazione: (2024)
A Portable Multi-GPU Solver for Collisional Plasmas with Coulombic Interactions
di: Almgren-Bell, James, et al.
Pubblicazione: (2025)
di: Almgren-Bell, James, et al.
Pubblicazione: (2025)
Taking GPU Programming Models to Task for Performance Portability
di: Davis, Joshua H., et al.
Pubblicazione: (2024)
di: Davis, Joshua H., et al.
Pubblicazione: (2024)
Universal Quantum Computer Simulation of 50 Qubits on Europe`s First Exascale Supercomputer Harnessing Its Heterogeneous CPU-GPU Architecture
di: De Raedt, Hans, et al.
Pubblicazione: (2025)
di: De Raedt, Hans, et al.
Pubblicazione: (2025)
Physics-Guided Deep Learning for Heat Pump Stress Detection: A Comprehensive Analysis on When2Heat Dataset
di: Alam, Md Shahabub, et al.
Pubblicazione: (2025)
di: Alam, Md Shahabub, et al.
Pubblicazione: (2025)
Portable PGAS ‐Based GPU ‐Accelerated Branch‐And‐Bound Algorithms at Scale
di: Guillaume Helbecque, et al.
Pubblicazione: (2025)
di: Guillaume Helbecque, et al.
Pubblicazione: (2025)
Leveraging AI for Productive and Trustworthy HPC Software: Challenges and Research Directions
di: Teranishi, Keita, et al.
Pubblicazione: (2025)
di: Teranishi, Keita, et al.
Pubblicazione: (2025)
High-Performance Portable GPU Primitives for Arbitrary Types and Operators in Julia
di: Pilliat, Emmanuel
Pubblicazione: (2026)
di: Pilliat, Emmanuel
Pubblicazione: (2026)
Towards Exascale Computation for Turbomachinery Flows
di: Fu, Yuhang, et al.
Pubblicazione: (2023)
di: Fu, Yuhang, et al.
Pubblicazione: (2023)
Beyond Exascale: Dataflow Domain Translation on a Cerebras Cluster
di: Oppelstrup, Tomas, et al.
Pubblicazione: (2025)
di: Oppelstrup, Tomas, et al.
Pubblicazione: (2025)
Toward Portable GPU Performance: Julia Recursive Implementation of TRMM and TRSM
di: Carrica, Vicki, et al.
Pubblicazione: (2025)
di: Carrica, Vicki, et al.
Pubblicazione: (2025)
Enabling GPU Portability into the Numba-JITed Monte Carlo Particle Transport Code MC/DC
di: Morgan, Joanna Piper, et al.
Pubblicazione: (2025)
di: Morgan, Joanna Piper, et al.
Pubblicazione: (2025)
GPU Acceleration and Portability of the TRIMEG Code for Gyrokinetic Plasma Simulations using OpenMP
di: Daneri, Giorgio
Pubblicazione: (2026)
di: Daneri, Giorgio
Pubblicazione: (2026)
Is RISC-V Ready for Machine Learning? Portable Gaussian Processes Using Asynchronous Tasks
di: Strack, Alexander, et al.
Pubblicazione: (2026)
di: Strack, Alexander, et al.
Pubblicazione: (2026)
Software for Exascale Computing - SPPEXA 2016-2019
Pubblicazione: (2020)
Pubblicazione: (2020)
Implementing Multi-GPU Scientific Computing Miniapps Across Performance Portable Frameworks
di: Villalobos, Johansell, et al.
Pubblicazione: (2025)
di: Villalobos, Johansell, et al.
Pubblicazione: (2025)
Data-Driven Analysis to Understand GPU Hardware Resource Usage of Optimizations
di: Islam, Tanzima Z., et al.
Pubblicazione: (2024)
di: Islam, Tanzima Z., et al.
Pubblicazione: (2024)
Exascale In-situ visualization for Astronomy & Cosmology
di: Tuccari, Nicola, et al.
Pubblicazione: (2025)
di: Tuccari, Nicola, et al.
Pubblicazione: (2025)
GPU-Portable Real-Space Density Functional Theory Implementation on Unified-Memory Architectures
di: Ito, Atsushi M.
Pubblicazione: (2025)
di: Ito, Atsushi M.
Pubblicazione: (2025)
Performant Unified GPU Kernels for Portable Singular Value Computation Across Hardware and Precision
di: Ringoot, Evelyne, et al.
Pubblicazione: (2025)
di: Ringoot, Evelyne, et al.
Pubblicazione: (2025)
Multi-Partner Project: Multi-GPU Performance Portability Analysis for CFD Simulations at Scale
di: Eleftherakis, Panagiotis-Eleftherios, et al.
Pubblicazione: (2026)
di: Eleftherakis, Panagiotis-Eleftherios, et al.
Pubblicazione: (2026)
Toward Reproducible and Standardized Computer Architecture Simulation with gem5
di: Pai, Kunal, et al.
Pubblicazione: (2025)
di: Pai, Kunal, et al.
Pubblicazione: (2025)
Minos: Systematically Classifying Performance and Power Characteristics of GPU Workloads on HPC Clusters
di: Jain, Rutwik, et al.
Pubblicazione: (2026)
di: Jain, Rutwik, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Mapping Spiking Neural Networks to Heterogeneous Crossbar Architectures using Integer Linear Programming
di: Pohl, Devin, et al.
Pubblicazione: (2025) -
Classic and Quantum Task-Based Intelligent Runtime for QIRs Running on Multiple QPUs
di: Miniskar, Narasinga Rao, et al.
Pubblicazione: (2026) -
Q-IRIS: The Evolution of the IRIS Task-Based Runtime to Enable Classical-Quantum Workflows
di: Miniskar, Narasinga Rao, et al.
Pubblicazione: (2025) -
Mojo: MLIR-Based Performance-Portable HPC Science Kernels on GPUs for the Python Ecosystem
di: Godoy, William F., et al.
Pubblicazione: (2025) -
Large language model evaluation for high‐performance computing software development
di: William F. Godoy, et al.
Pubblicazione: (2024)