A Continuous Benchmarking Infrastructure for High-Performance Computing Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Alt, Christoph, Lanser, Martin, Plewinski, Jonas, Janki, Atin, Klawonn, Axel, Köstler, Harald, Selzer, Michael, Rüde, Ulrich |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performance analysis of the free surface lattice Boltzmann implementation in waLBerla
by: Jonas Plewinski, et al.
Published: (2024)
by: Jonas Plewinski, et al.
Published: (2024)
Architecture Specific Generation of Large Scale Lattice Boltzmann Methods for Sparse Complex Geometries
by: Suffa, Philipp, et al.
Published: (2024)
by: Suffa, Philipp, et al.
Published: (2024)
AI Application Benchmarking: Power-Aware Performance Analysis for Vision and Language Models
by: Mayr, Martin, et al.
Published: (2026)
by: Mayr, Martin, et al.
Published: (2026)
Large-scale Multigrid with Adaptive Galerkin Coarsening
by: Böhm, Fabian, et al.
Published: (2025)
by: Böhm, Fabian, et al.
Published: (2025)
Towards Automated Algebraic Multigrid Preconditioner Design Using Genetic Programming for Large-Scale Laser Beam Welding Simulations
by: Parthasarathy, Dinesh, et al.
Published: (2024)
by: Parthasarathy, Dinesh, et al.
Published: (2024)
Nonlinear Monolithic Two-Level Schwarz Methods for the Navier-Stokes Equations
by: Klawonn, Axel, et al.
Published: (2024)
by: Klawonn, Axel, et al.
Published: (2024)
SProBench: Stream Processing Benchmark for High Performance Computing Infrastructure
by: Kulkarni, Apurv Deepak, et al.
Published: (2025)
by: Kulkarni, Apurv Deepak, et al.
Published: (2025)
Highly Scalable Two-level Monolithic Overlapping Schwarz Preconditioners for Thermo-elastoplastic Laser Beam Welding Problems
by: Bevilacqua, Tommaso, et al.
Published: (2025)
by: Bevilacqua, Tommaso, et al.
Published: (2025)
A Domain Decomposition-Based CNN-DNN Architecture for Model Parallel Training Applied to Image Recognition Problems
by: Klawonn, Axel, et al.
Published: (2023)
by: Klawonn, Axel, et al.
Published: (2023)
Model Parallel Training and Transfer Learning for Convolutional Neural Networks by Domain Decomposition
by: Klawonn, Axel, et al.
Published: (2024)
by: Klawonn, Axel, et al.
Published: (2024)
Domain-decomposed image classification algorithms using linear discriminant analysis and convolutional neural networks
by: Klawonn, Axel, et al.
Published: (2024)
by: Klawonn, Axel, et al.
Published: (2024)
Adaptive and frugal BDDC coarse spaces for virtual element discretizations of a Stokes problem with heterogeneous viscosity
by: Bevilacqua, Tommaso, et al.
Published: (2024)
by: Bevilacqua, Tommaso, et al.
Published: (2024)
Two-level nonlinear Schwarz methods - a parallel implementation with application to nonlinear elasticity and incompressible flow problems
by: Ho, Kyrill, et al.
Published: (2026)
by: Ho, Kyrill, et al.
Published: (2026)
Computational homogenization for aerogel-like polydisperse open-porous materials using neural network--based surrogate models on the microscale
by: Klawonn, Axel, et al.
Published: (2024)
by: Klawonn, Axel, et al.
Published: (2024)
OSCAR-P and aMLLibrary: Profiling and Predicting the Performance of FaaS-based Applications in Computing Continua
by: Sala, Roberto, et al.
Published: (2024)
by: Sala, Roberto, et al.
Published: (2024)
Towards Computational Performance Engineering for Unsupervised Concept Drift Detection -- Complexities, Benchmarking, Performance Analysis
by: Werner, Elias, et al.
Published: (2023)
by: Werner, Elias, et al.
Published: (2023)
Nonlinear Two-Level Schwarz Methods: A Parallel Implementation in FROSch
by: Heinlein, Alexander, et al.
Published: (2024)
by: Heinlein, Alexander, et al.
Published: (2024)
Performance of Confidential Computing GPUs
by: Ibarra, Antonio Martínez, et al.
Published: (2025)
by: Ibarra, Antonio Martínez, et al.
Published: (2025)
Efficiency and scalability of fully-resolved fluid-particle simulations on heterogeneous CPU-GPU architectures
by: Kemmler, Samuel, et al.
Published: (2023)
by: Kemmler, Samuel, et al.
Published: (2023)
Memshare: Memory Sharing for Multicore Computation in R with an Application to Feature Selection by Mutual Information using PDE
by: Thrun, Michael C., et al.
Published: (2025)
by: Thrun, Michael C., et al.
Published: (2025)
Performance Characterization of Containers in Edge Computing
by: Gupta, Ragini, et al.
Published: (2025)
by: Gupta, Ragini, et al.
Published: (2025)
Towards a parallel Schwarz solver framework for virtual elements using GDSW coarse spaces
by: Bevilacqua, Tommaso, et al.
Published: (2025)
by: Bevilacqua, Tommaso, et al.
Published: (2025)
Monolithic overlapping Schwarz preconditioners for nonlinear finite element simulations of laser beam welding processes
by: Bevilacqua, Tommaso, et al.
Published: (2024)
by: Bevilacqua, Tommaso, et al.
Published: (2024)
gpu tracker: Python Package for Tracking and Profiling GPU and Other Hardware Utilization in Both Desktop and High-Performance Computing Environments
by: Huckvale, Erik D., et al.
Published: (2024)
by: Huckvale, Erik D., et al.
Published: (2024)
Performance Characterization and Optimizations of Traditional ML Applications
by: Kumar, Harsh, et al.
Published: (2024)
by: Kumar, Harsh, et al.
Published: (2024)
High Performance Matrix Multiplication
by: Davis, Ethan
Published: (2025)
by: Davis, Ethan
Published: (2025)
Meta-Metrics and Best Practices for System-Level Inference Performance Benchmarking
by: Salaria, Shweta, et al.
Published: (2025)
by: Salaria, Shweta, et al.
Published: (2025)
Resource Allocation Influence on Application Performance in Sliced Testbeds
by: Moreira, Rodrigo, et al.
Published: (2024)
by: Moreira, Rodrigo, et al.
Published: (2024)
Evaluating the Performance of the DeepSeek Model in Confidential Computing Environment
by: Dong, Ben, et al.
Published: (2025)
by: Dong, Ben, et al.
Published: (2025)
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
by: Andersson, Måns I., et al.
Published: (2025)
by: Andersson, Måns I., et al.
Published: (2025)
The Multiserver-Job Stochastic Recurrence Equation for Cloud Computing Performance Evaluation
by: Baccelli, Francois, et al.
Published: (2026)
by: Baccelli, Francois, et al.
Published: (2026)
LCS.jl: A High-Performance, Multi-Platform Computational Model in Julia for Turbulent Particle-Laden Flows
by: Tominaga, Taketo, et al.
Published: (2026)
by: Tominaga, Taketo, et al.
Published: (2026)
Hierarchical Analyses Applied to Computer System Performance: Review and Call for Further Studies
by: Thomasian, Alexander
Published: (2024)
by: Thomasian, Alexander
Published: (2024)
Performance Optimization of 3D Stencil Computation on ARM Scalable Vector Extension
by: Chen, Hongguang
Published: (2025)
by: Chen, Hongguang
Published: (2025)
Accurate Performance Modeling And Uncertainty Analysis of Lossy Compression in Scientific Applications
by: Liu, Youyuan, et al.
Published: (2024)
by: Liu, Youyuan, et al.
Published: (2024)
SysOM-AI: Continuous Cross-Layer Performance Diagnosis for Production AI Training
by: Zheng, Yusheng, et al.
Published: (2026)
by: Zheng, Yusheng, et al.
Published: (2026)
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
by: Zhu, Jianwei, et al.
Published: (2024)
by: Zhu, Jianwei, et al.
Published: (2024)
Evaluating HPC-Style CPU Performance and Cost in Virtualized Cloud Infrastructures
by: Tharwani, Jay, et al.
Published: (2025)
by: Tharwani, Jay, et al.
Published: (2025)
The SAP Cloud Infrastructure Dataset: A Reality Check of Scheduling and Placement of VMs in Cloud Computing
by: Uhlig, Arno, et al.
Published: (2025)
by: Uhlig, Arno, et al.
Published: (2025)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
by: Ramesh, Risshab Srinivas
Published: (2024)
by: Ramesh, Risshab Srinivas
Published: (2024)
Similar Items
-
Performance analysis of the free surface lattice Boltzmann implementation in waLBerla
by: Jonas Plewinski, et al.
Published: (2024) -
Architecture Specific Generation of Large Scale Lattice Boltzmann Methods for Sparse Complex Geometries
by: Suffa, Philipp, et al.
Published: (2024) -
AI Application Benchmarking: Power-Aware Performance Analysis for Vision and Language Models
by: Mayr, Martin, et al.
Published: (2026) -
Large-scale Multigrid with Adaptive Galerkin Coarsening
by: Böhm, Fabian, et al.
Published: (2025) -
Towards Automated Algebraic Multigrid Preconditioner Design Using Genetic Programming for Large-Scale Laser Beam Welding Simulations
by: Parthasarathy, Dinesh, et al.
Published: (2024)