Building an Accelerated OpenFOAM Proof-of-Concept Application using Modern C++
Fuente:
arXiv
Saved in:
| Main Authors: | Malenza, Giulio, Stabile, Giovanni, Spiga, Filippo, Birke, Robert, Aldinucci, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Annotation-guided AoS-to-SoA conversions and GPU offloading with data views in C++
by: Radtke, Pawel K., et al.
Published: (2025)
by: Radtke, Pawel K., et al.
Published: (2025)
rcpptimer: Rcpp Tic-Toc Timer with OpenMP Support
by: Berrisch, Jonathan
Published: (2025)
by: Berrisch, Jonathan
Published: (2025)
RAO-SS: A Prototype of Run-time Auto-tuning Facility for Sparse Direct Solvers
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
A Test for FLOPs as a Discriminant for Linear Algebra Algorithms
by: Sankaran, Aravind, et al.
Published: (2022)
by: Sankaran, Aravind, et al.
Published: (2022)
Towards a Higher Roofline for Matrix-Vector Multiplication in Matrix-Free HOSFEM
by: Cao, Zijian, et al.
Published: (2025)
by: Cao, Zijian, et al.
Published: (2025)
Faster Base64 Encoding and Decoding Using AVX2 Instructions
by: Muła, Wojciech, et al.
Published: (2017)
by: Muła, Wojciech, et al.
Published: (2017)
GoldbachGPU: An Open Source GPU-Accelerated Framework for Verification of Goldbach's Conjecture
by: Llorente-Saguer, Isaac
Published: (2026)
by: Llorente-Saguer, Isaac
Published: (2026)
Performant Automatic BLAS Offloading on Unified Memory Architecture with OpenMP First-Touch Style Data Movement
by: Li, Junjie
Published: (2024)
by: Li, Junjie
Published: (2024)
MapReplay: Trace-Driven Benchmark Generation for Java HashMap
by: Schiavio, Filippo, et al.
Published: (2026)
by: Schiavio, Filippo, et al.
Published: (2026)
Accelerating High-Order Finite Element Simulations at Extreme Scale with FP64 Tensor Cores
by: Tu, Jiqun, et al.
Published: (2026)
by: Tu, Jiqun, et al.
Published: (2026)
Testing the Unknown: A Framework for OpenMP Testing via Random Program Generation
by: Laguna, Ignacio, et al.
Published: (2024)
by: Laguna, Ignacio, et al.
Published: (2024)
DGEMM without FP64 Arithmetic - Using FP64 Emulation and FP8 Tensor Cores with Ozaki Scheme
by: Mukunoki, Daichi
Published: (2025)
by: Mukunoki, Daichi
Published: (2025)
Hierarchical Recursive Precision for Accelerating Symmetric Linear Solves on MXUs
by: Carrica, Vicki, et al.
Published: (2026)
by: Carrica, Vicki, et al.
Published: (2026)
pyGinkgo: A Sparse Linear Algebra Operator Framework for Python
by: Tuteja, Keshvi, et al.
Published: (2025)
by: Tuteja, Keshvi, et al.
Published: (2025)
Portability of Fortran's `do concurrent' on GPUs
by: Caplan, Ronald M., et al.
Published: (2024)
by: Caplan, Ronald M., et al.
Published: (2024)
Acceleration of Tensor-Product Operations with Tensor Cores
by: Cui, Cu
Published: (2024)
by: Cui, Cu
Published: (2024)
Exploring energy consumption of AI frameworks on a 64-core RV64 Server CPU
by: Malenza, Giulio, et al.
Published: (2025)
by: Malenza, Giulio, et al.
Published: (2025)
Easy Acceleration with Distributed Arrays
by: Kepner, Jeremy, et al.
Published: (2025)
by: Kepner, Jeremy, et al.
Published: (2025)
An Empirical Study on the Performance and Energy Usage of Compiled Python Code
by: Stoico, Vincenzo, et al.
Published: (2025)
by: Stoico, Vincenzo, et al.
Published: (2025)
Input-Gen: Guided Generation of Stateful Inputs for Testing, Tuning, and Training
by: Ivanov, Ivan R., et al.
Published: (2024)
by: Ivanov, Ivan R., et al.
Published: (2024)
Runtime Verification on Abstract Finite State Models
by: Jevitha, KP, et al.
Published: (2024)
by: Jevitha, KP, et al.
Published: (2024)
Stencil-Lifting: Hierarchical Recursive Lifting System for Extracting Summary of Stencil Kernel in Legacy Codes
by: Li, Mingyi, et al.
Published: (2025)
by: Li, Mingyi, et al.
Published: (2025)
Xabclib:A Fully Auto-tuned Sparse Iterative Solver
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
Automated MPI-X code generation for scalable finite-difference solvers
by: Bisbas, George, et al.
Published: (2023)
by: Bisbas, George, et al.
Published: (2023)
On the energy efficiency of sparse matrix computations on multi-GPU clusters
by: Bernaschi, Massimo, et al.
Published: (2025)
by: Bernaschi, Massimo, et al.
Published: (2025)
NApy: Efficient Statistics in Python for Large-Scale Heterogeneous Data with Enhanced Support for Missing Data
by: Woller, Fabian, et al.
Published: (2025)
by: Woller, Fabian, et al.
Published: (2025)
A Communication Avoiding and Reducing Algorithm for Symmetric Eigenproblem for Very Small Matrices
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
Beating vDSP: A 138 GFLOPS Radix-8 Stockham FFT on Apple Silicon via Two-Tier Register-Threadgroup Memory Decomposition
by: Bergach, Mohamed Amine
Published: (2026)
by: Bergach, Mohamed Amine
Published: (2026)
Performance measurements of modern Fortran MPI applications with Score-P
by: Corbin, Gregor
Published: (2025)
by: Corbin, Gregor
Published: (2025)
Black-Scholes Option Pricing on Intel CPUs and GPUs: Implementation on SYCL and Optimization Techniques
by: Panova, Elena, et al.
Published: (2022)
by: Panova, Elena, et al.
Published: (2022)
Toward Capturing Genetic Epistasis From Multivariate Genome-Wide Association Studies Using Mixed-Precision Kernel Ridge Regression
by: Ltaief, Hatem, et al.
Published: (2024)
by: Ltaief, Hatem, et al.
Published: (2024)
Improving the Graph Challenge Reference Implementation
by: Voloshchuk, Inna, et al.
Published: (2026)
by: Voloshchuk, Inna, et al.
Published: (2026)
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
by: Kanakagiri, Raghavendra, et al.
Published: (2023)
by: Kanakagiri, Raghavendra, et al.
Published: (2023)
An Empirical Study on Method-Level Performance Evolution in Open-Source Java Projects
by: Shahedi, Kaveh, et al.
Published: (2025)
by: Shahedi, Kaveh, et al.
Published: (2025)
Anonymized Network Sensing Graph Challenge
by: Jananthan, Hayden, et al.
Published: (2024)
by: Jananthan, Hayden, et al.
Published: (2024)
Who Wins the Race? (R Vs Python) - An Exploratory Study on Energy Consumption of Machine Learning Algorithms
by: Chattaraj, Rajrupa, et al.
Published: (2025)
by: Chattaraj, Rajrupa, et al.
Published: (2025)
Library Liberation: Competitive Performance Matmul Through Compiler-composed Nanokernels
by: Thangamani, Arun, et al.
Published: (2025)
by: Thangamani, Arun, et al.
Published: (2025)
Open Source Prover in the Attic
by: Kovács, Zoltán, et al.
Published: (2024)
by: Kovács, Zoltán, et al.
Published: (2024)
FaaSter Troubleshooting -- Evaluating Distributed Tracing Approaches for Serverless Applications
by: Borges, Maria C., et al.
Published: (2021)
by: Borges, Maria C., et al.
Published: (2021)
gigiProfiler: Diagnosing Performance Issues by Uncovering Application Resource Bottlenecks
by: Hu, Yigong, et al.
Published: (2025)
by: Hu, Yigong, et al.
Published: (2025)
Similar Items
-
Annotation-guided AoS-to-SoA conversions and GPU offloading with data views in C++
by: Radtke, Pawel K., et al.
Published: (2025) -
rcpptimer: Rcpp Tic-Toc Timer with OpenMP Support
by: Berrisch, Jonathan
Published: (2025) -
RAO-SS: A Prototype of Run-time Auto-tuning Facility for Sparse Direct Solvers
by: Katagiri, Takahiro, et al.
Published: (2024) -
A Test for FLOPs as a Discriminant for Linear Algebra Algorithms
by: Sankaran, Aravind, et al.
Published: (2022) -
Towards a Higher Roofline for Matrix-Vector Multiplication in Matrix-Free HOSFEM
by: Cao, Zijian, et al.
Published: (2025)