Evaluating the impact of the L3 cache size of AMD EPYC CPUs on the performance of CFD applications
Fuente:
arXiv
Saved in:
| Main Authors: | Lawenda, Marcin, Szustak, Łukasz, Környei, László, Galeazzo, Flavio Cesar Cunha, Bratek, Paweł |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Profiling and optimization of multi-card GPU machine learning jobs
by: Lawenda, Marcin, et al.
Published: (2025)
by: Lawenda, Marcin, et al.
Published: (2025)
Efficient allocation of image recognition and LLM tasks on multi-GPU system
by: Lawenda, Marcin, et al.
Published: (2025)
by: Lawenda, Marcin, et al.
Published: (2025)
Profiling and Optimization of Multicard GPU Machine Learning Jobs
by: Marcin Lawenda, et al.
Published: (2025)
by: Marcin Lawenda, et al.
Published: (2025)
A high-performance and portable implementation of the SISSO method for CPUs and GPUs
by: Eibl, Sebastian, et al.
Published: (2025)
by: Eibl, Sebastian, et al.
Published: (2025)
A dynamic parallel method for performance optimization on hybrid CPUs
by: Yu, Luo, et al.
Published: (2024)
by: Yu, Luo, et al.
Published: (2024)
HiDALGO2 D3.2 Scalability, Optimization and Co-Design Activities
by: NIKAS, KONSTANTINOS, et al.
Published: (2025)
by: NIKAS, KONSTANTINOS, et al.
Published: (2025)
Energy efficiency of cache eviction algorithms for Zipf distributed objects
by: Sziklay, Emese, et al.
Published: (2025)
by: Sziklay, Emese, et al.
Published: (2025)
Efficient GPU implementation of randomized SVD and its applications
by: Struski, Łukasz, et al.
Published: (2021)
by: Struski, Łukasz, et al.
Published: (2021)
The Bicameral Cache: a split cache for vector architectures
by: Rebolledo, Susana, et al.
Published: (2024)
by: Rebolledo, Susana, et al.
Published: (2024)
AMD MI300X GPU Performance Analysis
by: Ambati, Chandrish, et al.
Published: (2025)
by: Ambati, Chandrish, et al.
Published: (2025)
Explainable Port Mapping Inference with Sparse Performance Counters for AMD's Zen Architectures
by: Ritter, Fabian, et al.
Published: (2024)
by: Ritter, Fabian, et al.
Published: (2024)
Comparison of Vectorization Capabilities of Different Compilers for X86 and ARM CPUs
by: Sakib, Nazmus, et al.
Published: (2025)
by: Sakib, Nazmus, et al.
Published: (2025)
Towards High-Performance and Portable Molecular Docking on CPUs through Vectorization
by: Accordi, Gianmarco, et al.
Published: (2025)
by: Accordi, Gianmarco, et al.
Published: (2025)
Performance Analysis of HPC applications on the Aurora Supercomputer: Exploring the Impact of HBM-Enabled Intel Xeon Max CPUs
by: Ibeid, Huda, et al.
Published: (2025)
by: Ibeid, Huda, et al.
Published: (2025)
Impact of Data-Oriented and Object-Oriented Design on Performance and Cache Utilization with Artificial Intelligence Algorithms in Multi-Threaded CPUs
by: Arantes, Gabriel M., et al.
Published: (2025)
by: Arantes, Gabriel M., et al.
Published: (2025)
Microarchitectural comparison and in-core modeling of state-of-the-art CPUs: Grace, Sapphire Rapids, and Genoa
by: Laukemann, Jan, et al.
Published: (2024)
by: Laukemann, Jan, et al.
Published: (2024)
CloverLeaf on Intel Multi-Core CPUs: A Case Study in Write-Allocate Evasion
by: Laukemann, Jan, et al.
Published: (2023)
by: Laukemann, Jan, et al.
Published: (2023)
THEAS: Efficient Power Management in Multi-Core CPUs via Cache-Aware Resource Scheduling
by: Muhammad, Said, et al.
Published: (2025)
by: Muhammad, Said, et al.
Published: (2025)
Introducing MareNostrum5: A European pre-exascale energy-efficient system designed to serve a broad spectrum of scientific workloads
by: Banchelli, Fabio, et al.
Published: (2025)
by: Banchelli, Fabio, et al.
Published: (2025)
SparAMX: Accelerating Compressed LLMs Token Generation on AMX-powered CPUs
by: AbouElhamayed, Ahmed F., et al.
Published: (2025)
by: AbouElhamayed, Ahmed F., et al.
Published: (2025)
Improved vectorization of OpenCV algorithms for RISC-V CPUs
by: Volokitin, V. D., et al.
Published: (2023)
by: Volokitin, V. D., et al.
Published: (2023)
Black-Scholes Option Pricing on Intel CPUs and GPUs: Implementation on SYCL and Optimization Techniques
by: Panova, Elena, et al.
Published: (2022)
by: Panova, Elena, et al.
Published: (2022)
Preliminary report: Initial evaluation of StdPar implementations on AMD GPUs for HPC
by: Lin, Wei-Chen, et al.
Published: (2024)
by: Lin, Wei-Chen, et al.
Published: (2024)
A relação entre a «performance» social e a «performance» económico-financeira
by: Daniel Taborda
Published: (2007)
by: Daniel Taborda
Published: (2007)
On the effects of logical database design on database size, query complexity, query performance, and energy consumption
by: Taipalus, Toni
Published: (2025)
by: Taipalus, Toni
Published: (2025)
Dissecting CPU-GPU Unified Physical Memory on AMD MI300A APUs
by: Wahlgren, Jacob, et al.
Published: (2025)
by: Wahlgren, Jacob, et al.
Published: (2025)
Bringing Auto-tuning to HIP: Analysis of Tuning Impact and Difficulty on AMD and Nvidia GPUs
by: Lurati, Milo, et al.
Published: (2024)
by: Lurati, Milo, et al.
Published: (2024)
Seamless acceleration of Fortran intrinsics via AMD AI engines
by: Brown, Nick, et al.
Published: (2025)
by: Brown, Nick, et al.
Published: (2025)
Evaluating Emerging AI/ML Accelerators: IPU, RDU, and NVIDIA/AMD GPUs
by: Peng, Hongwu, et al.
Published: (2023)
by: Peng, Hongwu, et al.
Published: (2023)
The impact of alliances and internal R&D on the firm's innovation and financial performance
by: Fábio de Oliveira Paula
Published: (2018)
by: Fábio de Oliveira Paula
Published: (2018)
The impact of leadership style on employees' performance
by: DEMBELE, Christine, et al.
Published: (2025)
by: DEMBELE, Christine, et al.
Published: (2025)
A Priori Loop Nest Normalization: Automatic Loop Scheduling in Complex Applications
by: Trümper, Lukas, et al.
Published: (2024)
by: Trümper, Lukas, et al.
Published: (2024)
Modeling Tradeoffs between mobility, cost, and performance in Edge Computing
by: Waseem, Muhammad Danish, et al.
Published: (2026)
by: Waseem, Muhammad Danish, et al.
Published: (2026)
Reconsidering the performance of DEVS modeling and simulation environments using the DEVStone benchmark
by: Risco-Martín, José L., et al.
Published: (2024)
by: Risco-Martín, José L., et al.
Published: (2024)
Resource Allocation Influence on Application Performance in Sliced Testbeds
by: Moreira, Rodrigo, et al.
Published: (2024)
by: Moreira, Rodrigo, et al.
Published: (2024)
Towards self-optimization of publish/subscribe IoT systems using continuous performance monitoring
by: Djahafi, Mohammed, et al.
Published: (2024)
by: Djahafi, Mohammed, et al.
Published: (2024)
Inference performance evaluation for LLMs on edge devices with a novel benchmarking framework and metric
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
EDAN: Towards Understanding Memory Parallelism and Latency Sensitivity in HPC
by: Shen, Siyuan, et al.
Published: (2025)
by: Shen, Siyuan, et al.
Published: (2025)
Assessing the impact of e-marketing on business performance and customer relationship in livingstone district, Zambia
by: Mukosa, Grace, et al.
Published: (2025)
by: Mukosa, Grace, et al.
Published: (2025)
Performance Evaluation of Subroutines Call in PHP
by: Kalmukov, Yordan
Published: (2026)
by: Kalmukov, Yordan
Published: (2026)
Similar Items
-
Profiling and optimization of multi-card GPU machine learning jobs
by: Lawenda, Marcin, et al.
Published: (2025) -
Efficient allocation of image recognition and LLM tasks on multi-GPU system
by: Lawenda, Marcin, et al.
Published: (2025) -
Profiling and Optimization of Multicard GPU Machine Learning Jobs
by: Marcin Lawenda, et al.
Published: (2025) -
A high-performance and portable implementation of the SISSO method for CPUs and GPUs
by: Eibl, Sebastian, et al.
Published: (2025) -
A dynamic parallel method for performance optimization on hybrid CPUs
by: Yu, Luo, et al.
Published: (2024)