A high-performance and portable implementation of the SISSO method for CPUs and GPUs
Fuente:
arXiv
Salvato in:
| Autori principali: | Eibl, Sebastian, Yao, Yi, Scheffler, Matthias, Rampp, Markus, Ghiringhelli, Luca M., Purcell, Thomas A. R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A dynamic parallel method for performance optimization on hybrid CPUs
di: Yu, Luo, et al.
Pubblicazione: (2024)
di: Yu, Luo, et al.
Pubblicazione: (2024)
Black-Scholes Option Pricing on Intel CPUs and GPUs: Implementation on SYCL and Optimization Techniques
di: Panova, Elena, et al.
Pubblicazione: (2022)
di: Panova, Elena, et al.
Pubblicazione: (2022)
Evaluating the impact of the L3 cache size of AMD EPYC CPUs on the performance of CFD applications
di: Lawenda, Marcin, et al.
Pubblicazione: (2025)
di: Lawenda, Marcin, et al.
Pubblicazione: (2025)
Fast Entropy Decoding for Sparse MVM on GPUs
di: Schätzle, Emil, et al.
Pubblicazione: (2026)
di: Schätzle, Emil, et al.
Pubblicazione: (2026)
Opening the Black Box: Performance Estimation during Code Generation for GPUs
di: Ernst, Dominik, et al.
Pubblicazione: (2021)
di: Ernst, Dominik, et al.
Pubblicazione: (2021)
Performance of Confidential Computing GPUs
di: Ibarra, Antonio Martínez, et al.
Pubblicazione: (2025)
di: Ibarra, Antonio Martínez, et al.
Pubblicazione: (2025)
Preliminary report: Initial evaluation of StdPar implementations on AMD GPUs for HPC
di: Lin, Wei-Chen, et al.
Pubblicazione: (2024)
di: Lin, Wei-Chen, et al.
Pubblicazione: (2024)
CloverLeaf on Intel Multi-Core CPUs: A Case Study in Write-Allocate Evasion
di: Laukemann, Jan, et al.
Pubblicazione: (2023)
di: Laukemann, Jan, et al.
Pubblicazione: (2023)
Time is Not Compute: Scaling Laws for Wall-Clock Constrained Training on Consumer GPUs
di: Liu, Yi
Pubblicazione: (2026)
di: Liu, Yi
Pubblicazione: (2026)
Unveiling the Core of Materials Properties via SISSO and Sensitivity Analysis
di: Foppa, Lucas, et al.
Pubblicazione: (2026)
di: Foppa, Lucas, et al.
Pubblicazione: (2026)
Comparison of Vectorization Capabilities of Different Compilers for X86 and ARM CPUs
di: Sakib, Nazmus, et al.
Pubblicazione: (2025)
di: Sakib, Nazmus, et al.
Pubblicazione: (2025)
Impact of Data-Oriented and Object-Oriented Design on Performance and Cache Utilization with Artificial Intelligence Algorithms in Multi-Threaded CPUs
di: Arantes, Gabriel M., et al.
Pubblicazione: (2025)
di: Arantes, Gabriel M., et al.
Pubblicazione: (2025)
Towards High-Performance and Portable Molecular Docking on CPUs through Vectorization
di: Accordi, Gianmarco, et al.
Pubblicazione: (2025)
di: Accordi, Gianmarco, et al.
Pubblicazione: (2025)
A substantive theory on the implementation process of operational performance improvement methods
di: Darlan José Roman
Pubblicazione: (2017)
di: Darlan José Roman
Pubblicazione: (2017)
A substantive theory on the implementation process of operational performance improvement methods
di: Darlan José Roman
Pubblicazione: (2017)
di: Darlan José Roman
Pubblicazione: (2017)
Automated PMC-based Power Modeling Methodology for Modern Mobile GPUs
di: Dash, Pranab, et al.
Pubblicazione: (2024)
di: Dash, Pranab, et al.
Pubblicazione: (2024)
Application of performance portability solutions for GPUs and many-core CPUs to track reconstruction kernels
di: Kwok, Ka Hei Martin, et al.
Pubblicazione: (2024)
di: Kwok, Ka Hei Martin, et al.
Pubblicazione: (2024)
Benchmarking GPUs on SVBRDF Extractor Model
di: Kandel, Narayan, et al.
Pubblicazione: (2023)
di: Kandel, Narayan, et al.
Pubblicazione: (2023)
FRSZ2 for In-Register Block Compression Inside GMRES on GPUs
di: Grützmacher, Thomas, et al.
Pubblicazione: (2024)
di: Grützmacher, Thomas, et al.
Pubblicazione: (2024)
Improved vectorization of OpenCV algorithms for RISC-V CPUs
di: Volokitin, V. D., et al.
Pubblicazione: (2023)
di: Volokitin, V. D., et al.
Pubblicazione: (2023)
Microarchitectural comparison and in-core modeling of state-of-the-art CPUs: Grace, Sapphire Rapids, and Genoa
di: Laukemann, Jan, et al.
Pubblicazione: (2024)
di: Laukemann, Jan, et al.
Pubblicazione: (2024)
PM2Lat: Highly Accurate and Generalized Prediction of DNN Execution Latency on GPUs
di: Le, Truong-Thanh, et al.
Pubblicazione: (2026)
di: Le, Truong-Thanh, et al.
Pubblicazione: (2026)
Deciphering boundary layer dynamics in high-Rayleigh-number convection using 3360 GPUs and a high-scaling in-situ workflow
di: Bode, Mathis, et al.
Pubblicazione: (2025)
di: Bode, Mathis, et al.
Pubblicazione: (2025)
THEAS: Efficient Power Management in Multi-Core CPUs via Cache-Aware Resource Scheduling
di: Muhammad, Said, et al.
Pubblicazione: (2025)
di: Muhammad, Said, et al.
Pubblicazione: (2025)
SparAMX: Accelerating Compressed LLMs Token Generation on AMX-powered CPUs
di: AbouElhamayed, Ahmed F., et al.
Pubblicazione: (2025)
di: AbouElhamayed, Ahmed F., et al.
Pubblicazione: (2025)
How to Rent GPUs on a Budget
di: Li, Zhouzi, et al.
Pubblicazione: (2024)
di: Li, Zhouzi, et al.
Pubblicazione: (2024)
A relação entre a «performance» social e a «performance» económico-financeira
di: Daniel Taborda
Pubblicazione: (2007)
di: Daniel Taborda
Pubblicazione: (2007)
DF-GNN: Dynamic Fusion Framework for Attention Graph Neural Networks on GPUs
di: Liu, Jiahui, et al.
Pubblicazione: (2024)
di: Liu, Jiahui, et al.
Pubblicazione: (2024)
A Study of Performance Portability in Plasma Physics Simulations
di: Ruzicka, Josef, et al.
Pubblicazione: (2024)
di: Ruzicka, Josef, et al.
Pubblicazione: (2024)
A Comprehensive Analysis of Process Energy Consumption on Multi-Socket Systems with GPUs
di: León-Vega, Luis G., et al.
Pubblicazione: (2024)
di: León-Vega, Luis G., et al.
Pubblicazione: (2024)
Performance Analysis of HPC applications on the Aurora Supercomputer: Exploring the Impact of HBM-Enabled Intel Xeon Max CPUs
di: Ibeid, Huda, et al.
Pubblicazione: (2025)
di: Ibeid, Huda, et al.
Pubblicazione: (2025)
Pushing the Envelope of LLM Inference on AI-PC and Intel GPUs
di: Georganas, Evangelos, et al.
Pubblicazione: (2025)
di: Georganas, Evangelos, et al.
Pubblicazione: (2025)
Alya towards Exascale: Optimal OpenACC Performance of the Navier-Stokes Finite Element Assembly on GPUs
di: Owen, Herbert, et al.
Pubblicazione: (2024)
di: Owen, Herbert, et al.
Pubblicazione: (2024)
Characterizing and Understanding HGNN Training on GPUs
di: Han, Dengke, et al.
Pubblicazione: (2024)
di: Han, Dengke, et al.
Pubblicazione: (2024)
GROMACS Unplugged: How Power Capping and Frequency Shapes Performance on GPUs
di: Afzal, Ayesha, et al.
Pubblicazione: (2025)
di: Afzal, Ayesha, et al.
Pubblicazione: (2025)
Cyclic Data Streaming on GPUs for Short Range Stencils Applied to Molecular Dynamics
di: Rose, Martin, et al.
Pubblicazione: (2025)
di: Rose, Martin, et al.
Pubblicazione: (2025)
Accelerating AI Performance using Anderson Extrapolation on GPUs
di: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Pubblicazione: (2024)
di: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Pubblicazione: (2024)
Fusing Depthwise and Pointwise Convolutions for Efficient Inference on GPUs
di: Qararyah, Fareed, et al.
Pubblicazione: (2024)
di: Qararyah, Fareed, et al.
Pubblicazione: (2024)
Bringing Auto-tuning to HIP: Analysis of Tuning Impact and Difficulty on AMD and Nvidia GPUs
di: Lurati, Milo, et al.
Pubblicazione: (2024)
di: Lurati, Milo, et al.
Pubblicazione: (2024)
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
di: Zhu, Jianwei, et al.
Pubblicazione: (2024)
di: Zhu, Jianwei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A dynamic parallel method for performance optimization on hybrid CPUs
di: Yu, Luo, et al.
Pubblicazione: (2024) -
Black-Scholes Option Pricing on Intel CPUs and GPUs: Implementation on SYCL and Optimization Techniques
di: Panova, Elena, et al.
Pubblicazione: (2022) -
Evaluating the impact of the L3 cache size of AMD EPYC CPUs on the performance of CFD applications
di: Lawenda, Marcin, et al.
Pubblicazione: (2025) -
Fast Entropy Decoding for Sparse MVM on GPUs
di: Schätzle, Emil, et al.
Pubblicazione: (2026) -
Opening the Black Box: Performance Estimation during Code Generation for GPUs
di: Ernst, Dominik, et al.
Pubblicazione: (2021)