Machine Learning-driven Autotuning of Graphics Processing Unit Accelerated Computational Fluid Dynamics for Enhanced Performance
Fuente:
arXiv
Guardado en:
| Autores principales: | Xue, Weicheng, Roy, Christohper John |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Characterizing Machine Learning Force Fields as Emerging Molecular Dynamics Workloads on Graphics Processing Units
por: De Alwis, Udari, et al.
Publicado: (2026)
por: De Alwis, Udari, et al.
Publicado: (2026)
Integrating ytopt and libEnsemble to Autotune OpenMC
por: Wu, Xingfu, et al.
Publicado: (2024)
por: Wu, Xingfu, et al.
Publicado: (2024)
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
por: Andersson, Måns I., et al.
Publicado: (2025)
por: Andersson, Måns I., et al.
Publicado: (2025)
Accelerating Machine Learning Queries with Linear Algebra Query Processing
por: Sun, Wenbo, et al.
Publicado: (2023)
por: Sun, Wenbo, et al.
Publicado: (2023)
Understanding the Impact of Synchronous, Asynchronous, and Hybrid In-Situ Techniques in Computational Fluid Dynamics Applications
por: Ju, Yi, et al.
Publicado: (2024)
por: Ju, Yi, et al.
Publicado: (2024)
HPC Application Parameter Autotuning on Edge Devices: A Bandit Learning Approach
por: Hossain, Abrar, et al.
Publicado: (2025)
por: Hossain, Abrar, et al.
Publicado: (2025)
Optimizing Near Field Computation in the MLFMA Algorithm with Data Redundancy and Performance Modeling on a Single GPU
por: Sadeghi, Morteza, et al.
Publicado: (2024)
por: Sadeghi, Morteza, et al.
Publicado: (2024)
SProBench: Stream Processing Benchmark for High Performance Computing Infrastructure
por: Kulkarni, Apurv Deepak, et al.
Publicado: (2025)
por: Kulkarni, Apurv Deepak, et al.
Publicado: (2025)
Enhanced Scalability in Assessing Quantum Integer Factorization Performance
por: Lee, Junseo
Publicado: (2023)
por: Lee, Junseo
Publicado: (2023)
An Autotuning-based Optimization Framework for Mixed-kernel SVM Classifications in Smart Pixel Datasets and Heterojunction Transistors
por: Wu, Xingfu, et al.
Publicado: (2024)
por: Wu, Xingfu, et al.
Publicado: (2024)
A Precision Emulation Approach to the GPU Acceleration of Ab Initio Electronic Structure Calculations
por: Liu, Hang, et al.
Publicado: (2026)
por: Liu, Hang, et al.
Publicado: (2026)
Performance Implications of Multi-Chiplet Neural Processing Units on Autonomous Driving Perception
por: Odema, Mohanad, et al.
Publicado: (2024)
por: Odema, Mohanad, et al.
Publicado: (2024)
Accelerating Particle-in-Cell Monte Carlo Simulations with MPI, OpenMP/OpenACC and Asynchronous Multi-GPU Programming
por: Williams, Jeremy J., et al.
Publicado: (2024)
por: Williams, Jeremy J., et al.
Publicado: (2024)
A Machine Learning accelerated geophysical fluid solver
por: Bai, Yang
Publicado: (2026)
por: Bai, Yang
Publicado: (2026)
Towards Universal Performance Modeling for Machine Learning Training on Multi-GPU Platforms
por: Lin, Zhongyi, et al.
Publicado: (2024)
por: Lin, Zhongyi, et al.
Publicado: (2024)
Accelerating the Tesseract Decoder for Quantum Error Correction
por: Grbic, Dragana, et al.
Publicado: (2026)
por: Grbic, Dragana, et al.
Publicado: (2026)
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
por: Lysenstøen, Christian
Publicado: (2026)
por: Lysenstøen, Christian
Publicado: (2026)
Non-Asymptotic Performance Analysis of DOA Estimation Based on Real-Valued Root-MUSIC
por: Liu, Junyang, et al.
Publicado: (2025)
por: Liu, Junyang, et al.
Publicado: (2025)
TINA: Acceleration of Non-NN Signal Processing Algorithms Using NN Accelerators
por: Boerkamp, Christiaan, et al.
Publicado: (2024)
por: Boerkamp, Christiaan, et al.
Publicado: (2024)
Advanced Scheduling Strategies for Distributed Quantum Computing Jobs
por: Ni, Gongyu, et al.
Publicado: (2026)
por: Ni, Gongyu, et al.
Publicado: (2026)
GPU-Accelerated Parallel Selected Inversion for Structured Matrices Using sTiles
por: Fattah, Esmail Abdul, et al.
Publicado: (2025)
por: Fattah, Esmail Abdul, et al.
Publicado: (2025)
A Data-driven ML Approach for Maximizing Performance in LLM-Adapter Serving
por: Agullo, Ferran, et al.
Publicado: (2025)
por: Agullo, Ferran, et al.
Publicado: (2025)
Multi-GPU Hybrid Particle-in-Cell Monte Carlo Simulations for Exascale Computing Systems
por: Williams, Jeremy J., et al.
Publicado: (2026)
por: Williams, Jeremy J., et al.
Publicado: (2026)
Characterizing the Performance of the Implicit Massively Parallel Particle-in-Cell iPIC3D Code
por: Williams, Jeremy J., et al.
Publicado: (2024)
por: Williams, Jeremy J., et al.
Publicado: (2024)
Solving Combinatorial Optimization Problems on a Photonic Quantum Computer
por: Slysz, Mateusz, et al.
Publicado: (2024)
por: Slysz, Mateusz, et al.
Publicado: (2024)
Performance of Confidential Computing GPUs
por: Ibarra, Antonio Martínez, et al.
Publicado: (2025)
por: Ibarra, Antonio Martínez, et al.
Publicado: (2025)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
por: Ramesh, Risshab Srinivas
Publicado: (2024)
por: Ramesh, Risshab Srinivas
Publicado: (2024)
Scheduling in Quantum Satellite Networks: Fairness and Performance Optimization
por: Dikshit, Ashutosh Jayant, et al.
Publicado: (2025)
por: Dikshit, Ashutosh Jayant, et al.
Publicado: (2025)
Cost-Performance Evaluation of General Compute Instances: AWS, Azure, GCP, and OCI
por: Tharwani, Jay, et al.
Publicado: (2024)
por: Tharwani, Jay, et al.
Publicado: (2024)
Integrating High Performance In-Memory Data Streaming and In-Situ Visualization in Hybrid MPI+OpenMP PIC MC Simulations Towards Exascale
por: Williams, Jeremy J., et al.
Publicado: (2025)
por: Williams, Jeremy J., et al.
Publicado: (2025)
Performance Characterization of Containers in Edge Computing
por: Gupta, Ragini, et al.
Publicado: (2025)
por: Gupta, Ragini, et al.
Publicado: (2025)
Phantora: Maximizing Code Reuse in Simulation-based Machine Learning System Performance Estimation
por: Qin, Jianxing, et al.
Publicado: (2025)
por: Qin, Jianxing, et al.
Publicado: (2025)
Performance metrics for the continuous distribution of entanglement in multi-user quantum networks
por: Iñesta, Álvaro G., et al.
Publicado: (2023)
por: Iñesta, Álvaro G., et al.
Publicado: (2023)
Enhancing Traffic Sign Recognition On The Performance Based On Yolov8
por: Ibrahim, Baba, et al.
Publicado: (2025)
por: Ibrahim, Baba, et al.
Publicado: (2025)
Performance Optimization in Stream Processing Systems: Experiment-Driven Configuration Tuning for Kafka Streams
por: Chen, David, et al.
Publicado: (2026)
por: Chen, David, et al.
Publicado: (2026)
Efficient Transpilation of OpenQASM 3.0 Dynamic Circuits to CUDA-Q: Performance and Expressiveness Advantages
por: Kulkarni, Vinooth, et al.
Publicado: (2026)
por: Kulkarni, Vinooth, et al.
Publicado: (2026)
Deciphering boundary layer dynamics in high-Rayleigh-number convection using 3360 GPUs and a high-scaling in-situ workflow
por: Bode, Mathis, et al.
Publicado: (2025)
por: Bode, Mathis, et al.
Publicado: (2025)
Towards Run Time Estimation of the Gaussian Chemistry Code for SEAGrid Science Gateway
por: Beltre, Angel, et al.
Publicado: (2019)
por: Beltre, Angel, et al.
Publicado: (2019)
nekCRF: A next generation high-order reactive low Mach flow solver for direct numerical simulations
por: Kerkemeier, Stefan, et al.
Publicado: (2024)
por: Kerkemeier, Stefan, et al.
Publicado: (2024)
An Analysis of Performance Bottlenecks in MRI Pre-Processing
por: Dugré, Mathieu, et al.
Publicado: (2024)
por: Dugré, Mathieu, et al.
Publicado: (2024)
Ejemplares similares
-
Characterizing Machine Learning Force Fields as Emerging Molecular Dynamics Workloads on Graphics Processing Units
por: De Alwis, Udari, et al.
Publicado: (2026) -
Integrating ytopt and libEnsemble to Autotune OpenMC
por: Wu, Xingfu, et al.
Publicado: (2024) -
Portable High-Performance Kernel Generation for a Computational Fluid Dynamics Code with DaCe
por: Andersson, Måns I., et al.
Publicado: (2025) -
Accelerating Machine Learning Queries with Linear Algebra Query Processing
por: Sun, Wenbo, et al.
Publicado: (2023) -
Understanding the Impact of Synchronous, Asynchronous, and Hybrid In-Situ Techniques in Computational Fluid Dynamics Applications
por: Ju, Yi, et al.
Publicado: (2024)