Efficient and scalable atmospheric dynamics simulations using non-conforming meshes
Fuente:
arXiv
Salvato in:
| Autori principali: | Orlando, Giuseppe, Benacchio, Tommaso, Bonaventura, Luca |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving the scalability of a high-order atmospheric dynamics solver based on the deal.II library
di: Orlando, Giuseppe, et al.
Pubblicazione: (2025)
di: Orlando, Giuseppe, et al.
Pubblicazione: (2025)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
di: Higgins, Andrew J., et al.
Pubblicazione: (2025)
di: Higgins, Andrew J., et al.
Pubblicazione: (2025)
Efficient Hardware Accelerator Based on Medium Granularity Dataflow for SpTRSV
di: Chen, Qian, et al.
Pubblicazione: (2024)
di: Chen, Qian, et al.
Pubblicazione: (2024)
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
di: Baioni, Paolo Joseph, et al.
Pubblicazione: (2024)
di: Baioni, Paolo Joseph, et al.
Pubblicazione: (2024)
Parallel simulation and adaptive mesh refinement for 3D elastostatic contact mechanics problems between deformable bodies
di: Epalle, Alexandre, et al.
Pubblicazione: (2025)
di: Epalle, Alexandre, et al.
Pubblicazione: (2025)
Robust and accurate simulations of flows over orography using non-conforming meshes
di: Orlando, Giuseppe, et al.
Pubblicazione: (2024)
di: Orlando, Giuseppe, et al.
Pubblicazione: (2024)
Serinv: A Scalable Library for the Selected Inversion of Block-Tridiagonal with Arrowhead Matrices
di: Maillou, Vincent, et al.
Pubblicazione: (2025)
di: Maillou, Vincent, et al.
Pubblicazione: (2025)
Code Generation for Near-Roofline Finite Element Actions on GPUs from Symbolic Variational Forms
di: Kulkarni, Kaushik, et al.
Pubblicazione: (2025)
di: Kulkarni, Kaushik, et al.
Pubblicazione: (2025)
PICO: Performance Insights for Collective Operations
di: Pasqualoni, Saverio, et al.
Pubblicazione: (2025)
di: Pasqualoni, Saverio, et al.
Pubblicazione: (2025)
A dynamic parallel method for performance optimization on hybrid CPUs
di: Yu, Luo, et al.
Pubblicazione: (2024)
di: Yu, Luo, et al.
Pubblicazione: (2024)
Accelerating Gaussian beam tracing method with dynamic parallelism on graphics processing units
di: Sheng, Zhang, et al.
Pubblicazione: (2025)
di: Sheng, Zhang, et al.
Pubblicazione: (2025)
ADELIA: Automatic Differentiation for Efficient Laplace Inference Approximations
di: Boudaoud, Afif, et al.
Pubblicazione: (2026)
di: Boudaoud, Afif, et al.
Pubblicazione: (2026)
Automated MPI-X code generation for scalable finite-difference solvers
di: Bisbas, George, et al.
Pubblicazione: (2023)
di: Bisbas, George, et al.
Pubblicazione: (2023)
Towards a Scalable and Efficient PGAS-based Distributed OpenMP
di: Shan, Baodi, et al.
Pubblicazione: (2024)
di: Shan, Baodi, et al.
Pubblicazione: (2024)
Efficient allocation of image recognition and LLM tasks on multi-GPU system
di: Lawenda, Marcin, et al.
Pubblicazione: (2025)
di: Lawenda, Marcin, et al.
Pubblicazione: (2025)
Efficient GPU-Centered Singular Value Decomposition Using the Divide-and-Conquer Method
di: Liu, Shifang, et al.
Pubblicazione: (2025)
di: Liu, Shifang, et al.
Pubblicazione: (2025)
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
di: Wang, Yuxin, et al.
Pubblicazione: (2023)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
di: Ather, Hammad, et al.
Pubblicazione: (2024)
di: Ather, Hammad, et al.
Pubblicazione: (2024)
Efficient Serverless Cold Start: Reducing Library Loading Overhead by Profile-guided Optimization
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2025)
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2025)
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
di: Lin, Mao, et al.
Pubblicazione: (2026)
di: Lin, Mao, et al.
Pubblicazione: (2026)
Efficient Fault Localization in a Cloud Stack Using End-to-End Application Service Topology
di: Mathews, Dhanya R, et al.
Pubblicazione: (2025)
di: Mathews, Dhanya R, et al.
Pubblicazione: (2025)
THEAS: Efficient Power Management in Multi-Core CPUs via Cache-Aware Resource Scheduling
di: Muhammad, Said, et al.
Pubblicazione: (2025)
di: Muhammad, Said, et al.
Pubblicazione: (2025)
Towards a Peer-to-Peer Data Distribution Layer for Efficient and Collaborative Resource Optimization of Distributed Dataflow Applications
di: Scheinert, Dominik, et al.
Pubblicazione: (2023)
di: Scheinert, Dominik, et al.
Pubblicazione: (2023)
Scaling the memory wall using mixed-precision -- HPG-MxP on an exascale machine
di: Kashi, Aditya, et al.
Pubblicazione: (2025)
di: Kashi, Aditya, et al.
Pubblicazione: (2025)
Towards a GPU-Parallelization of the neXtSIM-DG Dynamical Core
di: Jendersie, Robert, et al.
Pubblicazione: (2024)
di: Jendersie, Robert, et al.
Pubblicazione: (2024)
Two-Stage Block Orthogonalization to Improve Performance of $s$-step GMRES
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2024)
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2024)
SUNDIALS Time Integrators for Exascale Applications with Many Independent ODE Systems
di: Balos, Cody J., et al.
Pubblicazione: (2024)
di: Balos, Cody J., et al.
Pubblicazione: (2024)
Neural Acceleration of Incomplete Cholesky Preconditioners
di: Booth, Joshua Dennis, et al.
Pubblicazione: (2024)
di: Booth, Joshua Dennis, et al.
Pubblicazione: (2024)
A Parallel in Time Algorithm Based on ParaExp for Optimal Control Problems
di: Kwok, Felix, et al.
Pubblicazione: (2024)
di: Kwok, Felix, et al.
Pubblicazione: (2024)
Cucheb: A GPU implementation of the filtered Lanczos procedure
di: Aurentz, Jared L., et al.
Pubblicazione: (2024)
di: Aurentz, Jared L., et al.
Pubblicazione: (2024)
GPU Accelerated Implicit Kinetic Meshfree Method based on Modified LU-SGS
di: Verma, Mayuri, et al.
Pubblicazione: (2024)
di: Verma, Mayuri, et al.
Pubblicazione: (2024)
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
di: Rudi, Johann, et al.
Pubblicazione: (2024)
di: Rudi, Johann, et al.
Pubblicazione: (2024)
Error Analysis of Matrix Multiplication Emulation Using Ozaki-II Scheme
di: Uchino, Yuki, et al.
Pubblicazione: (2026)
di: Uchino, Yuki, et al.
Pubblicazione: (2026)
Teaching An Old Dog New Tricks: Porting Legacy Code to Heterogeneous Compute Architectures With Automated Code Translation
di: Nytko, Nicolas, et al.
Pubblicazione: (2025)
di: Nytko, Nicolas, et al.
Pubblicazione: (2025)
A simple GPU implementation of spectral-element methods for solving 3D Poisson type equations on rectangular domains and its applications
di: Liu, Xinyu, et al.
Pubblicazione: (2023)
di: Liu, Xinyu, et al.
Pubblicazione: (2023)
On some orthogonalization schemes in Tensor Train format
di: Coulaud, Olivier, et al.
Pubblicazione: (2022)
di: Coulaud, Olivier, et al.
Pubblicazione: (2022)
Asymptotic Analysis of a Leader Election Algorithm
di: Lavault, Christian, et al.
Pubblicazione: (2006)
di: Lavault, Christian, et al.
Pubblicazione: (2006)
Random-sketching Techniques to Enhance the Numerical Stability of Block Orthogonalization Algorithms for s-step GMRES
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2025)
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2025)
RAPTOR: Practical Numerical Profiling of Scientific Applications
di: Hoerold, Faveo, et al.
Pubblicazione: (2025)
di: Hoerold, Faveo, et al.
Pubblicazione: (2025)
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension
di: Coleman, Evan
Pubblicazione: (2026)
di: Coleman, Evan
Pubblicazione: (2026)
Documenti analoghi
-
Improving the scalability of a high-order atmospheric dynamics solver based on the deal.II library
di: Orlando, Giuseppe, et al.
Pubblicazione: (2025) -
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
di: Higgins, Andrew J., et al.
Pubblicazione: (2025) -
Efficient Hardware Accelerator Based on Medium Granularity Dataflow for SpTRSV
di: Chen, Qian, et al.
Pubblicazione: (2024) -
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
di: Baioni, Paolo Joseph, et al.
Pubblicazione: (2024) -
Parallel simulation and adaptive mesh refinement for 3D elastostatic contact mechanics problems between deformable bodies
di: Epalle, Alexandre, et al.
Pubblicazione: (2025)