Improving the scalability of a high-order atmospheric dynamics solver based on the deal.II library
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Orlando, Giuseppe, Benacchio, Tommaso, Bonaventura, Luca |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient and scalable atmospheric dynamics simulations using non-conforming meshes
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024)
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
von: Baioni, Paolo Joseph, et al.
Veröffentlicht: (2024)
von: Baioni, Paolo Joseph, et al.
Veröffentlicht: (2024)
Efficient Hardware Accelerator Based on Medium Granularity Dataflow for SpTRSV
von: Chen, Qian, et al.
Veröffentlicht: (2024)
von: Chen, Qian, et al.
Veröffentlicht: (2024)
Automated MPI-X code generation for scalable finite-difference solvers
von: Bisbas, George, et al.
Veröffentlicht: (2023)
von: Bisbas, George, et al.
Veröffentlicht: (2023)
Serinv: A Scalable Library for the Selected Inversion of Block-Tridiagonal with Arrowhead Matrices
von: Maillou, Vincent, et al.
Veröffentlicht: (2025)
von: Maillou, Vincent, et al.
Veröffentlicht: (2025)
Two-Stage Block Orthogonalization to Improve Performance of $s$-step GMRES
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2024)
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2024)
Error Analysis of Matrix Multiplication Emulation Using Ozaki-II Scheme
von: Uchino, Yuki, et al.
Veröffentlicht: (2026)
von: Uchino, Yuki, et al.
Veröffentlicht: (2026)
Code Generation for Near-Roofline Finite Element Actions on GPUs from Symbolic Variational Forms
von: Kulkarni, Kaushik, et al.
Veröffentlicht: (2025)
von: Kulkarni, Kaushik, et al.
Veröffentlicht: (2025)
PICO: Performance Insights for Collective Operations
von: Pasqualoni, Saverio, et al.
Veröffentlicht: (2025)
von: Pasqualoni, Saverio, et al.
Veröffentlicht: (2025)
GPU Accelerated Implicit Kinetic Meshfree Method based on Modified LU-SGS
von: Verma, Mayuri, et al.
Veröffentlicht: (2024)
von: Verma, Mayuri, et al.
Veröffentlicht: (2024)
A dynamic parallel method for performance optimization on hybrid CPUs
von: Yu, Luo, et al.
Veröffentlicht: (2024)
von: Yu, Luo, et al.
Veröffentlicht: (2024)
RAID Organizations for Improved Reliability and Performance: A Not Entirely Unbiased Tutorial
von: Thomasian, Alexander
Veröffentlicht: (2023)
von: Thomasian, Alexander
Veröffentlicht: (2023)
Accelerating Gaussian beam tracing method with dynamic parallelism on graphics processing units
von: Sheng, Zhang, et al.
Veröffentlicht: (2025)
von: Sheng, Zhang, et al.
Veröffentlicht: (2025)
Parallel I/O Characterization and Optimization on Large-Scale HPC Systems: A 360-Degree Survey
von: Ather, Hammad, et al.
Veröffentlicht: (2024)
von: Ather, Hammad, et al.
Veröffentlicht: (2024)
mLR: Scalable Laminography Reconstruction based on Memoization
von: Ma, Bin, et al.
Veröffentlicht: (2025)
von: Ma, Bin, et al.
Veröffentlicht: (2025)
Towards a Scalable and Efficient PGAS-based Distributed OpenMP
von: Shan, Baodi, et al.
Veröffentlicht: (2024)
von: Shan, Baodi, et al.
Veröffentlicht: (2024)
Scaling Large-scale GNN Training to Thousands of Processors on CPU-based Supercomputers
von: Zhuang, Chen, et al.
Veröffentlicht: (2024)
von: Zhuang, Chen, et al.
Veröffentlicht: (2024)
Dissecting the software-based measurement of CPU energy consumption: a comparative analysis
von: Raffin, Guillaume, et al.
Veröffentlicht: (2024)
von: Raffin, Guillaume, et al.
Veröffentlicht: (2024)
Unleashing the Power of Preemptive Priority-based Scheduling for Real-Time GPU Tasks
von: Wang, Yidi, et al.
Veröffentlicht: (2024)
von: Wang, Yidi, et al.
Veröffentlicht: (2024)
Optimized thread-block arrangement in a GPU implementation of a linear solver for atmospheric chemistry mechanisms
von: Ruiz, Christian Guzman, et al.
Veröffentlicht: (2024)
von: Ruiz, Christian Guzman, et al.
Veröffentlicht: (2024)
Towards a GPU-Parallelization of the neXtSIM-DG Dynamical Core
von: Jendersie, Robert, et al.
Veröffentlicht: (2024)
von: Jendersie, Robert, et al.
Veröffentlicht: (2024)
Parallel simulation and adaptive mesh refinement for 3D elastostatic contact mechanics problems between deformable bodies
von: Epalle, Alexandre, et al.
Veröffentlicht: (2025)
von: Epalle, Alexandre, et al.
Veröffentlicht: (2025)
Teaching An Old Dog New Tricks: Porting Legacy Code to Heterogeneous Compute Architectures With Automated Code Translation
von: Nytko, Nicolas, et al.
Veröffentlicht: (2025)
von: Nytko, Nicolas, et al.
Veröffentlicht: (2025)
SUNDIALS Time Integrators for Exascale Applications with Many Independent ODE Systems
von: Balos, Cody J., et al.
Veröffentlicht: (2024)
von: Balos, Cody J., et al.
Veröffentlicht: (2024)
Neural Acceleration of Incomplete Cholesky Preconditioners
von: Booth, Joshua Dennis, et al.
Veröffentlicht: (2024)
von: Booth, Joshua Dennis, et al.
Veröffentlicht: (2024)
A simple GPU implementation of spectral-element methods for solving 3D Poisson type equations on rectangular domains and its applications
von: Liu, Xinyu, et al.
Veröffentlicht: (2023)
von: Liu, Xinyu, et al.
Veröffentlicht: (2023)
On some orthogonalization schemes in Tensor Train format
von: Coulaud, Olivier, et al.
Veröffentlicht: (2022)
von: Coulaud, Olivier, et al.
Veröffentlicht: (2022)
Asymptotic Analysis of a Leader Election Algorithm
von: Lavault, Christian, et al.
Veröffentlicht: (2006)
von: Lavault, Christian, et al.
Veröffentlicht: (2006)
A Parallel in Time Algorithm Based on ParaExp for Optimal Control Problems
von: Kwok, Felix, et al.
Veröffentlicht: (2024)
von: Kwok, Felix, et al.
Veröffentlicht: (2024)
Cucheb: A GPU implementation of the filtered Lanczos procedure
von: Aurentz, Jared L., et al.
Veröffentlicht: (2024)
von: Aurentz, Jared L., et al.
Veröffentlicht: (2024)
Random-sketching Techniques to Enhance the Numerical Stability of Block Orthogonalization Algorithms for s-step GMRES
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2025)
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2025)
RAPTOR: Practical Numerical Profiling of Scientific Applications
von: Hoerold, Faveo, et al.
Veröffentlicht: (2025)
von: Hoerold, Faveo, et al.
Veröffentlicht: (2025)
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension
von: Coleman, Evan
Veröffentlicht: (2026)
von: Coleman, Evan
Veröffentlicht: (2026)
Algebraic Temporal Blocking for Sparse Iterative Solvers on Multi-Core CPUs
von: Alappat, Christie, et al.
Veröffentlicht: (2023)
von: Alappat, Christie, et al.
Veröffentlicht: (2023)
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
von: Rudi, Johann, et al.
Veröffentlicht: (2024)
von: Rudi, Johann, et al.
Veröffentlicht: (2024)
Modifying the Asynchronous Jacobi Method for Data Corruption Resilience
von: Vogl, Christopher J., et al.
Veröffentlicht: (2022)
von: Vogl, Christopher J., et al.
Veröffentlicht: (2022)
Constructive community race: full-density spiking neural network model drives neuromorphic computing
von: Senk, Johanna, et al.
Veröffentlicht: (2025)
von: Senk, Johanna, et al.
Veröffentlicht: (2025)
Randomized algorithms for distributed computation of principal component analysis and singular value decomposition
von: Li, Huamin, et al.
Veröffentlicht: (2016)
von: Li, Huamin, et al.
Veröffentlicht: (2016)
Improved Analysis of the Accelerated Noisy Power Method with Applications to Decentralized PCA
von: Aguié, Pierre, et al.
Veröffentlicht: (2026)
von: Aguié, Pierre, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Efficient and scalable atmospheric dynamics simulations using non-conforming meshes
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024) -
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025) -
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
von: Baioni, Paolo Joseph, et al.
Veröffentlicht: (2024) -
Efficient Hardware Accelerator Based on Medium Granularity Dataflow for SpTRSV
von: Chen, Qian, et al.
Veröffentlicht: (2024) -
Automated MPI-X code generation for scalable finite-difference solvers
von: Bisbas, George, et al.
Veröffentlicht: (2023)