Gespeichert in:
| Hauptverfasser: | Ramirez-Hidalgo, Gustavo, He, Lianhua, Zhang, Ke-Long |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2407.08092 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multiple right hand side multigrid for domain wall fermions with a multigrid preconditioned block conjugate gradient algorithm
von: Boyle, Peter A
Veröffentlicht: (2024)
von: Boyle, Peter A
Veröffentlicht: (2024)
Performance-Portable Optimization and Analysis of Multiple Right-Hand Sides in a Lattice QCD Solver
von: Long, Shiting, et al.
Veröffentlicht: (2026)
von: Long, Shiting, et al.
Veröffentlicht: (2026)
Accelerating Lattice QCD Simulations using GPUs
von: Matthaei, Tilmann
Veröffentlicht: (2024)
von: Matthaei, Tilmann
Veröffentlicht: (2024)
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
von: Baioni, Paolo Joseph, et al.
Veröffentlicht: (2024)
von: Baioni, Paolo Joseph, et al.
Veröffentlicht: (2024)
A simple GPU implementation of spectral-element methods for solving 3D Poisson type equations on rectangular domains and its applications
von: Liu, Xinyu, et al.
Veröffentlicht: (2023)
von: Liu, Xinyu, et al.
Veröffentlicht: (2023)
Energy Efficiency trends in HPC: what high-energy and astrophysicists need to know
von: Suarez, Estela, et al.
Veröffentlicht: (2025)
von: Suarez, Estela, et al.
Veröffentlicht: (2025)
Towards a GPU-Parallelization of the neXtSIM-DG Dynamical Core
von: Jendersie, Robert, et al.
Veröffentlicht: (2024)
von: Jendersie, Robert, et al.
Veröffentlicht: (2024)
Two-Stage Block Orthogonalization to Improve Performance of $s$-step GMRES
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2024)
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2024)
SUNDIALS Time Integrators for Exascale Applications with Many Independent ODE Systems
von: Balos, Cody J., et al.
Veröffentlicht: (2024)
von: Balos, Cody J., et al.
Veröffentlicht: (2024)
Neural Acceleration of Incomplete Cholesky Preconditioners
von: Booth, Joshua Dennis, et al.
Veröffentlicht: (2024)
von: Booth, Joshua Dennis, et al.
Veröffentlicht: (2024)
A Parallel in Time Algorithm Based on ParaExp for Optimal Control Problems
von: Kwok, Felix, et al.
Veröffentlicht: (2024)
von: Kwok, Felix, et al.
Veröffentlicht: (2024)
Cucheb: A GPU implementation of the filtered Lanczos procedure
von: Aurentz, Jared L., et al.
Veröffentlicht: (2024)
von: Aurentz, Jared L., et al.
Veröffentlicht: (2024)
GPU Accelerated Implicit Kinetic Meshfree Method based on Modified LU-SGS
von: Verma, Mayuri, et al.
Veröffentlicht: (2024)
von: Verma, Mayuri, et al.
Veröffentlicht: (2024)
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
von: Rudi, Johann, et al.
Veröffentlicht: (2024)
von: Rudi, Johann, et al.
Veröffentlicht: (2024)
Error Analysis of Matrix Multiplication Emulation Using Ozaki-II Scheme
von: Uchino, Yuki, et al.
Veröffentlicht: (2026)
von: Uchino, Yuki, et al.
Veröffentlicht: (2026)
Parallel simulation and adaptive mesh refinement for 3D elastostatic contact mechanics problems between deformable bodies
von: Epalle, Alexandre, et al.
Veröffentlicht: (2025)
von: Epalle, Alexandre, et al.
Veröffentlicht: (2025)
Teaching An Old Dog New Tricks: Porting Legacy Code to Heterogeneous Compute Architectures With Automated Code Translation
von: Nytko, Nicolas, et al.
Veröffentlicht: (2025)
von: Nytko, Nicolas, et al.
Veröffentlicht: (2025)
On some orthogonalization schemes in Tensor Train format
von: Coulaud, Olivier, et al.
Veröffentlicht: (2022)
von: Coulaud, Olivier, et al.
Veröffentlicht: (2022)
Asymptotic Analysis of a Leader Election Algorithm
von: Lavault, Christian, et al.
Veröffentlicht: (2006)
von: Lavault, Christian, et al.
Veröffentlicht: (2006)
Random-sketching Techniques to Enhance the Numerical Stability of Block Orthogonalization Algorithms for s-step GMRES
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2025)
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2025)
RAPTOR: Practical Numerical Profiling of Scientific Applications
von: Hoerold, Faveo, et al.
Veröffentlicht: (2025)
von: Hoerold, Faveo, et al.
Veröffentlicht: (2025)
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension
von: Coleman, Evan
Veröffentlicht: (2026)
von: Coleman, Evan
Veröffentlicht: (2026)
Algebraic Temporal Blocking for Sparse Iterative Solvers on Multi-Core CPUs
von: Alappat, Christie, et al.
Veröffentlicht: (2023)
von: Alappat, Christie, et al.
Veröffentlicht: (2023)
Modifying the Asynchronous Jacobi Method for Data Corruption Resilience
von: Vogl, Christopher J., et al.
Veröffentlicht: (2022)
von: Vogl, Christopher J., et al.
Veröffentlicht: (2022)
Randomized algorithms for distributed computation of principal component analysis and singular value decomposition
von: Li, Huamin, et al.
Veröffentlicht: (2016)
von: Li, Huamin, et al.
Veröffentlicht: (2016)
Efficient and scalable atmospheric dynamics simulations using non-conforming meshes
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024)
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024)
Real-time Bayesian inference at extreme scale: A digital twin for tsunami early warning applied to the Cascadia subduction zone
von: Henneking, Stefan, et al.
Veröffentlicht: (2025)
von: Henneking, Stefan, et al.
Veröffentlicht: (2025)
Improving the scalability of a high-order atmospheric dynamics solver based on the deal.II library
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2025)
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2025)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
TTrace: Lightweight Error Checking and Diagnosis for Distributed Training
von: Jiang, Haitian, et al.
Veröffentlicht: (2025)
von: Jiang, Haitian, et al.
Veröffentlicht: (2025)
TriMe++: Multi-threaded triangular meshing in two dimensions
von: Lu, Jiayin, et al.
Veröffentlicht: (2023)
von: Lu, Jiayin, et al.
Veröffentlicht: (2023)
Accelerating Diffusion Models with Parallel Sampling: Inference at Sub-Linear Time Complexity
von: Chen, Haoxuan, et al.
Veröffentlicht: (2024)
von: Chen, Haoxuan, et al.
Veröffentlicht: (2024)
Distributed computing for physics-based data-driven reduced modeling at scale: Application to a rotating detonation rocket engine
von: Farcas, Ionut-Gabriel, et al.
Veröffentlicht: (2024)
von: Farcas, Ionut-Gabriel, et al.
Veröffentlicht: (2024)
Impact of EIP-4844 on Ethereum: Consensus Security, Ethereum Usage, Rollup Transaction Dynamics, and Blob Gas Fee Markets
von: Park, Seongwan, et al.
Veröffentlicht: (2024)
von: Park, Seongwan, et al.
Veröffentlicht: (2024)
SUperman: Efficient Permanent Computation on GPUs
von: Elbek, Deniz, et al.
Veröffentlicht: (2025)
von: Elbek, Deniz, et al.
Veröffentlicht: (2025)
Decomposing Solution Sets of Polynomial Systems: A New Parallel Monodromy Breakup Algorithm
von: Leykin, Anton, et al.
Veröffentlicht: (2005)
von: Leykin, Anton, et al.
Veröffentlicht: (2005)
A Distributed Block Chebyshev-Davidson Algorithm for Parallel Spectral Clustering
von: Pang, Qiyuan, et al.
Veröffentlicht: (2022)
von: Pang, Qiyuan, et al.
Veröffentlicht: (2022)
Parallelization of Software Systems Test Case Selection Algorithm Based on Singular Value Decomposition
von: Moghaddam, Mahdi Movahedian
Veröffentlicht: (2022)
von: Moghaddam, Mahdi Movahedian
Veröffentlicht: (2022)
DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration
von: Lai, Shih-Yu, et al.
Veröffentlicht: (2026)
von: Lai, Shih-Yu, et al.
Veröffentlicht: (2026)
Distributed Matrix-Vector Multiplication: A Convolutional Coding Approach
von: Das, Anindya Bijoy, et al.
Veröffentlicht: (2019)
von: Das, Anindya Bijoy, et al.
Veröffentlicht: (2019)
Ähnliche Einträge
-
Multiple right hand side multigrid for domain wall fermions with a multigrid preconditioned block conjugate gradient algorithm
von: Boyle, Peter A
Veröffentlicht: (2024) -
Performance-Portable Optimization and Analysis of Multiple Right-Hand Sides in a Lattice QCD Solver
von: Long, Shiting, et al.
Veröffentlicht: (2026) -
Accelerating Lattice QCD Simulations using GPUs
von: Matthaei, Tilmann
Veröffentlicht: (2024) -
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
von: Baioni, Paolo Joseph, et al.
Veröffentlicht: (2024) -
A simple GPU implementation of spectral-element methods for solving 3D Poisson type equations on rectangular domains and its applications
von: Liu, Xinyu, et al.
Veröffentlicht: (2023)