Salvato in:
| Autori principali: | Li, Huamin, Kluger, Yuval, Tygert, Mark |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2016
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/1612.08709 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Random-sketching Techniques to Enhance the Numerical Stability of Block Orthogonalization Algorithms for s-step GMRES
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2025)
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2025)
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension
di: Coleman, Evan
Pubblicazione: (2026)
di: Coleman, Evan
Pubblicazione: (2026)
Exponential convergence of a distributed divide-and-conquer algorithm for constrained convex optimization on networks
di: Emirov, Nazar, et al.
Pubblicazione: (2025)
di: Emirov, Nazar, et al.
Pubblicazione: (2025)
Error Analysis of Matrix Multiplication Emulation Using Ozaki-II Scheme
di: Uchino, Yuki, et al.
Pubblicazione: (2026)
di: Uchino, Yuki, et al.
Pubblicazione: (2026)
Towards a GPU-Parallelization of the neXtSIM-DG Dynamical Core
di: Jendersie, Robert, et al.
Pubblicazione: (2024)
di: Jendersie, Robert, et al.
Pubblicazione: (2024)
Parallel simulation and adaptive mesh refinement for 3D elastostatic contact mechanics problems between deformable bodies
di: Epalle, Alexandre, et al.
Pubblicazione: (2025)
di: Epalle, Alexandre, et al.
Pubblicazione: (2025)
Two-Stage Block Orthogonalization to Improve Performance of $s$-step GMRES
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2024)
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2024)
Teaching An Old Dog New Tricks: Porting Legacy Code to Heterogeneous Compute Architectures With Automated Code Translation
di: Nytko, Nicolas, et al.
Pubblicazione: (2025)
di: Nytko, Nicolas, et al.
Pubblicazione: (2025)
SUNDIALS Time Integrators for Exascale Applications with Many Independent ODE Systems
di: Balos, Cody J., et al.
Pubblicazione: (2024)
di: Balos, Cody J., et al.
Pubblicazione: (2024)
Neural Acceleration of Incomplete Cholesky Preconditioners
di: Booth, Joshua Dennis, et al.
Pubblicazione: (2024)
di: Booth, Joshua Dennis, et al.
Pubblicazione: (2024)
A simple GPU implementation of spectral-element methods for solving 3D Poisson type equations on rectangular domains and its applications
di: Liu, Xinyu, et al.
Pubblicazione: (2023)
di: Liu, Xinyu, et al.
Pubblicazione: (2023)
On some orthogonalization schemes in Tensor Train format
di: Coulaud, Olivier, et al.
Pubblicazione: (2022)
di: Coulaud, Olivier, et al.
Pubblicazione: (2022)
Asymptotic Analysis of a Leader Election Algorithm
di: Lavault, Christian, et al.
Pubblicazione: (2006)
di: Lavault, Christian, et al.
Pubblicazione: (2006)
A Parallel in Time Algorithm Based on ParaExp for Optimal Control Problems
di: Kwok, Felix, et al.
Pubblicazione: (2024)
di: Kwok, Felix, et al.
Pubblicazione: (2024)
Cucheb: A GPU implementation of the filtered Lanczos procedure
di: Aurentz, Jared L., et al.
Pubblicazione: (2024)
di: Aurentz, Jared L., et al.
Pubblicazione: (2024)
RAPTOR: Practical Numerical Profiling of Scientific Applications
di: Hoerold, Faveo, et al.
Pubblicazione: (2025)
di: Hoerold, Faveo, et al.
Pubblicazione: (2025)
GPU Accelerated Implicit Kinetic Meshfree Method based on Modified LU-SGS
di: Verma, Mayuri, et al.
Pubblicazione: (2024)
di: Verma, Mayuri, et al.
Pubblicazione: (2024)
Algebraic Temporal Blocking for Sparse Iterative Solvers on Multi-Core CPUs
di: Alappat, Christie, et al.
Pubblicazione: (2023)
di: Alappat, Christie, et al.
Pubblicazione: (2023)
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
di: Rudi, Johann, et al.
Pubblicazione: (2024)
di: Rudi, Johann, et al.
Pubblicazione: (2024)
Modifying the Asynchronous Jacobi Method for Data Corruption Resilience
di: Vogl, Christopher J., et al.
Pubblicazione: (2022)
di: Vogl, Christopher J., et al.
Pubblicazione: (2022)
Distributed computing for physics-based data-driven reduced modeling at scale: Application to a rotating detonation rocket engine
di: Farcas, Ionut-Gabriel, et al.
Pubblicazione: (2024)
di: Farcas, Ionut-Gabriel, et al.
Pubblicazione: (2024)
Parallel GPU-Accelerated Randomized Construction of Approximate Cholesky Preconditioners
di: Liang, Tianyu, et al.
Pubblicazione: (2025)
di: Liang, Tianyu, et al.
Pubblicazione: (2025)
Efficient and scalable atmospheric dynamics simulations using non-conforming meshes
di: Orlando, Giuseppe, et al.
Pubblicazione: (2024)
di: Orlando, Giuseppe, et al.
Pubblicazione: (2024)
Real-time Bayesian inference at extreme scale: A digital twin for tsunami early warning applied to the Cascadia subduction zone
di: Henneking, Stefan, et al.
Pubblicazione: (2025)
di: Henneking, Stefan, et al.
Pubblicazione: (2025)
Improving the scalability of a high-order atmospheric dynamics solver based on the deal.II library
di: Orlando, Giuseppe, et al.
Pubblicazione: (2025)
di: Orlando, Giuseppe, et al.
Pubblicazione: (2025)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
di: Higgins, Andrew J., et al.
Pubblicazione: (2025)
di: Higgins, Andrew J., et al.
Pubblicazione: (2025)
Portable, Massively Parallel Implementation of a Material Point Method for Compressible Flows
di: Baioni, Paolo Joseph, et al.
Pubblicazione: (2024)
di: Baioni, Paolo Joseph, et al.
Pubblicazione: (2024)
A Domain Decomposition-based Solver for Acoustic Wave propagation in Two-Dimensional Random Media
di: Vasudevan, Sudhi Sharma Padillath
Pubblicazione: (2025)
di: Vasudevan, Sudhi Sharma Padillath
Pubblicazione: (2025)
GPU-Parallelizable Randomized Sketch-and-Precondition for Linear Regression using Sparse Sign Sketches
di: Chen, Tyler, et al.
Pubblicazione: (2025)
di: Chen, Tyler, et al.
Pubblicazione: (2025)
Parallel Sparse and Data-Sparse Factorization-based Linear Solvers
di: Li, Xiaoye Sherry, et al.
Pubblicazione: (2026)
di: Li, Xiaoye Sherry, et al.
Pubblicazione: (2026)
TTrace: Lightweight Error Checking and Diagnosis for Distributed Training
di: Jiang, Haitian, et al.
Pubblicazione: (2025)
di: Jiang, Haitian, et al.
Pubblicazione: (2025)
SUperman: Efficient Permanent Computation on GPUs
di: Elbek, Deniz, et al.
Pubblicazione: (2025)
di: Elbek, Deniz, et al.
Pubblicazione: (2025)
Decomposing Solution Sets of Polynomial Systems: A New Parallel Monodromy Breakup Algorithm
di: Leykin, Anton, et al.
Pubblicazione: (2005)
di: Leykin, Anton, et al.
Pubblicazione: (2005)
A Distributed Block Chebyshev-Davidson Algorithm for Parallel Spectral Clustering
di: Pang, Qiyuan, et al.
Pubblicazione: (2022)
di: Pang, Qiyuan, et al.
Pubblicazione: (2022)
Parallelization of Software Systems Test Case Selection Algorithm Based on Singular Value Decomposition
di: Moghaddam, Mahdi Movahedian
Pubblicazione: (2022)
di: Moghaddam, Mahdi Movahedian
Pubblicazione: (2022)
DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration
di: Lai, Shih-Yu, et al.
Pubblicazione: (2026)
di: Lai, Shih-Yu, et al.
Pubblicazione: (2026)
Distributed Matrix-Vector Multiplication: A Convolutional Coding Approach
di: Das, Anindya Bijoy, et al.
Pubblicazione: (2019)
di: Das, Anindya Bijoy, et al.
Pubblicazione: (2019)
Accelerating Diffusion Models with Parallel Sampling: Inference at Sub-Linear Time Complexity
di: Chen, Haoxuan, et al.
Pubblicazione: (2024)
di: Chen, Haoxuan, et al.
Pubblicazione: (2024)
Improved Analysis of the Accelerated Noisy Power Method with Applications to Decentralized PCA
di: Aguié, Pierre, et al.
Pubblicazione: (2026)
di: Aguié, Pierre, et al.
Pubblicazione: (2026)
Fully-Automated Code Generation for Efficient Computation of Sparse Matrix Permanents on GPUs
di: Elbek, Deniz, et al.
Pubblicazione: (2025)
di: Elbek, Deniz, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Random-sketching Techniques to Enhance the Numerical Stability of Block Orthogonalization Algorithms for s-step GMRES
di: Yamazaki, Ichitaro, et al.
Pubblicazione: (2025) -
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension
di: Coleman, Evan
Pubblicazione: (2026) -
Exponential convergence of a distributed divide-and-conquer algorithm for constrained convex optimization on networks
di: Emirov, Nazar, et al.
Pubblicazione: (2025) -
Error Analysis of Matrix Multiplication Emulation Using Ozaki-II Scheme
di: Uchino, Yuki, et al.
Pubblicazione: (2026) -
Towards a GPU-Parallelization of the neXtSIM-DG Dynamical Core
di: Jendersie, Robert, et al.
Pubblicazione: (2024)