On Advanced Monte Carlo Methods for Linear Algebra on Advanced Accelerator Architectures
Fuente:
arXiv
Saved in:
| Main Authors: | Lebedev, Anton, Alexandrov, Vassil |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HYLU: Hybrid Parallel Sparse LU Factorization
by: Chen, Xiaoming
Published: (2025)
by: Chen, Xiaoming
Published: (2025)
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
by: Homola, Jakub, et al.
Published: (2025)
by: Homola, Jakub, et al.
Published: (2025)
A Bayesian Optimization through Sequential Monte Carlo and Statistical Physics-Inspired Techniques
by: Lebedev, Anton, et al.
Published: (2024)
by: Lebedev, Anton, et al.
Published: (2024)
Sensor Placement for Tsunami Early Warning via Large-Scale Bayesian Optimal Experimental Design
by: Venkat, Sreeram, et al.
Published: (2026)
by: Venkat, Sreeram, et al.
Published: (2026)
Using matrices in post-processing phase of CFD simulations
by: Argentini, Gianluca
Published: (2004)
by: Argentini, Gianluca
Published: (2004)
Chebyshev Accelerated Subspace Eigensolver for Pseudo-hermitian Hamiltonians
by: Di Napoli, Edoardo, et al.
Published: (2026)
by: Di Napoli, Edoardo, et al.
Published: (2026)
A Comparative Analysis of Distributed Linear Solvers under Data Heterogeneity
by: Velasevic, Boris, et al.
Published: (2023)
by: Velasevic, Boris, et al.
Published: (2023)
Precision-Aware Iterative Algorithms Based on Group-Shared Exponents of Floating-Point Numbers
by: Gao, Jianhua, et al.
Published: (2024)
by: Gao, Jianhua, et al.
Published: (2024)
Speed, power and cost implications for GPU acceleration of Computational Fluid Dynamics on HPC systems
by: Cooper-Baldock, Zachary, et al.
Published: (2024)
by: Cooper-Baldock, Zachary, et al.
Published: (2024)
Distributed and heterogeneous tensor-vector contraction algorithms for high performance computing
by: Martinez-Ferrer, Pedro J., et al.
Published: (2025)
by: Martinez-Ferrer, Pedro J., et al.
Published: (2025)
NM-SpMM: Accelerating Matrix Multiplication Using N:M Sparsity with GPGPU
by: Ma, Cong, et al.
Published: (2025)
by: Ma, Cong, et al.
Published: (2025)
A Task Parallel Orthonormalization Multigrid Method For Multiphase Elliptic Problems
by: Toprak, Teoman, et al.
Published: (2025)
by: Toprak, Teoman, et al.
Published: (2025)
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
by: Kanakagiri, Raghavendra, et al.
Published: (2023)
by: Kanakagiri, Raghavendra, et al.
Published: (2023)
Solving Large Rank-Deficient Linear Least-Squares Problems on Shared-Memory CPU Architectures and GPU Architectures
by: Chillarón, Mónica, et al.
Published: (2024)
by: Chillarón, Mónica, et al.
Published: (2024)
PackSELL: A Sparse Matrix Format for Precision-Agnostic High-Performance SpMV
by: Suzuki, Kengo, et al.
Published: (2026)
by: Suzuki, Kengo, et al.
Published: (2026)
Workflow for High-Fidelity Dynamic Analysis of Structures with Pile Foundation
by: Pakzad, Amin, et al.
Published: (2025)
by: Pakzad, Amin, et al.
Published: (2025)
Canonicalization of Batched Einstein Summations for Tuning Retrieval
by: Kulkarni, Kaushik, et al.
Published: (2026)
by: Kulkarni, Kaushik, et al.
Published: (2026)
Challenging Portability Paradigms: FPGA Acceleration Using SYCL and OpenCL
by: de Castro, Manuel, et al.
Published: (2024)
by: de Castro, Manuel, et al.
Published: (2024)
Adaptive time step selection for Spectral Deferred Correction
by: Saupe, Thomas, et al.
Published: (2024)
by: Saupe, Thomas, et al.
Published: (2024)
Resilience Against Soft Faults through Adaptivity in Spectral Deferred Correction
by: Saupe, Thomas, et al.
Published: (2024)
by: Saupe, Thomas, et al.
Published: (2024)
Shortest paths search method based on the projective description of unweighted mixed graphs
by: Melent'ev, V. A.
Published: (2023)
by: Melent'ev, V. A.
Published: (2023)
Enabling Practical Transparent Checkpointing for MPI: A Topological Sort Approach
by: Xu, Yao, et al.
Published: (2024)
by: Xu, Yao, et al.
Published: (2024)
Optimizing Fine-Grained Parallelism Through Dynamic Load Balancing on Multi-Socket Many-Core Systems
by: Wang, Wenyi, et al.
Published: (2025)
by: Wang, Wenyi, et al.
Published: (2025)
Joint Training on AMD and NVIDIA GPUs
by: Hu, Jon, et al.
Published: (2026)
by: Hu, Jon, et al.
Published: (2026)
AutoTSMM: An Auto-tuning Framework for Building High-Performance Tall-and-Skinny Matrix-Matrix Multiplication on CPUs
by: Li, Chendi, et al.
Published: (2022)
by: Li, Chendi, et al.
Published: (2022)
Communication-Efficient, 2D Parallel Stochastic Gradient Descent for Distributed-Memory Optimization
by: Devarakonda, Aditya, et al.
Published: (2025)
by: Devarakonda, Aditya, et al.
Published: (2025)
TriADA: Massively Parallel Trilinear Matrix-by-Tensor Multiply-Add Algorithm and Device Architecture for the Acceleration of 3D Discrete Transformations
by: Sedukhin, Stanislav, et al.
Published: (2025)
by: Sedukhin, Stanislav, et al.
Published: (2025)
Comparison of substructured non-overlapping domain decomposition and overlapping additive Schwarz methods for large-scale Helmholtz problems with multiple sources
by: Martin, Boris, et al.
Published: (2025)
by: Martin, Boris, et al.
Published: (2025)
Scalable Dual Coordinate Descent for Kernel Methods
by: Shao, Zishan, et al.
Published: (2024)
by: Shao, Zishan, et al.
Published: (2024)
A Systematic Literature Survey of Sparse Matrix-Vector Multiplication
by: Gao, Jianhua, et al.
Published: (2024)
by: Gao, Jianhua, et al.
Published: (2024)
Optimization of Approximate Maps for Linear Systems Arising in Discretized PDEs
by: Islam, Rishad, et al.
Published: (2024)
by: Islam, Rishad, et al.
Published: (2024)
Communication-Efficient and Memory-Aware Parallel Bootstrapping using MPI
by: Zhang, Di
Published: (2025)
by: Zhang, Di
Published: (2025)
GPU Accelerated Implicit Kinetic Meshfree Method based on Modified LU-SGS
by: Verma, Mayuri, et al.
Published: (2024)
by: Verma, Mayuri, et al.
Published: (2024)
Serinv: A Scalable Library for the Selected Inversion of Block-Tridiagonal with Arrowhead Matrices
by: Maillou, Vincent, et al.
Published: (2025)
by: Maillou, Vincent, et al.
Published: (2025)
A Nested Krylov Method Using Half-Precision Arithmetic
by: Suzuki, Kengo, et al.
Published: (2025)
by: Suzuki, Kengo, et al.
Published: (2025)
Efficient Parallel Scheduling for Sparse Triangular Solvers
by: Böhnlein, Toni, et al.
Published: (2025)
by: Böhnlein, Toni, et al.
Published: (2025)
Two Iterative Algorithms for Solving Systems of Simultaneous Linear Algebraic Equations with Real Matrices of Coefficients
by: Kondratiev, A. S., et al.
Published: (2005)
by: Kondratiev, A. S., et al.
Published: (2005)
Stream parallel skeleton optimization
by: Aldinucci, Marco, et al.
Published: (2024)
by: Aldinucci, Marco, et al.
Published: (2024)
StreamFlow: cross-breeding cloud with HPC
by: Colonnelli, Iacopo, et al.
Published: (2020)
by: Colonnelli, Iacopo, et al.
Published: (2020)
A Two-Level Direct Solver for the Hierarchical Poincaré-Steklov Method
by: Kump, Joseph, et al.
Published: (2025)
by: Kump, Joseph, et al.
Published: (2025)
Similar Items
-
HYLU: Hybrid Parallel Sparse LU Factorization
by: Chen, Xiaoming
Published: (2025) -
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
by: Homola, Jakub, et al.
Published: (2025) -
A Bayesian Optimization through Sequential Monte Carlo and Statistical Physics-Inspired Techniques
by: Lebedev, Anton, et al.
Published: (2024) -
Sensor Placement for Tsunami Early Warning via Large-Scale Bayesian Optimal Experimental Design
by: Venkat, Sreeram, et al.
Published: (2026) -
Using matrices in post-processing phase of CFD simulations
by: Argentini, Gianluca
Published: (2004)