Performance evaluation of accelerated complex multiple-precision LU decomposition
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Kouya, Tomonori |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Performance evaluation of accelerated real and complex multiple-precision sparse matrix-vector multiplication
von: Kouya, Tomonori
Veröffentlicht: (2024)
von: Kouya, Tomonori
Veröffentlicht: (2024)
Acceleration of multi-component multiple-precision arithmetic with branch-free algorithms and SIMD vectorization
von: Kouya, Tomonori
Veröffentlicht: (2026)
von: Kouya, Tomonori
Veröffentlicht: (2026)
Assessing the Performance of Mixed-Precision ILU(0)-Preconditioned Multiple-Precision Real and Complex Krylov Subspace Methods
von: Kouya, Tomonori
Veröffentlicht: (2025)
von: Kouya, Tomonori
Veröffentlicht: (2025)
LRAMM -- Low precision approximates GEMM via RSVD
von: Gu, Hongyaoxing
Veröffentlicht: (2024)
von: Gu, Hongyaoxing
Veröffentlicht: (2024)
A Technical Survey of Sparse Linear Solvers in Electronic Design Automation
von: Rai, Nityanand
Veröffentlicht: (2025)
von: Rai, Nityanand
Veröffentlicht: (2025)
sTiles: An Accelerated Computational Framework for Sparse Factorizations of Structured Matrices
von: Fattah, Esmail Abdul, et al.
Veröffentlicht: (2025)
von: Fattah, Esmail Abdul, et al.
Veröffentlicht: (2025)
On General Linearly Implicit Quantized State System Methods
von: Bergonzi, Mariana, et al.
Veröffentlicht: (2025)
von: Bergonzi, Mariana, et al.
Veröffentlicht: (2025)
Scalable Binary CUR Low-Rank Approximation Algorithm
von: Su, Bowen
Veröffentlicht: (2025)
von: Su, Bowen
Veröffentlicht: (2025)
HPL-MxP Benchmark: Mixed-Precision Algorithms, Iterative Refinement, and Scalable Data Generation
von: Dongarra, Jack, et al.
Veröffentlicht: (2025)
von: Dongarra, Jack, et al.
Veröffentlicht: (2025)
Online Pseudo-average Shifting Attention(PASA) for Robust Low-precision LLM Inference: Algorithms and Numerical Analysis
von: Cheng, Long, et al.
Veröffentlicht: (2025)
von: Cheng, Long, et al.
Veröffentlicht: (2025)
Performance evaluation of mixed-precision Runge-Kutta methods for the solution of partial differential equations
von: Dravins, Ivo, et al.
Veröffentlicht: (2024)
von: Dravins, Ivo, et al.
Veröffentlicht: (2024)
Accelerating AI Performance using Anderson Extrapolation on GPUs
von: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Veröffentlicht: (2024)
von: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Veröffentlicht: (2024)
Identifying Best Practice Melting Patterns in Induction Furnaces: A Data-Driven Approach Using Time Series KMeans Clustering and Multi-Criteria Decision Making
von: Howard, Daniel Anthony, et al.
Veröffentlicht: (2024)
von: Howard, Daniel Anthony, et al.
Veröffentlicht: (2024)
Cache Blocking for Flux Reconstruction: Extension to Navier-Stokes Equations and Anti-aliasing
von: Akkurt, Semih, et al.
Veröffentlicht: (2023)
von: Akkurt, Semih, et al.
Veröffentlicht: (2023)
A High Performance GPU CountSketch Implementation and Its Application to Multisketching and Least Squares Problems
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
von: Higgins, Andrew J., et al.
Veröffentlicht: (2025)
A method for accelerating low precision operations by sparse matrix multiplication
von: Gu, Hongyaoxing
Veröffentlicht: (2024)
von: Gu, Hongyaoxing
Veröffentlicht: (2024)
Diagonally-Addressed Matrix Nicknack: How to improve SpMV performance
von: Saak, Jens, et al.
Veröffentlicht: (2023)
von: Saak, Jens, et al.
Veröffentlicht: (2023)
Efficient and scalable atmospheric dynamics simulations using non-conforming meshes
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024)
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2024)
A Robust Sustainability Assessment Methodology for Aircraft Parts: Application to a Fuselage Panel
von: Anagnostopoulou, Aikaterini A., et al.
Veröffentlicht: (2024)
von: Anagnostopoulou, Aikaterini A., et al.
Veröffentlicht: (2024)
StructMG: A Fast and Scalable Structured Algebraic Multigrid
von: Zong, Yi, et al.
Veröffentlicht: (2025)
von: Zong, Yi, et al.
Veröffentlicht: (2025)
Improving the scalability of a high-order atmospheric dynamics solver based on the deal.II library
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2025)
von: Orlando, Giuseppe, et al.
Veröffentlicht: (2025)
Fast wavefield evaluation method based on modified proxy-surface-accelerated interpolative decomposition for two-dimensional scattering problems
von: Matsumoto, Yasuhiro
Veröffentlicht: (2024)
von: Matsumoto, Yasuhiro
Veröffentlicht: (2024)
Scaling the memory wall using mixed-precision -- HPG-MxP on an exascale machine
von: Kashi, Aditya, et al.
Veröffentlicht: (2025)
von: Kashi, Aditya, et al.
Veröffentlicht: (2025)
Efficient Hardware Accelerator Based on Medium Granularity Dataflow for SpTRSV
von: Chen, Qian, et al.
Veröffentlicht: (2024)
von: Chen, Qian, et al.
Veröffentlicht: (2024)
Mapping Sparse Triangular Solves to GPUs via Fine-grained Domain Decomposition
von: Gondhalekar, Atharva, et al.
Veröffentlicht: (2025)
von: Gondhalekar, Atharva, et al.
Veröffentlicht: (2025)
Efficient approximations of matrix multiplication using truncated decompositions
von: Kar, Suvendu, et al.
Veröffentlicht: (2025)
von: Kar, Suvendu, et al.
Veröffentlicht: (2025)
ShyLU node: On-node Scalable Solvers and Preconditioners Recent Progresses and Current Performance
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2025)
von: Yamazaki, Ichitaro, et al.
Veröffentlicht: (2025)
A Novel Paradigm Shift for Next-Generation: Symbiotic Backscatter Rate-Splitting Multiple Access Systems
von: Vu, Thai-Hoc, et al.
Veröffentlicht: (2024)
von: Vu, Thai-Hoc, et al.
Veröffentlicht: (2024)
Acceleration of Tensor-Product Operations with Tensor Cores
von: Cui, Cu
Veröffentlicht: (2024)
von: Cui, Cu
Veröffentlicht: (2024)
A Multi-server Scheduling Framework for Resource Allocation in Wireless Multi-carrier Networks
von: Zhang, Ying Jun
Veröffentlicht: (2006)
von: Zhang, Ying Jun
Veröffentlicht: (2006)
An accelerated Levin-Clenshaw-Curtis method for the evaluation of highly oscillatory integrals
von: Iserles, Arieh, et al.
Veröffentlicht: (2024)
von: Iserles, Arieh, et al.
Veröffentlicht: (2024)
Low-Rank Approximation by Randomly Pivoted LU
von: Gilles, Marc Aurèle, et al.
Veröffentlicht: (2026)
von: Gilles, Marc Aurèle, et al.
Veröffentlicht: (2026)
Serinv: A Scalable Library for the Selected Inversion of Block-Tridiagonal with Arrowhead Matrices
von: Maillou, Vincent, et al.
Veröffentlicht: (2025)
von: Maillou, Vincent, et al.
Veröffentlicht: (2025)
Algorithms and optimizations for global non-linear hybrid fluid-kinetic finite element stellarator simulations
von: Greco, Luca Venerando
Veröffentlicht: (2025)
von: Greco, Luca Venerando
Veröffentlicht: (2025)
Faster arbitrary-precision dot product and matrix multiplication
von: Johansson, Fredrik
Veröffentlicht: (2019)
von: Johansson, Fredrik
Veröffentlicht: (2019)
An accelerated direct solver for scalar wave scattering by multiple transmissive inclusions in two dimensions
von: Matsumoto, Yasuhiro
Veröffentlicht: (2026)
von: Matsumoto, Yasuhiro
Veröffentlicht: (2026)
A neural network kernel decomposition for learning multiple steady states in parameterized dynamical systems
von: Zhang, Yimeng, et al.
Veröffentlicht: (2023)
von: Zhang, Yimeng, et al.
Veröffentlicht: (2023)
Generalized cyclic symmetric decompositions for the matrix multiplication tensor
von: Vermeylen, Charlotte, et al.
Veröffentlicht: (2024)
von: Vermeylen, Charlotte, et al.
Veröffentlicht: (2024)
A GPU accelerated mixed-precision Finite Difference informed Random Walker (FDiRW) solver for strongly inhomogeneous diffusion problems
von: Mao, Zirui, et al.
Veröffentlicht: (2024)
von: Mao, Zirui, et al.
Veröffentlicht: (2024)
Speed, power and cost implications for GPU acceleration of Computational Fluid Dynamics on HPC systems
von: Cooper-Baldock, Zachary, et al.
Veröffentlicht: (2024)
von: Cooper-Baldock, Zachary, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Performance evaluation of accelerated real and complex multiple-precision sparse matrix-vector multiplication
von: Kouya, Tomonori
Veröffentlicht: (2024) -
Acceleration of multi-component multiple-precision arithmetic with branch-free algorithms and SIMD vectorization
von: Kouya, Tomonori
Veröffentlicht: (2026) -
Assessing the Performance of Mixed-Precision ILU(0)-Preconditioned Multiple-Precision Real and Complex Krylov Subspace Methods
von: Kouya, Tomonori
Veröffentlicht: (2025) -
LRAMM -- Low precision approximates GEMM via RSVD
von: Gu, Hongyaoxing
Veröffentlicht: (2024) -
A Technical Survey of Sparse Linear Solvers in Electronic Design Automation
von: Rai, Nityanand
Veröffentlicht: (2025)