Saved in:
| Main Authors: | Veneva, Milena, Imamura, Toshiyuki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.05938 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ML-Based Optimum Sub-system Size Heuristic for the GPU Implementation of the Tridiagonal Partition Method
by: Veneva, Milena
Published: (2025)
by: Veneva, Milena
Published: (2025)
Floating Point Compression of Hierarchical Matrix Formats and its Impact on Matrix-Vector Multiplication
by: Kriemann, Ronald
Published: (2024)
by: Kriemann, Ronald
Published: (2024)
Parallel Gauss-Jordan Elimination and System Reduction for Efficient Circuit Simulation
by: Noveski, Filip, et al.
Published: (2026)
by: Noveski, Filip, et al.
Published: (2026)
Mixed-Precision Performance Portability of FFT-Based GPU-Accelerated Algorithms for Block-Triangular Toeplitz Matrices
by: Venkat, Sreeram, et al.
Published: (2025)
by: Venkat, Sreeram, et al.
Published: (2025)
A Task Parallel Orthonormalization Multigrid Method For Multiphase Elliptic Problems
by: Toprak, Teoman, et al.
Published: (2025)
by: Toprak, Teoman, et al.
Published: (2025)
Racing to Idle: Energy Efficiency of Matrix Multiplication on Heterogeneous CPU and GPU Architectures
by: Ansari, Mufakir Qamar, et al.
Published: (2025)
by: Ansari, Mufakir Qamar, et al.
Published: (2025)
RandNet-Parareal: a time-parallel PDE solver using Random Neural Networks
by: Gattiglio, Guglielmo, et al.
Published: (2024)
by: Gattiglio, Guglielmo, et al.
Published: (2024)
Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU
by: Ansari, Mufakir Qamar, et al.
Published: (2025)
by: Ansari, Mufakir Qamar, et al.
Published: (2025)
Parallelization Strategies for the Randomized Kaczmarz Algorithm on Large-Scale Dense Systems
by: Ferreira, Inês, et al.
Published: (2024)
by: Ferreira, Inês, et al.
Published: (2024)
CLAIRE: Scalable GPU-Accelerated Algorithms for Diffeomorphic Image Registration in 3D
by: Mang, Andreas
Published: (2024)
by: Mang, Andreas
Published: (2024)
Scalable Mean-Variance Portfolio Optimization via Subspace Embeddings and GPU-Friendly Nesterov-Accelerated Projected Gradient
by: Niu, Yi-Shuai, et al.
Published: (2026)
by: Niu, Yi-Shuai, et al.
Published: (2026)
Symbolic Algorithm for Solving SLAEs with Multi-Diagonal Coefficient Matrices
by: Veneva, Milena
Published: (2024)
by: Veneva, Milena
Published: (2024)
A multigrid reduction framework for domains with symmetries
by: Alsalti-Baldellou, Àdel, et al.
Published: (2024)
by: Alsalti-Baldellou, Àdel, et al.
Published: (2024)
The Performance of Low-Synchronization Variants of Reorthogonalized Block Classical Gram--Schmidt
by: Carson, Erin, et al.
Published: (2025)
by: Carson, Erin, et al.
Published: (2025)
Code Generation for Near-Roofline Finite Element Actions on GPUs from Symbolic Variational Forms
by: Kulkarni, Kaushik, et al.
Published: (2025)
by: Kulkarni, Kaushik, et al.
Published: (2025)
Adaptive time step selection for Spectral Deferred Correction
by: Saupe, Thomas, et al.
Published: (2024)
by: Saupe, Thomas, et al.
Published: (2024)
Resilience Against Soft Faults through Adaptivity in Spectral Deferred Correction
by: Saupe, Thomas, et al.
Published: (2024)
by: Saupe, Thomas, et al.
Published: (2024)
nuGPR: GPU-Accelerated Gaussian Process Regression with Iterative Algorithms and Low-Rank Approximations
by: Zhao, Ziqi, et al.
Published: (2025)
by: Zhao, Ziqi, et al.
Published: (2025)
Nearest Neighbors GParareal: Improving Scalability of Gaussian Processes for Parallel-in-Time Solvers
by: Gattiglio, Guglielmo, et al.
Published: (2024)
by: Gattiglio, Guglielmo, et al.
Published: (2024)
Parallel performance of shared memory parallel spectral deferred corrections
by: Freese, Philip, et al.
Published: (2024)
by: Freese, Philip, et al.
Published: (2024)
Parametrization and convergence of a primal-dual block-coordinate approach to linearly-constrained nonsmooth optimization
by: Bilenne, Olivier
Published: (2024)
by: Bilenne, Olivier
Published: (2024)
Prob-GParareal: A Probabilistic Numerical Parallel-in-Time Solver for Differential Equations
by: Gattiglio, Guglielmo, et al.
Published: (2025)
by: Gattiglio, Guglielmo, et al.
Published: (2025)
High-performance matrix-free unfitted finite element operator evaluation
by: Bergbauer, Maximilian, et al.
Published: (2024)
by: Bergbauer, Maximilian, et al.
Published: (2024)
Matrix-Free Evaluation of High-Order Shifted Boundary Finite Element Operators
by: Wichrowski, Michał
Published: (2025)
by: Wichrowski, Michał
Published: (2025)
Scalable Dual Coordinate Descent for Kernel Methods
by: Shao, Zishan, et al.
Published: (2024)
by: Shao, Zishan, et al.
Published: (2024)
A Virtual Processor brings back the Free Lunch
by: Kutschbach, Haymo
Published: (2026)
by: Kutschbach, Haymo
Published: (2026)
M2L Translation Operators for Kernel Independent Fast Multipole Methods on Modern Architectures
by: Kailasa, Srinath, et al.
Published: (2024)
by: Kailasa, Srinath, et al.
Published: (2024)
A Proximal-Gradient Method for Constrained Optimization
by: Dai, Yutong, et al.
Published: (2024)
by: Dai, Yutong, et al.
Published: (2024)
Small errors in random zeroth-order optimization are imaginary
by: Jongeneel, Wouter, et al.
Published: (2021)
by: Jongeneel, Wouter, et al.
Published: (2021)
A Proximal-Gradient Method for Solving Regularized Optimization Problems with General Constraints
by: Curtis, Frank E., et al.
Published: (2025)
by: Curtis, Frank E., et al.
Published: (2025)
Low-Memory Numerical Certification
by: Breiding, Paul, et al.
Published: (2026)
by: Breiding, Paul, et al.
Published: (2026)
On the Relationships among GPU-Accelerated First-Order Methods for Solving Linear Programming
by: Chen, Kaihuang, et al.
Published: (2025)
by: Chen, Kaihuang, et al.
Published: (2025)
Sensor Placement for Tsunami Early Warning via Large-Scale Bayesian Optimal Experimental Design
by: Venkat, Sreeram, et al.
Published: (2026)
by: Venkat, Sreeram, et al.
Published: (2026)
Accelerated primal dual fixed point algorithm
by: Zhu, Ya-Nan
Published: (2025)
by: Zhu, Ya-Nan
Published: (2025)
Scaling the memory wall using mixed-precision -- HPG-MxP on an exascale machine
by: Kashi, Aditya, et al.
Published: (2025)
by: Kashi, Aditya, et al.
Published: (2025)
Bare-Metal Tensor Virtualization: Overcoming the Memory Wall in Edge-AI Inference on ARM64
by: Kilictas, Bugra, et al.
Published: (2026)
by: Kilictas, Bugra, et al.
Published: (2026)
Memory- and compute-optimized geometric multigrid GMGPolar for curvilinear coordinate representations -- Applications to fusion plasma
by: Litz, Julian, et al.
Published: (2025)
by: Litz, Julian, et al.
Published: (2025)
A Parareal Algorithm with Low-Rank Coarse Solvers
by: Gander, Martin J., et al.
Published: (2025)
by: Gander, Martin J., et al.
Published: (2025)
Iterative Methods in GPU-Resident Linear Solvers for Nonlinear Constrained Optimization
by: Świrydowicz, Kasia, et al.
Published: (2024)
by: Świrydowicz, Kasia, et al.
Published: (2024)
A Continuous Energy Ising Machine Leveraging Difference-of-Convex Programming
by: Banerjee, Debraj, et al.
Published: (2025)
by: Banerjee, Debraj, et al.
Published: (2025)
Similar Items
-
ML-Based Optimum Sub-system Size Heuristic for the GPU Implementation of the Tridiagonal Partition Method
by: Veneva, Milena
Published: (2025) -
Floating Point Compression of Hierarchical Matrix Formats and its Impact on Matrix-Vector Multiplication
by: Kriemann, Ronald
Published: (2024) -
Parallel Gauss-Jordan Elimination and System Reduction for Efficient Circuit Simulation
by: Noveski, Filip, et al.
Published: (2026) -
Mixed-Precision Performance Portability of FFT-Based GPU-Accelerated Algorithms for Block-Triangular Toeplitz Matrices
by: Venkat, Sreeram, et al.
Published: (2025) -
A Task Parallel Orthonormalization Multigrid Method For Multiphase Elliptic Problems
by: Toprak, Teoman, et al.
Published: (2025)