Linear Complexity $\mathcal{H}^2$ Direct Solver for Fine-Grained Parallel Architectures
Fuente:
arXiv
Saved in:
| Main Authors: | Boukaram, Wajih, Keyes, David, Li, Sherry, Liu, Yang, Turkiyyah, George |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Parallel Sparse and Data-Sparse Factorization-based Linear Solvers
by: Li, Xiaoye Sherry, et al.
Published: (2026)
by: Li, Xiaoye Sherry, et al.
Published: (2026)
SDSL-Solver: Scalable Distributed Sparse Linear Solvers for Large-Scale Interior Point Methods
by: Yang, Shaofeng, et al.
Published: (2026)
by: Yang, Shaofeng, et al.
Published: (2026)
Optimizing Long-context LLM Serving via Fine-grained Sequence Parallelism
by: Li, Cong, et al.
Published: (2025)
by: Li, Cong, et al.
Published: (2025)
High-Performance Statistical Computing (HPSC): Challenges, Opportunities, and Future Directions
by: Abdulah, Sameh, et al.
Published: (2025)
by: Abdulah, Sameh, et al.
Published: (2025)
Fine-Grained Energy Prediction For Parallellized LLM Inference With PIE-P
by: Dutt, Anurag, et al.
Published: (2025)
by: Dutt, Anurag, et al.
Published: (2025)
Multi-GPU Acceleration of PALABOS Fluid Solver using C++ Standard Parallelism
by: Latt, Jonas, et al.
Published: (2025)
by: Latt, Jonas, et al.
Published: (2025)
Parallel Approximations for High-Dimensional Multivariate Normal Probability Computation in Confidence Region Detection Applications
by: Zhang, Xiran, et al.
Published: (2024)
by: Zhang, Xiran, et al.
Published: (2024)
A Parallel and Highly-Portable HPC Poisson Solver: Preconditioned Bi-CGSTAB with alpaka
by: Pennati, Luca, et al.
Published: (2025)
by: Pennati, Luca, et al.
Published: (2025)
FedFQ: Federated Learning with Fine-Grained Quantization
by: Li, Haowei, et al.
Published: (2024)
by: Li, Haowei, et al.
Published: (2024)
Accelerating Mixed-Precision Out-of-Core Cholesky Factorization with Static Task Scheduling
by: Ren, Jie, et al.
Published: (2024)
by: Ren, Jie, et al.
Published: (2024)
LiveR: Fine-Grained Elasticity via Live Reconfiguration for Model Training
by: Liu, Haoyuan, et al.
Published: (2026)
by: Liu, Haoyuan, et al.
Published: (2026)
How Fast Can Graph Computations Go on Fine-grained Parallel Architectures
by: Wang, Yuqing, et al.
Published: (2025)
by: Wang, Yuqing, et al.
Published: (2025)
MemFine: Memory-Aware Fine-Grained Scheduling for MoE Training
by: Zhao, Lu, et al.
Published: (2025)
by: Zhao, Lu, et al.
Published: (2025)
Exploring Fine-grained Task Parallelism on Simultaneous Multithreading Cores
by: Los, Denis, et al.
Published: (2024)
by: Los, Denis, et al.
Published: (2024)
Heterogeneous Federated Fine-Tuning with Parallel One-Rank Adaptation
by: Zhang, Zikai, et al.
Published: (2026)
by: Zhang, Zikai, et al.
Published: (2026)
The Merit of Simple Policies: Buying Performance With Parallelism and System Architecture
by: Yildiz, Mert, et al.
Published: (2025)
by: Yildiz, Mert, et al.
Published: (2025)
PACE Solver Description: twin_width_fmi
by: Balaban, David, et al.
Published: (2025)
by: Balaban, David, et al.
Published: (2025)
torch-sla: Differentiable Sparse Linear Algebra with Adjoint Solvers and Sparse Tensor Parallelism for PyTorch
by: Chi, Mingyuan, et al.
Published: (2026)
by: Chi, Mingyuan, et al.
Published: (2026)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
by: Mo, Zizhao, et al.
Published: (2025)
by: Mo, Zizhao, et al.
Published: (2025)
Efficient MoE Inference with Fine-Grained Scheduling of Disaggregated Expert Parallelism
by: Pan, Xinglin, et al.
Published: (2025)
by: Pan, Xinglin, et al.
Published: (2025)
Benchmarking the Parallel 1D Heat Equation Solver in Chapel, Charm++, C++, HPX, Go, Julia, Python, Rust, Swift, and Java
by: Diehl, Patrick, et al.
Published: (2023)
by: Diehl, Patrick, et al.
Published: (2023)
Parallel AIG Refactoring via Conflict Breaking
by: Cai, Ye, et al.
Published: (2024)
by: Cai, Ye, et al.
Published: (2024)
Fast Sparse Matrix Permutation for Mesh-Based Direct Solvers
by: Zarebavami, Behrooz, et al.
Published: (2026)
by: Zarebavami, Behrooz, et al.
Published: (2026)
A Simulated Annealing Approach to Identical Parallel Machine Scheduling
by: Li, Jiaxing, et al.
Published: (2024)
by: Li, Jiaxing, et al.
Published: (2024)
A Spark Optimizer for Adaptive, Fine-Grained Parameter Tuning
by: Lyu, Chenghao, et al.
Published: (2024)
by: Lyu, Chenghao, et al.
Published: (2024)
A Framework for Fine-Grained Synchronization of Dependent GPU Kernels
by: Jangda, Abhinav, et al.
Published: (2023)
by: Jangda, Abhinav, et al.
Published: (2023)
Towards Fine-Grained Scalability for Stateful Stream Processing Systems
by: Qing, Yunfan, et al.
Published: (2025)
by: Qing, Yunfan, et al.
Published: (2025)
TD-Pipe: Temporally-Disaggregated Pipeline Parallelism Architecture for High-Throughput LLM Inference
by: Zhang, Hongbin, et al.
Published: (2025)
by: Zhang, Hongbin, et al.
Published: (2025)
Distributed Variational Quantum Linear Solver
by: Lu, Chao, et al.
Published: (2026)
by: Lu, Chao, et al.
Published: (2026)
Fine-grained MoE Load Balancing with Linear Programming
by: Zhao, Chenqi, et al.
Published: (2025)
by: Zhao, Chenqi, et al.
Published: (2025)
Persistent HyTM via Fast Path Fine-Grained Locking
by: Coccimiglio, Gaetano, et al.
Published: (2025)
by: Coccimiglio, Gaetano, et al.
Published: (2025)
A Massively Parallel Performance Portable Free-space Spectral Poisson Solver
by: Mayani, Sonali, et al.
Published: (2024)
by: Mayani, Sonali, et al.
Published: (2024)
GPU-Accelerated Modified Bessel Function of the Second Kind for Gaussian Processes
by: Geng, Zipei, et al.
Published: (2025)
by: Geng, Zipei, et al.
Published: (2025)
ParaLiNGAM: Parallel Causal Structure Learning for Linear non-Gaussian Acyclic Models
by: Shahbazinia, Amirhossein, et al.
Published: (2021)
by: Shahbazinia, Amirhossein, et al.
Published: (2021)
FASER: Fine-Grained Phase Management for Speculative Decoding in Dynamic LLM Serving
by: Chen, Wenyan, et al.
Published: (2026)
by: Chen, Wenyan, et al.
Published: (2026)
Shared Memory-Aware Latency-Sensitive Message Aggregation for Fine-Grained Communication
by: Chandrasekar, Kavitha, et al.
Published: (2024)
by: Chandrasekar, Kavitha, et al.
Published: (2024)
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
by: Cheng, Minyu, et al.
Published: (2026)
by: Cheng, Minyu, et al.
Published: (2026)
Fine-Grained Vectorized Merge Sorting on RISC-V: From Register to Cache
by: Zhang, Jin, et al.
Published: (2024)
by: Zhang, Jin, et al.
Published: (2024)
Parallel Online Directed Acyclic Graph Exploration for Atlasing Soft-Matter Assembly Configuration Spaces
by: Prabhu, Rahul, et al.
Published: (2024)
by: Prabhu, Rahul, et al.
Published: (2024)
Training Through Failure: Effects of Data Consistency in Parallel Machine Learning Training
by: Cao, Ray, et al.
Published: (2024)
by: Cao, Ray, et al.
Published: (2024)
Similar Items
-
Parallel Sparse and Data-Sparse Factorization-based Linear Solvers
by: Li, Xiaoye Sherry, et al.
Published: (2026) -
SDSL-Solver: Scalable Distributed Sparse Linear Solvers for Large-Scale Interior Point Methods
by: Yang, Shaofeng, et al.
Published: (2026) -
Optimizing Long-context LLM Serving via Fine-grained Sequence Parallelism
by: Li, Cong, et al.
Published: (2025) -
High-Performance Statistical Computing (HPSC): Challenges, Opportunities, and Future Directions
by: Abdulah, Sameh, et al.
Published: (2025) -
Fine-Grained Energy Prediction For Parallellized LLM Inference With PIE-P
by: Dutt, Anurag, et al.
Published: (2025)