Accelerating Mixed-Precision Out-of-Core Cholesky Factorization with Static Task Scheduling
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Jie, Ltaief, Hatem, Abdulah, Sameh, Keyes, David E. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GPU-Accelerated Modified Bessel Function of the Second Kind for Gaussian Processes
by: Geng, Zipei, et al.
Published: (2025)
by: Geng, Zipei, et al.
Published: (2025)
GPU-Accelerated Vecchia Approximations of Gaussian Processes for Geospatial Data using Batched Matrix Computations
by: Pan, Qilong, et al.
Published: (2024)
by: Pan, Qilong, et al.
Published: (2024)
Parallel Approximations for High-Dimensional Multivariate Normal Probability Computation in Confidence Region Detection Applications
by: Zhang, Xiran, et al.
Published: (2024)
by: Zhang, Xiran, et al.
Published: (2024)
Scaled Block Vecchia Approximation for High-Dimensional Gaussian Process Emulation on GPUs
by: Pan, Qilong, et al.
Published: (2025)
by: Pan, Qilong, et al.
Published: (2025)
High-Performance Statistical Computing (HPSC): Challenges, Opportunities, and Future Directions
by: Abdulah, Sameh, et al.
Published: (2025)
by: Abdulah, Sameh, et al.
Published: (2025)
GPU Accelerated Sparse Cholesky Factorization
by: Karsavuran, M. Ozan, et al.
Published: (2024)
by: Karsavuran, M. Ozan, et al.
Published: (2024)
Cross-Layer Energy Analysis of Multimodal Training on Grace Hopper Superchips
by: Ahmed, Mahmoud, et al.
Published: (2026)
by: Ahmed, Mahmoud, et al.
Published: (2026)
RCOMPSs: A Scalable Runtime System for R Code Execution on Manycore Systems
by: Zhang, Xiran, et al.
Published: (2025)
by: Zhang, Xiran, et al.
Published: (2025)
A Novel Approach to Translate Structural Aggregation Queries to MapReduce Code
by: Abdelmoniem, Ahmed M., et al.
Published: (2025)
by: Abdelmoniem, Ahmed M., et al.
Published: (2025)
Neural Acceleration of Incomplete Cholesky Preconditioners
by: Booth, Joshua Dennis, et al.
Published: (2024)
by: Booth, Joshua Dennis, et al.
Published: (2024)
Efficient Task Graph Scheduling for Parallel QR Factorization in SLSQP
by: Chatterjee, Soumyajit, et al.
Published: (2025)
by: Chatterjee, Soumyajit, et al.
Published: (2025)
A Granularity Characterization of Task Scheduling Effectiveness
by: Anvari, Sana Taghipour, et al.
Published: (2026)
by: Anvari, Sana Taghipour, et al.
Published: (2026)
Fast Algorithms for Scheduling Many-body Correlation Functions on Accelerators
by: Selvitopi, Oguz, et al.
Published: (2025)
by: Selvitopi, Oguz, et al.
Published: (2025)
FedPBS: Proximal-Balanced Scaling Federated Learning Model for Robust Personalized Training for Non-IID Data
by: AbouNassar, Eman M., et al.
Published: (2026)
by: AbouNassar, Eman M., et al.
Published: (2026)
Taming GPU Underutilization via Static Partitioning and Fine-grained CPU Offloading
by: Schieffer, Gabin, et al.
Published: (2026)
by: Schieffer, Gabin, et al.
Published: (2026)
LMDeploy Accelerates Mixed-Precision LLM Inference with TurboMind
by: Zhang, Li, et al.
Published: (2025)
by: Zhang, Li, et al.
Published: (2025)
DaggerFFT: A Distributed FFT Framework Using Task Scheduling in Julia
by: Anvari, Sana Taghipour, et al.
Published: (2026)
by: Anvari, Sana Taghipour, et al.
Published: (2026)
Task Scheduling in Geo-Distributed Computing: A Survey
by: Wu, Yujian, et al.
Published: (2025)
by: Wu, Yujian, et al.
Published: (2025)
Fast and Scalable Mixed Precision Euclidean Distance Calculations Using GPU Tensor Cores
by: Curless, Brian, et al.
Published: (2025)
by: Curless, Brian, et al.
Published: (2025)
ARGO: An Auto-Tuning Runtime System for Scalable GNN Training on Multi-Core Processor
by: Lin, Yi-Chien, et al.
Published: (2024)
by: Lin, Yi-Chien, et al.
Published: (2024)
Alternative Mixed Integer Linear Programming Optimization for Joint Job Scheduling and Data Allocation in Grid Computing
by: Feng, Shengyu, et al.
Published: (2025)
by: Feng, Shengyu, et al.
Published: (2025)
Scheduling Coflows in Multi-Core OCS Networks with Performance Guarantee
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
Linear Complexity $\mathcal{H}^2$ Direct Solver for Fine-Grained Parallel Architectures
by: Boukaram, Wajih, et al.
Published: (2025)
by: Boukaram, Wajih, et al.
Published: (2025)
PISA: An Adversarial Approach To Comparing Task Graph Scheduling Algorithms
by: Coleman, Jared, et al.
Published: (2024)
by: Coleman, Jared, et al.
Published: (2024)
Parameterized Task Graph Scheduling Algorithm for Comparing Algorithmic Components
by: Coleman, Jared, et al.
Published: (2024)
by: Coleman, Jared, et al.
Published: (2024)
INSPIRIT: Optimizing Heterogeneous Task Scheduling through Adaptive Priority in Task-based Runtime Systems
by: Wang, Yiqing, et al.
Published: (2024)
by: Wang, Yiqing, et al.
Published: (2024)
Hyperion: Hierarchical Scheduling for Parallel LLM Acceleration in Multi-tier Networks
by: Ma, Mulei, et al.
Published: (2025)
by: Ma, Mulei, et al.
Published: (2025)
Data-Locality-Aware Task Assignment and Scheduling for Distributed Job Executions
by: Zhao, Hailiang, et al.
Published: (2024)
by: Zhao, Hailiang, et al.
Published: (2024)
A Performance Analysis of Task Scheduling for UQ Workflows on HPC Systems
by: Loi, Chung Ming, et al.
Published: (2025)
by: Loi, Chung Ming, et al.
Published: (2025)
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
by: Naman, Pranjal, et al.
Published: (2026)
by: Naman, Pranjal, et al.
Published: (2026)
QoS Aware Mixed-Criticality Task Scheduling in Vehicular Edge Cloud System
by: Sarkar, Suvarthi, et al.
Published: (2024)
by: Sarkar, Suvarthi, et al.
Published: (2024)
Efficient Scheduling of Vehicular Tasks on Edge Systems with Green Energy and Battery Storage
by: Sarkar, Suvarthi, et al.
Published: (2024)
by: Sarkar, Suvarthi, et al.
Published: (2024)
WOW: Workflow-Aware Data Movement and Task Scheduling for Dynamic Scientific Workflows
by: Lehmann, Fabian, et al.
Published: (2025)
by: Lehmann, Fabian, et al.
Published: (2025)
Energy-Aware Scheduling Strategies for Partially-Replicable Task Chains on Heterogeneous Processors
by: Idouar, Yacine, et al.
Published: (2025)
by: Idouar, Yacine, et al.
Published: (2025)
Preemption Aware Task Scheduling for Priority and Deadline Constrained DNN Inference Task Offloading in Homogeneous Mobile-Edge Networks
by: Cotter, Jamie, et al.
Published: (2025)
by: Cotter, Jamie, et al.
Published: (2025)
PICO: Accelerating All k-Core Paradigms on GPU
by: Zhao, Chen, et al.
Published: (2024)
by: Zhao, Chen, et al.
Published: (2024)
O(K)-Approximation Coflow Scheduling in K-Core Optical Circuit Switching Networks
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
Eliminating Timing Anomalies in Scheduling Periodic Segmented Self-Suspending Tasks with Release Jitter
by: Lin, Ching-Chi, et al.
Published: (2024)
by: Lin, Ching-Chi, et al.
Published: (2024)
GCAPS: GPU Context-Aware Preemptive Priority-based Scheduling for Real-Time Tasks
by: Wang, Yidi, et al.
Published: (2024)
by: Wang, Yidi, et al.
Published: (2024)
A Reinforcement Learning-Driven Task Scheduling Algorithm for Multi-Tenant Distributed Systems
by: Zhang, Xiaopei, et al.
Published: (2025)
by: Zhang, Xiaopei, et al.
Published: (2025)
Similar Items
-
GPU-Accelerated Modified Bessel Function of the Second Kind for Gaussian Processes
by: Geng, Zipei, et al.
Published: (2025) -
GPU-Accelerated Vecchia Approximations of Gaussian Processes for Geospatial Data using Batched Matrix Computations
by: Pan, Qilong, et al.
Published: (2024) -
Parallel Approximations for High-Dimensional Multivariate Normal Probability Computation in Confidence Region Detection Applications
by: Zhang, Xiran, et al.
Published: (2024) -
Scaled Block Vecchia Approximation for High-Dimensional Gaussian Process Emulation on GPUs
by: Pan, Qilong, et al.
Published: (2025) -
High-Performance Statistical Computing (HPSC): Challenges, Opportunities, and Future Directions
by: Abdulah, Sameh, et al.
Published: (2025)