Saved in:
Bibliographic Details
Main Authors: Sun, Wenbo, Hai, Rihan
Format: Preprint
Published: 2025
Subjects:
Online Access:https://arxiv.org/abs/2502.01985
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916597029928960
author Sun, Wenbo
Hai, Rihan
author_facet Sun, Wenbo
Hai, Rihan
contents The machine learning (ML) training over disparate data sources traditionally involves materialization, which can impose substantial time and space overhead due to data movement and replication. Factorized learning, which leverages direct computation on disparate sources through linear algebra (LA) rewriting, has emerged as a viable alternative to improve computational efficiency. However, the adaptation of factorized learning to leverage the full capabilities of modern LA-friendly hardware like GPUs has been limited, often requiring manual intervention for algorithm compatibility. This paper introduces Ilargi, a novel factorized learning framework that utilizes matrix-represented data integration (DI) metadata to facilitate automatic factorization across CPU and GPU environments without the need for costly relational joins. Ilargi incorporates an ML-based cost estimator to intelligently selects between factorization and materialization based on data properties, algorithm complexity, hardware environments, and their interactions. This strategy ensures up to 8.9x speedups on GPUs and achieves over 20% acceleration in batch ML training workloads, thereby enhancing the practicability of ML training across diverse data integration scenarios and hardware platforms. To our knowledge, this work is the very first effort in GPU-compatible factorized learning.
format Preprint
id arxiv_https___arxiv_org_abs_2502_01985
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Ilargi: a GPU Compatible Factorized ML Model Training Framework
Sun, Wenbo
Hai, Rihan
Machine Learning
Distributed, Parallel, and Cluster Computing
The machine learning (ML) training over disparate data sources traditionally involves materialization, which can impose substantial time and space overhead due to data movement and replication. Factorized learning, which leverages direct computation on disparate sources through linear algebra (LA) rewriting, has emerged as a viable alternative to improve computational efficiency. However, the adaptation of factorized learning to leverage the full capabilities of modern LA-friendly hardware like GPUs has been limited, often requiring manual intervention for algorithm compatibility. This paper introduces Ilargi, a novel factorized learning framework that utilizes matrix-represented data integration (DI) metadata to facilitate automatic factorization across CPU and GPU environments without the need for costly relational joins. Ilargi incorporates an ML-based cost estimator to intelligently selects between factorization and materialization based on data properties, algorithm complexity, hardware environments, and their interactions. This strategy ensures up to 8.9x speedups on GPUs and achieves over 20% acceleration in batch ML training workloads, thereby enhancing the practicability of ML training across diverse data integration scenarios and hardware platforms. To our knowledge, this work is the very first effort in GPU-compatible factorized learning.
title Ilargi: a GPU Compatible Factorized ML Model Training Framework
topic Machine Learning
Distributed, Parallel, and Cluster Computing
url https://arxiv.org/abs/2502.01985