CaPGNN: Optimizing Parallel Graph Neural Network Training with Joint Caching and Resource-Aware Graph Partitioning
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Xianfeng, Zou, Yi, Shi, Zheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GSplit: Scaling Graph Neural Network Training on Large Graphs via Split-Parallelism
by: Polisetty, Sandeep, et al.
Published: (2023)
by: Polisetty, Sandeep, et al.
Published: (2023)
An Experimental Comparison of Partitioning Strategies for Distributed Graph Neural Network Training
by: Merkel, Nikolai, et al.
Published: (2023)
by: Merkel, Nikolai, et al.
Published: (2023)
Heta: Distributed Training of Heterogeneous Graph Neural Networks
by: Zhong, Yuchen, et al.
Published: (2024)
by: Zhong, Yuchen, et al.
Published: (2024)
EmbedPart: Embedding-Driven Graph Partitioning for Scalable Graph Neural Network Training
by: Merkel, Nikolai, et al.
Published: (2026)
by: Merkel, Nikolai, et al.
Published: (2026)
Graph Neural Networks as Ordering Heuristics for Parallel Graph Coloring
by: Langedal, Kenneth, et al.
Published: (2024)
by: Langedal, Kenneth, et al.
Published: (2024)
WindGP: Efficient Graph Partitioning on Heterogenous Machines
by: Zeng, Li, et al.
Published: (2024)
by: Zeng, Li, et al.
Published: (2024)
10Cache: Heterogeneous Resource-Aware Tensor Caching and Migration for LLM Training
by: Afroz, Sabiha, et al.
Published: (2025)
by: Afroz, Sabiha, et al.
Published: (2025)
A Cascaded Graph Neural Network for Joint Root Cause Localization and Analysis in Edge Computing Environments
by: Fernando, Duneesha, et al.
Published: (2026)
by: Fernando, Duneesha, et al.
Published: (2026)
RapidGNN: Communication Efficient Large-Scale Distributed Training of Graph Neural Networks
by: Niam, Arefin, et al.
Published: (2025)
by: Niam, Arefin, et al.
Published: (2025)
NeutronTP: Load-Balanced Distributed Full-Graph GNN Training with Tensor Parallelism
by: Ai, Xin, et al.
Published: (2024)
by: Ai, Xin, et al.
Published: (2024)
pdGRASS: A Fast Parallel Density-Aware Algorithm for Graph Spectral Sparsification
by: Zhao, Tiancheng, et al.
Published: (2025)
by: Zhao, Tiancheng, et al.
Published: (2025)
Parallel Unconstrained Local Search for Partitioning Irregular Graphs
by: Maas, Nikolai, et al.
Published: (2023)
by: Maas, Nikolai, et al.
Published: (2023)
OptimES: Optimizing Federated Learning Using Remote Embeddings for Graph Neural Networks
by: Naman, Pranjal, et al.
Published: (2025)
by: Naman, Pranjal, et al.
Published: (2025)
CDFGNN: a Systematic Design of Cache-based Distributed Full-Batch Graph Neural Network Training with Communication Reduction
by: Zhang, Shuai, et al.
Published: (2024)
by: Zhang, Shuai, et al.
Published: (2024)
Deterministic Parallel High-Quality Hypergraph Partitioning
by: Krause, Robert, et al.
Published: (2025)
by: Krause, Robert, et al.
Published: (2025)
CacheFlow: Efficient LLM Serving with 3D-Parallel KV Cache Restoration
by: Nian, Sean, et al.
Published: (2026)
by: Nian, Sean, et al.
Published: (2026)
Optimizing Distributed Training Approaches for Scaling Neural Networks
by: Baligodugula, Vishnu Vardhan, et al.
Published: (2025)
by: Baligodugula, Vishnu Vardhan, et al.
Published: (2025)
CUTTANA: Scalable Graph Partitioning for Faster Distributed Graph Databases and Analytics
by: Hajidehi, Milad Rezaei, et al.
Published: (2023)
by: Hajidehi, Milad Rezaei, et al.
Published: (2023)
CFP: Efficient Optimization of Intra-Operator Parallelism Plans for Large Model Training
by: Hu, Weifang, et al.
Published: (2025)
by: Hu, Weifang, et al.
Published: (2025)
Parallel Order-Based Core Maintenance in Dynamic Graphs
by: Guo, Bin, et al.
Published: (2022)
by: Guo, Bin, et al.
Published: (2022)
Optimizing Hardware Resource Partitioning and Job Allocations on Modern GPUs under Power Caps
by: Arima, Eishi, et al.
Published: (2024)
by: Arima, Eishi, et al.
Published: (2024)
Efficient Task Graph Scheduling for Parallel QR Factorization in SLSQP
by: Chatterjee, Soumyajit, et al.
Published: (2025)
by: Chatterjee, Soumyajit, et al.
Published: (2025)
Distributed And Parallel Low-Diameter Decompositions for Arbitrary and Restricted Graphs
by: Dou, Jinfeng, et al.
Published: (2024)
by: Dou, Jinfeng, et al.
Published: (2024)
Faster Parallel Triangular Maximally Filtered Graphs and Hierarchical Clustering
by: Raphael, Steven, et al.
Published: (2024)
by: Raphael, Steven, et al.
Published: (2024)
Leiden-Fusion Partitioning Method for Effective Distributed Training of Graph Embeddings
by: Bai, Yuhe, et al.
Published: (2024)
by: Bai, Yuhe, et al.
Published: (2024)
Cortex: Achieving Low-Latency, Cost-Efficient Remote Data Access For LLM via Semantic-Aware Knowledge Caching
by: Ruan, Chaoyi, et al.
Published: (2025)
by: Ruan, Chaoyi, et al.
Published: (2025)
Fully Distributed Online Training of Graph Neural Networks in Networked Systems
by: Olshevskyi, Rostyslav, et al.
Published: (2024)
by: Olshevskyi, Rostyslav, et al.
Published: (2024)
ReInc: Scaling Training of Dynamic Graph Neural Networks
by: Guan, Mingyu, et al.
Published: (2025)
by: Guan, Mingyu, et al.
Published: (2025)
Graph Neural Network Training Systems: A Performance Comparison of Full-Graph and Mini-Batch
by: Bajaj, Saurabh, et al.
Published: (2024)
by: Bajaj, Saurabh, et al.
Published: (2024)
A Parallel and Distributed Rust Library for Core Decomposition on Large Graphs
by: Rucci, Davide, et al.
Published: (2025)
by: Rucci, Davide, et al.
Published: (2025)
A Semantic Partitioning Method for Large-Scale Training of Knowledge Graph Embeddings
by: Bai, Yuhe
Published: (2025)
by: Bai, Yuhe
Published: (2025)
Bandwidth-Aware and Cost-Efficient Pipeline Parallel Scheduling in Geo-Distributed LLM Training
by: Zhang, Han, et al.
Published: (2026)
by: Zhang, Han, et al.
Published: (2026)
Spatiotemporal Traffic Prediction in Distributed Backend Systems via Graph Neural Networks
by: Qiu, Zhimin, et al.
Published: (2025)
by: Qiu, Zhimin, et al.
Published: (2025)
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
by: Naman, Pranjal, et al.
Published: (2026)
by: Naman, Pranjal, et al.
Published: (2026)
CondenseGraph: Communication-Efficient Distributed GNN Training via On-the-Fly Graph Condensation
by: Zhang, Zizhao, et al.
Published: (2026)
by: Zhang, Zizhao, et al.
Published: (2026)
UniEP: Unified Expert-Parallel MoE MegaKernel for LLM Training
by: Zheng, Size, et al.
Published: (2026)
by: Zheng, Size, et al.
Published: (2026)
LuWu: An End-to-End In-Network Out-of-Core Optimizer for 100B-Scale Model-in-Network Data-Parallel Training on Distributed GPUs
by: Sun, Mo, et al.
Published: (2024)
by: Sun, Mo, et al.
Published: (2024)
Grappa: Gradient-Only Communication for Scalable Graph Neural Network Training
by: Xu, Chongyang, et al.
Published: (2026)
by: Xu, Chongyang, et al.
Published: (2026)
DawnPiper: A Memory-scablable Pipeline Parallel Training Framework
by: Peng, Xuan, et al.
Published: (2025)
by: Peng, Xuan, et al.
Published: (2025)
Partition Detection in Byzantine Networks
by: Bromberg, Yérom-David, et al.
Published: (2024)
by: Bromberg, Yérom-David, et al.
Published: (2024)
Similar Items
-
GSplit: Scaling Graph Neural Network Training on Large Graphs via Split-Parallelism
by: Polisetty, Sandeep, et al.
Published: (2023) -
An Experimental Comparison of Partitioning Strategies for Distributed Graph Neural Network Training
by: Merkel, Nikolai, et al.
Published: (2023) -
Heta: Distributed Training of Heterogeneous Graph Neural Networks
by: Zhong, Yuchen, et al.
Published: (2024) -
EmbedPart: Embedding-Driven Graph Partitioning for Scalable Graph Neural Network Training
by: Merkel, Nikolai, et al.
Published: (2026) -
Graph Neural Networks as Ordering Heuristics for Parallel Graph Coloring
by: Langedal, Kenneth, et al.
Published: (2024)