Optimizing Distributed Training Approaches for Scaling Neural Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Baligodugula, Vishnu Vardhan, Amsaad, Fathi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RapidGNN: Communication Efficient Large-Scale Distributed Training of Graph Neural Networks
por: Niam, Arefin, et al.
Publicado: (2025)
por: Niam, Arefin, et al.
Publicado: (2025)
Heta: Distributed Training of Heterogeneous Graph Neural Networks
por: Zhong, Yuchen, et al.
Publicado: (2024)
por: Zhong, Yuchen, et al.
Publicado: (2024)
LuWu: An End-to-End In-Network Out-of-Core Optimizer for 100B-Scale Model-in-Network Data-Parallel Training on Distributed GPUs
por: Sun, Mo, et al.
Publicado: (2024)
por: Sun, Mo, et al.
Publicado: (2024)
DeepCompile: A Compiler-Driven Approach to Optimizing Distributed Deep Learning Training
por: Tanaka, Masahiro, et al.
Publicado: (2025)
por: Tanaka, Masahiro, et al.
Publicado: (2025)
Armada: Memory-Efficient Distributed Training of Large-Scale Graph Neural Networks
por: Waleffe, Roger, et al.
Publicado: (2025)
por: Waleffe, Roger, et al.
Publicado: (2025)
BANG: Billion-Scale Approximate Nearest Neighbor Search using a Single GPU
por: V., Karthik, et al.
Publicado: (2024)
por: V., Karthik, et al.
Publicado: (2024)
MalleTrain: Deep Neural Network Training on Unfillable Supercomputer Nodes
por: Ma, Xiaolong, et al.
Publicado: (2024)
por: Ma, Xiaolong, et al.
Publicado: (2024)
Poplar: Efficient Scaling of Distributed DNN Training on Heterogeneous GPU Clusters
por: Zhang, WenZheng, et al.
Publicado: (2024)
por: Zhang, WenZheng, et al.
Publicado: (2024)
CaPGNN: Optimizing Parallel Graph Neural Network Training with Joint Caching and Resource-Aware Graph Partitioning
por: Song, Xianfeng, et al.
Publicado: (2025)
por: Song, Xianfeng, et al.
Publicado: (2025)
PRISM: Probabilistic Runtime Insights and Scalable Performance Modeling for Large-Scale Distributed Training
por: Golden, Alicia, et al.
Publicado: (2025)
por: Golden, Alicia, et al.
Publicado: (2025)
Optimizing Frequent Checkpointing via Low-Cost Differential for Distributed Training Systems
por: Yao, Chenxuan, et al.
Publicado: (2025)
por: Yao, Chenxuan, et al.
Publicado: (2025)
Nezha: Breaking Multi-Rail Network Barriers for Distributed DNN Training
por: Yu, Enda, et al.
Publicado: (2024)
por: Yu, Enda, et al.
Publicado: (2024)
Effectiveness of Distributed Gradient Descent with Local Steps for Overparameterized Models
por: Zhu, Heng, et al.
Publicado: (2024)
por: Zhu, Heng, et al.
Publicado: (2024)
Optimizing High-Throughput Distributed Data Pipelines for Reproducible Deep Learning at Scale
por: Mittal, Kashish, et al.
Publicado: (2026)
por: Mittal, Kashish, et al.
Publicado: (2026)
Distributed Generative Inference of LLM at Internet Scales with Multi-Dimensional Communication Optimization
por: Chen, Jiu, et al.
Publicado: (2026)
por: Chen, Jiu, et al.
Publicado: (2026)
Multi-Resolution Model Fusion for Accelerating the Convolutional Neural Network Training
por: Wang, Kewei, et al.
Publicado: (2025)
por: Wang, Kewei, et al.
Publicado: (2025)
Fully Distributed Online Training of Graph Neural Networks in Networked Systems
por: Olshevskyi, Rostyslav, et al.
Publicado: (2024)
por: Olshevskyi, Rostyslav, et al.
Publicado: (2024)
Embedded Distributed Inference of Deep Neural Networks: A Systematic Review
por: Peccia, Federico Nicolás, et al.
Publicado: (2024)
por: Peccia, Federico Nicolás, et al.
Publicado: (2024)
Diagonal Scaling: A Multi-Dimensional Resource Model and Optimization Framework for Distributed Databases
por: Abdullah, Shahir, et al.
Publicado: (2025)
por: Abdullah, Shahir, et al.
Publicado: (2025)
ReInc: Scaling Training of Dynamic Graph Neural Networks
por: Guan, Mingyu, et al.
Publicado: (2025)
por: Guan, Mingyu, et al.
Publicado: (2025)
Distributed Convolutional Neural Network Training on Mobile and Edge Clusters
por: Rama, Pranav, et al.
Publicado: (2024)
por: Rama, Pranav, et al.
Publicado: (2024)
ACE-Sync: An Adaptive Cloud-Edge Synchronization Framework for Communication-Efficient Large-Scale Distributed Model Training
por: Yang, Yi, et al.
Publicado: (2025)
por: Yang, Yi, et al.
Publicado: (2025)
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
por: Naman, Pranjal, et al.
Publicado: (2026)
por: Naman, Pranjal, et al.
Publicado: (2026)
Spatiotemporal Traffic Prediction in Distributed Backend Systems via Graph Neural Networks
por: Qiu, Zhimin, et al.
Publicado: (2025)
por: Qiu, Zhimin, et al.
Publicado: (2025)
A Distributed Approach for Persistent Homology Computation on a Large Scale
por: Ceccaroni, Riccardo, et al.
Publicado: (2024)
por: Ceccaroni, Riccardo, et al.
Publicado: (2024)
Efficient Distributed MLLM Training with Cornstarch
por: Jang, Insu, et al.
Publicado: (2025)
por: Jang, Insu, et al.
Publicado: (2025)
An Experimental Comparison of Partitioning Strategies for Distributed Graph Neural Network Training
por: Merkel, Nikolai, et al.
Publicado: (2023)
por: Merkel, Nikolai, et al.
Publicado: (2023)
NEST: Network- and Memory-Aware Device Placement For Distributed Deep Learning
por: Wang, Irene, et al.
Publicado: (2026)
por: Wang, Irene, et al.
Publicado: (2026)
Distributed Constrained Combinatorial Optimization leveraging Hypergraph Neural Networks
por: Heydaribeni, Nasimeh, et al.
Publicado: (2023)
por: Heydaribeni, Nasimeh, et al.
Publicado: (2023)
Echo: Simulating Distributed Training At Scale
por: Feng, Yicheng, et al.
Publicado: (2024)
por: Feng, Yicheng, et al.
Publicado: (2024)
Stable-MoE: Lyapunov-based Token Routing for Distributed Mixture-of-Experts Training over Edge Networks
por: Shi, Long, et al.
Publicado: (2025)
por: Shi, Long, et al.
Publicado: (2025)
OptimES: Optimizing Federated Learning Using Remote Embeddings for Graph Neural Networks
por: Naman, Pranjal, et al.
Publicado: (2025)
por: Naman, Pranjal, et al.
Publicado: (2025)
SDT-GNN: Streaming-based Distributed Training Framework for Graph Neural Networks
por: Huang, Xin, et al.
Publicado: (2024)
por: Huang, Xin, et al.
Publicado: (2024)
Galvatron: Automatic Distributed Training for Large Transformer Models
por: Gumaan, Esmail
Publicado: (2025)
por: Gumaan, Esmail
Publicado: (2025)
Addressing Variable Heterogeneity in Distributed Multimodal Training with Entrain
por: Jang, Insu, et al.
Publicado: (2026)
por: Jang, Insu, et al.
Publicado: (2026)
Accelerating Distributed MoE Training and Inference with Lina
por: Li, Jiamin, et al.
Publicado: (2022)
por: Li, Jiamin, et al.
Publicado: (2022)
Collaborative UAVs Multi-task Video Processing Optimization Based on Enhanced Distributed Actor-Critic Networks
por: Rong, Ziqi, et al.
Publicado: (2024)
por: Rong, Ziqi, et al.
Publicado: (2024)
GSplit: Scaling Graph Neural Network Training on Large Graphs via Split-Parallelism
por: Polisetty, Sandeep, et al.
Publicado: (2023)
por: Polisetty, Sandeep, et al.
Publicado: (2023)
Federated Neural Radiance Field for Distributed Intelligence
por: Zhang, Yintian, et al.
Publicado: (2024)
por: Zhang, Yintian, et al.
Publicado: (2024)
MegatronApp: Efficient and Comprehensive Management on Distributed LLM Training
por: Zhao, Bohan, et al.
Publicado: (2025)
por: Zhao, Bohan, et al.
Publicado: (2025)
Ejemplares similares
-
RapidGNN: Communication Efficient Large-Scale Distributed Training of Graph Neural Networks
por: Niam, Arefin, et al.
Publicado: (2025) -
Heta: Distributed Training of Heterogeneous Graph Neural Networks
por: Zhong, Yuchen, et al.
Publicado: (2024) -
LuWu: An End-to-End In-Network Out-of-Core Optimizer for 100B-Scale Model-in-Network Data-Parallel Training on Distributed GPUs
por: Sun, Mo, et al.
Publicado: (2024) -
DeepCompile: A Compiler-Driven Approach to Optimizing Distributed Deep Learning Training
por: Tanaka, Masahiro, et al.
Publicado: (2025) -
Armada: Memory-Efficient Distributed Training of Large-Scale Graph Neural Networks
por: Waleffe, Roger, et al.
Publicado: (2025)