PruneX: A Hierarchical Communication-Efficient System for Distributed CNN Training with Structured Pruning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Olama, Alireza, Lundell, Andreas, Hajj, Izzat El, Lilius, Johan, Björkqvist, Jerker |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Federated Learning With L0 Constraint Via Probabilistic Gates For Sparsity
von: Huthasana, Krishna Harsha Kovelakuntla, et al.
Veröffentlicht: (2025)
von: Huthasana, Krishna Harsha Kovelakuntla, et al.
Veröffentlicht: (2025)
Faster Vertex Cover Algorithms on GPUs with Component-Aware Parallel Branching
von: Amro, Hussein, et al.
Veröffentlicht: (2025)
von: Amro, Hussein, et al.
Veröffentlicht: (2025)
Pruning Blockchain Protocols for Efficient Access Control in IoT Systems
von: Huang, Yongtao, et al.
Veröffentlicht: (2024)
von: Huang, Yongtao, et al.
Veröffentlicht: (2024)
Parallelizing Maximal Clique Enumeration on GPUs
von: Almasri, Mohammad, et al.
Veröffentlicht: (2022)
von: Almasri, Mohammad, et al.
Veröffentlicht: (2022)
PacTrain: Pruning and Adaptive Sparse Gradient Compression for Efficient Collective Communication in Distributed Deep Learning
von: Wang, Yisu, et al.
Veröffentlicht: (2025)
von: Wang, Yisu, et al.
Veröffentlicht: (2025)
ALPHA-PIM: Analysis of Linear Algebraic Processing for High-Performance Graph Applications on a Real Processing-In-Memory System
von: Barkhordar, Marzieh, et al.
Veröffentlicht: (2026)
von: Barkhordar, Marzieh, et al.
Veröffentlicht: (2026)
Distributed Bilevel Optimization with Dual Pruning for Resource-limited Clients
von: Li, Mingyi, et al.
Veröffentlicht: (2025)
von: Li, Mingyi, et al.
Veröffentlicht: (2025)
Efficient Column-Wise N:M Pruning on RISC-V CPU
von: Chu, Chi-Wei, et al.
Veröffentlicht: (2025)
von: Chu, Chi-Wei, et al.
Veröffentlicht: (2025)
Memory-Efficient Federated Fine-Tuning of Large Language Models via Layer Pruning
von: Wu, Yebo, et al.
Veröffentlicht: (2025)
von: Wu, Yebo, et al.
Veröffentlicht: (2025)
Environment-Aware Dynamic Pruning for Pipelined Edge Inference
von: O'Quinn, Austin, et al.
Veröffentlicht: (2025)
von: O'Quinn, Austin, et al.
Veröffentlicht: (2025)
Rafture: Erasure-coded Raft with Post-Dissemination Pruning
von: Kerur, Rithwik, et al.
Veröffentlicht: (2026)
von: Kerur, Rithwik, et al.
Veröffentlicht: (2026)
ColonyOS -- A Meta-Operating System for Distributed Computing Across Heterogeneous Platform
von: Kristiansson, Johan
Veröffentlicht: (2024)
von: Kristiansson, Johan
Veröffentlicht: (2024)
Unity is Power: Semi-Asynchronous Collaborative Training of Large-Scale Models with Structured Pruning in Resource-Limited Clients
von: Li, Yan, et al.
Veröffentlicht: (2024)
von: Li, Yan, et al.
Veröffentlicht: (2024)
RapidGNN: Communication Efficient Large-Scale Distributed Training of Graph Neural Networks
von: Niam, Arefin, et al.
Veröffentlicht: (2025)
von: Niam, Arefin, et al.
Veröffentlicht: (2025)
CondenseGraph: Communication-Efficient Distributed GNN Training via On-the-Fly Graph Condensation
von: Zhang, Zizhao, et al.
Veröffentlicht: (2026)
von: Zhang, Zizhao, et al.
Veröffentlicht: (2026)
GreenDyGNN: Runtime-Adaptive Energy-Efficient Communication for Distributed GNN Training
von: Niam, Arefin, et al.
Veröffentlicht: (2026)
von: Niam, Arefin, et al.
Veröffentlicht: (2026)
MTGenRec: An Efficient Distributed Training System for Generative Recommendation Models in Meituan
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025)
Efficient Distributed MLLM Training with Cornstarch
von: Jang, Insu, et al.
Veröffentlicht: (2025)
von: Jang, Insu, et al.
Veröffentlicht: (2025)
Federated Learning based on Pruning and Recovery
von: Ma, Chengjie
Veröffentlicht: (2024)
von: Ma, Chengjie
Veröffentlicht: (2024)
FlexFL: Heterogeneous Federated Learning via APoZ-Guided Flexible Pruning in Uncertain Scenarios
von: Chen, Zekai, et al.
Veröffentlicht: (2024)
von: Chen, Zekai, et al.
Veröffentlicht: (2024)
Efficient Hierarchical Storage Management Framework Empowered by Reinforcement Learning
von: Zhang, Tianru, et al.
Veröffentlicht: (2022)
von: Zhang, Tianru, et al.
Veröffentlicht: (2022)
ACE-Sync: An Adaptive Cloud-Edge Synchronization Framework for Communication-Efficient Large-Scale Distributed Model Training
von: Yang, Yi, et al.
Veröffentlicht: (2025)
von: Yang, Yi, et al.
Veröffentlicht: (2025)
Enhanced Pruning for Distributed Closeness Centrality under Multi-Packet Messaging
von: Manya, Patrick D., et al.
Veröffentlicht: (2025)
von: Manya, Patrick D., et al.
Veröffentlicht: (2025)
FedAPTA: Federated Multi-task Learning for Heterogeneous Devices with Adaptive Layer-wise Pruning and Task-aware Aggregation
von: Yu, Zhen, et al.
Veröffentlicht: (2025)
von: Yu, Zhen, et al.
Veröffentlicht: (2025)
Lagom: Unleashing the Power of Communication and Computation Overlapping for Distributed LLM Training
von: Xu, Guanbin, et al.
Veröffentlicht: (2026)
von: Xu, Guanbin, et al.
Veröffentlicht: (2026)
MegatronApp: Efficient and Comprehensive Management on Distributed LLM Training
von: Zhao, Bohan, et al.
Veröffentlicht: (2025)
von: Zhao, Bohan, et al.
Veröffentlicht: (2025)
sVIRGO: A Scalable Virtual Tree Hierarchical Framework for Distributed Systems
von: Huang, Lican
Veröffentlicht: (2026)
von: Huang, Lican
Veröffentlicht: (2026)
DeFT: Mitigating Data Dependencies for Flexible Communication Scheduling in Distributed Training
von: Meng, Lin, et al.
Veröffentlicht: (2025)
von: Meng, Lin, et al.
Veröffentlicht: (2025)
Poplar: Efficient Scaling of Distributed DNN Training on Heterogeneous GPU Clusters
von: Zhang, WenZheng, et al.
Veröffentlicht: (2024)
von: Zhang, WenZheng, et al.
Veröffentlicht: (2024)
Efficient Training of Large Language Models on Distributed Infrastructures: A Survey
von: Duan, Jiangfei, et al.
Veröffentlicht: (2024)
von: Duan, Jiangfei, et al.
Veröffentlicht: (2024)
AMSP: Reducing Communication Overhead of ZeRO for Efficient LLM Training
von: Chen, Qiaoling, et al.
Veröffentlicht: (2023)
von: Chen, Qiaoling, et al.
Veröffentlicht: (2023)
Hiding Communication Cost in Distributed LLM Training via Micro-batch Co-execution
von: Wang, Haiquan, et al.
Veröffentlicht: (2024)
von: Wang, Haiquan, et al.
Veröffentlicht: (2024)
Communication-Efficient Distributed Learning via Sparse and Adaptive Stochastic Gradient
von: Deng, Xiaoge, et al.
Veröffentlicht: (2021)
von: Deng, Xiaoge, et al.
Veröffentlicht: (2021)
FedMef: Towards Memory-efficient Federated Dynamic Pruning
von: Huang, Hong, et al.
Veröffentlicht: (2024)
von: Huang, Hong, et al.
Veröffentlicht: (2024)
Efficient Distributed Algorithms for Shape Reduction via Reconfigurable Circuits
von: Almalki, Nada, et al.
Veröffentlicht: (2025)
von: Almalki, Nada, et al.
Veröffentlicht: (2025)
Distributed Hierarchical Machine Learning for Joint Resource Allocation and Slice Selection in In-Network Edge Systems
von: Rashid, Sulaiman Muhammad, et al.
Veröffentlicht: (2025)
von: Rashid, Sulaiman Muhammad, et al.
Veröffentlicht: (2025)
SAFL: Structure-Aware Personalized Federated Learning via Client-Specific Clustering and SCSI-Guided Model Pruning
von: Li, Nan, et al.
Veröffentlicht: (2025)
von: Li, Nan, et al.
Veröffentlicht: (2025)
CO2: Efficient Distributed Training with Full Communication-Computation Overlap
von: Sun, Weigao, et al.
Veröffentlicht: (2024)
von: Sun, Weigao, et al.
Veröffentlicht: (2024)
SparDL: Distributed Deep Learning Training with Efficient Sparse Communication
von: Zhao, Minjun, et al.
Veröffentlicht: (2023)
von: Zhao, Minjun, et al.
Veröffentlicht: (2023)
Bandwidth-Aware and Cost-Efficient Pipeline Parallel Scheduling in Geo-Distributed LLM Training
von: Zhang, Han, et al.
Veröffentlicht: (2026)
von: Zhang, Han, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Federated Learning With L0 Constraint Via Probabilistic Gates For Sparsity
von: Huthasana, Krishna Harsha Kovelakuntla, et al.
Veröffentlicht: (2025) -
Faster Vertex Cover Algorithms on GPUs with Component-Aware Parallel Branching
von: Amro, Hussein, et al.
Veröffentlicht: (2025) -
Pruning Blockchain Protocols for Efficient Access Control in IoT Systems
von: Huang, Yongtao, et al.
Veröffentlicht: (2024) -
Parallelizing Maximal Clique Enumeration on GPUs
von: Almasri, Mohammad, et al.
Veröffentlicht: (2022) -
PacTrain: Pruning and Adaptive Sparse Gradient Compression for Efficient Collective Communication in Distributed Deep Learning
von: Wang, Yisu, et al.
Veröffentlicht: (2025)