Hierarchical Training of Deep Neural Networks Using Early Exiting
Fuente:
arXiv
Saved in:
| Main Authors: | Sepehri, Yamin, Pad, Pedram, Yüzügüler, Ahmet Caner, Frossard, Pascal, Dunbar, L. Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PriPHiT: Privacy-Preserving Hierarchical Training of Deep Neural Networks
by: Sepehri, Yamin, et al.
Published: (2024)
by: Sepehri, Yamin, et al.
Published: (2024)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
When Does Global Attention Help? A Unified Empirical Study on Atomistic Graph Learning
by: Chowdhury, Arindam, et al.
Published: (2025)
by: Chowdhury, Arindam, et al.
Published: (2025)
ADF-LoRA: Alternating Low-Rank Aggregation for Decentralized Federated Fine-Tuning
by: Wang, Xiaoyu, et al.
Published: (2025)
by: Wang, Xiaoyu, et al.
Published: (2025)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
by: Polyakov, Igor, et al.
Published: (2025)
by: Polyakov, Igor, et al.
Published: (2025)
SPARK: Igniting Communication-Efficient Decentralized Learning via Stage-wise Projected NTK and Accelerated Regularization
by: Xia, Li
Published: (2025)
by: Xia, Li
Published: (2025)
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
by: Lan, Guangchen, et al.
Published: (2024)
by: Lan, Guangchen, et al.
Published: (2024)
DAGER: Exact Gradient Inversion for Large Language Models
by: Petrov, Ivo, et al.
Published: (2024)
by: Petrov, Ivo, et al.
Published: (2024)
XFED: Non-Collusive Model Poisoning Attack Against Byzantine-Robust Federated Classifiers
by: Mouri, Israt Jahan, et al.
Published: (2026)
by: Mouri, Israt Jahan, et al.
Published: (2026)
Training Diffusion Models with Federated Learning
by: de Goede, Matthijs, et al.
Published: (2024)
by: de Goede, Matthijs, et al.
Published: (2024)
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
by: Rosendal, Daan, et al.
Published: (2026)
by: Rosendal, Daan, et al.
Published: (2026)
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
by: Motta, Steven, et al.
Published: (2026)
by: Motta, Steven, et al.
Published: (2026)
Quantize Once, Train Fast: Allreduce-Compatible Compression with Provable Guarantees
by: Xin, Jihao, et al.
Published: (2023)
by: Xin, Jihao, et al.
Published: (2023)
Training LLMs on HPC Systems: Best Practices from the OpenGPT-X Project
by: Penke, Carolin, et al.
Published: (2025)
by: Penke, Carolin, et al.
Published: (2025)
A Comparative Analysis of Distributed Linear Solvers under Data Heterogeneity
by: Velasevic, Boris, et al.
Published: (2023)
by: Velasevic, Boris, et al.
Published: (2023)
Agentic Compilation: Mitigating the LLM Rerun Crisis for Minimized-Inference-Cost Web Automation
by: Chundru, Jagadeesh
Published: (2026)
by: Chundru, Jagadeesh
Published: (2026)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
by: Murimi, Almond Kiruthu
Published: (2025)
by: Murimi, Almond Kiruthu
Published: (2025)
Aergia: Leveraging Heterogeneity in Federated Learning Systems
by: Cox, Bart, et al.
Published: (2022)
by: Cox, Bart, et al.
Published: (2022)
Roadmap for Edge AI: A Dagstuhl Perspective
by: Ding, Aaron Yi, et al.
Published: (2021)
by: Ding, Aaron Yi, et al.
Published: (2021)
Towards Optimal Heterogeneous Client Sampling in Multi-Model Federated Learning
by: Zhang, Haoran, et al.
Published: (2025)
by: Zhang, Haoran, et al.
Published: (2025)
Parameterizing Federated Continual Learning for Reproducible Research
by: Cox, Bart, et al.
Published: (2024)
by: Cox, Bart, et al.
Published: (2024)
Asynchronous Byzantine Federated Learning
by: Cox, Bart, et al.
Published: (2024)
by: Cox, Bart, et al.
Published: (2024)
Hyper-parameter Optimization for Federated Learning with Step-wise Adaptive Mechanism
by: Saadati, Yasaman, et al.
Published: (2024)
by: Saadati, Yasaman, et al.
Published: (2024)
Connecting Large Language Model Agent to High Performance Computing Resource
by: Ma, Heng, et al.
Published: (2025)
by: Ma, Heng, et al.
Published: (2025)
Comparison of Autoscaling Frameworks for Containerised Machine-Learning-Applications in a Local and Cloud Environment
by: Schroeder, Christian, et al.
Published: (2023)
by: Schroeder, Christian, et al.
Published: (2023)
Asynchronous Multi-Server Federated Learning for Geo-Distributed Clients
by: Zuo, Yuncong, et al.
Published: (2024)
by: Zuo, Yuncong, et al.
Published: (2024)
Tram-FL: Routing-based Model Training for Decentralized Federated Learning
by: Maejima, Kota, et al.
Published: (2023)
by: Maejima, Kota, et al.
Published: (2023)
AutoDDL: Automatic Distributed Deep Learning with Near-Optimal Bandwidth Cost
by: Chen, Jinfan, et al.
Published: (2023)
by: Chen, Jinfan, et al.
Published: (2023)
FedFMS: Exploring Federated Foundation Models for Medical Image Segmentation
by: Liu, Yuxi, et al.
Published: (2024)
by: Liu, Yuxi, et al.
Published: (2024)
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
by: Chen, Mu-Chi, et al.
Published: (2025)
by: Chen, Mu-Chi, et al.
Published: (2025)
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
by: Li, Zonghang, et al.
Published: (2025)
by: Li, Zonghang, et al.
Published: (2025)
Federated Learning for Traffic Flow Prediction with Synthetic Data Augmentation
by: Orozco, Fermin, et al.
Published: (2024)
by: Orozco, Fermin, et al.
Published: (2024)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
by: Nguyen, Thinh, et al.
Published: (2025)
by: Nguyen, Thinh, et al.
Published: (2025)
A Full Compression Pipeline for Green Federated Learning in Communication-Constrained Environments
by: Colybes, Elouan, et al.
Published: (2026)
by: Colybes, Elouan, et al.
Published: (2026)
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
by: Xu, Youxuan, et al.
Published: (2025)
by: Xu, Youxuan, et al.
Published: (2025)
SI-ChainFL: Shapley-Incentivized Secure Federated Learning for High-Speed Rail Data Sharing
by: Zhao, Mingjie, et al.
Published: (2026)
by: Zhao, Mingjie, et al.
Published: (2026)
Learning In Chaos: Efficient Autoscaling and Self-Healing for Multi-Party Distributed Training
by: Feng, Wenjiao, et al.
Published: (2025)
by: Feng, Wenjiao, et al.
Published: (2025)
Naeural AI OS -- Decentralized ubiquitous computing MLOps execution engine
by: Bleotiu, Cristian, et al.
Published: (2023)
by: Bleotiu, Cristian, et al.
Published: (2023)
WLB-LLM: Workload-Balanced 4D Parallelism for Large Language Model Training
by: Wang, Zheng, et al.
Published: (2025)
by: Wang, Zheng, et al.
Published: (2025)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026)
by: Tang, Wenjie, et al.
Published: (2026)
Similar Items
-
PriPHiT: Privacy-Preserving Hierarchical Training of Deep Neural Networks
by: Sepehri, Yamin, et al.
Published: (2024) -
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
by: Zhang, Guilin, et al.
Published: (2025) -
When Does Global Attention Help? A Unified Empirical Study on Atomistic Graph Learning
by: Chowdhury, Arindam, et al.
Published: (2025) -
ADF-LoRA: Alternating Low-Rank Aggregation for Decentralized Federated Fine-Tuning
by: Wang, Xiaoyu, et al.
Published: (2025) -
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
by: Polyakov, Igor, et al.
Published: (2025)