AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Guilin, Guo, Wulan, Tan, Ziqi, Jiang, Hailong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
CarbonEdge: Carbon-Aware Deep Learning Inference Framework for Sustainable Edge Computing
von: Zhang, Guilin, et al.
Veröffentlicht: (2026)
von: Zhang, Guilin, et al.
Veröffentlicht: (2026)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
von: Polyakov, Igor, et al.
Veröffentlicht: (2025)
von: Polyakov, Igor, et al.
Veröffentlicht: (2025)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
von: Motta, Steven, et al.
Veröffentlicht: (2026)
von: Motta, Steven, et al.
Veröffentlicht: (2026)
ADF-LoRA: Alternating Low-Rank Aggregation for Decentralized Federated Fine-Tuning
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
von: Rosendal, Daan, et al.
Veröffentlicht: (2026)
von: Rosendal, Daan, et al.
Veröffentlicht: (2026)
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
SPARK: Igniting Communication-Efficient Decentralized Learning via Stage-wise Projected NTK and Accelerated Regularization
von: Xia, Li
Veröffentlicht: (2025)
von: Xia, Li
Veröffentlicht: (2025)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
Comparison of Autoscaling Frameworks for Containerised Machine-Learning-Applications in a Local and Cloud Environment
von: Schroeder, Christian, et al.
Veröffentlicht: (2023)
von: Schroeder, Christian, et al.
Veröffentlicht: (2023)
Roadmap for Edge AI: A Dagstuhl Perspective
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
von: Lan, Guangchen, et al.
Veröffentlicht: (2024)
von: Lan, Guangchen, et al.
Veröffentlicht: (2024)
Mobile Traffic Prediction at the Edge Through Distributed and Deep Transfer Learning
von: Petrella, Alfredo, et al.
Veröffentlicht: (2023)
von: Petrella, Alfredo, et al.
Veröffentlicht: (2023)
Hyper-parameter Optimization for Federated Learning with Step-wise Adaptive Mechanism
von: Saadati, Yasaman, et al.
Veröffentlicht: (2024)
von: Saadati, Yasaman, et al.
Veröffentlicht: (2024)
Connecting Large Language Model Agent to High Performance Computing Resource
von: Ma, Heng, et al.
Veröffentlicht: (2025)
von: Ma, Heng, et al.
Veröffentlicht: (2025)
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023)
Agentic Compilation: Mitigating the LLM Rerun Crisis for Minimized-Inference-Cost Web Automation
von: Chundru, Jagadeesh
Veröffentlicht: (2026)
von: Chundru, Jagadeesh
Veröffentlicht: (2026)
XFED: Non-Collusive Model Poisoning Attack Against Byzantine-Robust Federated Classifiers
von: Mouri, Israt Jahan, et al.
Veröffentlicht: (2026)
von: Mouri, Israt Jahan, et al.
Veröffentlicht: (2026)
DAGER: Exact Gradient Inversion for Large Language Models
von: Petrov, Ivo, et al.
Veröffentlicht: (2024)
von: Petrov, Ivo, et al.
Veröffentlicht: (2024)
A Full Compression Pipeline for Green Federated Learning in Communication-Constrained Environments
von: Colybes, Elouan, et al.
Veröffentlicht: (2026)
von: Colybes, Elouan, et al.
Veröffentlicht: (2026)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
Federated Learning Model Aggregation in Heterogenous Aerial and Space Networks
von: Dong, Fan, et al.
Veröffentlicht: (2023)
von: Dong, Fan, et al.
Veröffentlicht: (2023)
EWSJF: An Adaptive Scheduler with Hybrid Partitioning for Mixed-Workload LLM Inference
von: Sidik, Bronislav, et al.
Veröffentlicht: (2026)
von: Sidik, Bronislav, et al.
Veröffentlicht: (2026)
Aergia: Leveraging Heterogeneity in Federated Learning Systems
von: Cox, Bart, et al.
Veröffentlicht: (2022)
von: Cox, Bart, et al.
Veröffentlicht: (2022)
Towards Optimal Heterogeneous Client Sampling in Multi-Model Federated Learning
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
Parameterizing Federated Continual Learning for Reproducible Research
von: Cox, Bart, et al.
Veröffentlicht: (2024)
von: Cox, Bart, et al.
Veröffentlicht: (2024)
Asynchronous Byzantine Federated Learning
von: Cox, Bart, et al.
Veröffentlicht: (2024)
von: Cox, Bart, et al.
Veröffentlicht: (2024)
Training Diffusion Models with Federated Learning
von: de Goede, Matthijs, et al.
Veröffentlicht: (2024)
von: de Goede, Matthijs, et al.
Veröffentlicht: (2024)
Quantize Once, Train Fast: Allreduce-Compatible Compression with Provable Guarantees
von: Xin, Jihao, et al.
Veröffentlicht: (2023)
von: Xin, Jihao, et al.
Veröffentlicht: (2023)
Asynchronous Multi-Server Federated Learning for Geo-Distributed Clients
von: Zuo, Yuncong, et al.
Veröffentlicht: (2024)
von: Zuo, Yuncong, et al.
Veröffentlicht: (2024)
CodeCRDT: Observation-Driven Coordination for Multi-Agent LLM Code Generation
von: Pugachev, Sergey
Veröffentlicht: (2025)
von: Pugachev, Sergey
Veröffentlicht: (2025)
Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol
von: Parmar, Abhinav Singh
Veröffentlicht: (2026)
von: Parmar, Abhinav Singh
Veröffentlicht: (2026)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026)
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026)
A Comparative Analysis of Distributed Linear Solvers under Data Heterogeneity
von: Velasevic, Boris, et al.
Veröffentlicht: (2023)
von: Velasevic, Boris, et al.
Veröffentlicht: (2023)
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
von: Li, Zonghang, et al.
Veröffentlicht: (2025)
von: Li, Zonghang, et al.
Veröffentlicht: (2025)
TPI-LLM: Serving 70B-scale LLMs Efficiently on Low-resource Edge Devices
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
AutoDDL: Automatic Distributed Deep Learning with Near-Optimal Bandwidth Cost
von: Chen, Jinfan, et al.
Veröffentlicht: (2023)
von: Chen, Jinfan, et al.
Veröffentlicht: (2023)
Combinatorial Client-Master Multiagent Deep Reinforcement Learning for Task Offloading in Mobile Edge Computing
von: Gebrekidan, Tesfay Zemuy, et al.
Veröffentlicht: (2024)
von: Gebrekidan, Tesfay Zemuy, et al.
Veröffentlicht: (2024)
When Does Global Attention Help? A Unified Empirical Study on Atomistic Graph Learning
von: Chowdhury, Arindam, et al.
Veröffentlicht: (2025)
von: Chowdhury, Arindam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025) -
CarbonEdge: Carbon-Aware Deep Learning Inference Framework for Sustainable Edge Computing
von: Zhang, Guilin, et al.
Veröffentlicht: (2026) -
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
von: Polyakov, Igor, et al.
Veröffentlicht: (2025) -
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025) -
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
von: Motta, Steven, et al.
Veröffentlicht: (2026)