Resource Allocation for Stable LLM Training in Mobile Edge Computing
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Chang, Zhao, Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Resource Allocation in Large Language Model Integrated 6G Vehicular Networks
di: Liu, Chang, et al.
Pubblicazione: (2024)
di: Liu, Chang, et al.
Pubblicazione: (2024)
OptPipe: Memory- and Scheduling-Optimized Pipeline Parallelism for LLM Training
di: Li, Hongpei, et al.
Pubblicazione: (2025)
di: Li, Hongpei, et al.
Pubblicazione: (2025)
Distributed Difference of Convex Optimization
di: Khatana, Vivek, et al.
Pubblicazione: (2024)
di: Khatana, Vivek, et al.
Pubblicazione: (2024)
Uncoded Storage Coded Transmission Elastic Computing with Straggler Tolerance in Heterogeneous Systems
di: Zhong, Xi, et al.
Pubblicazione: (2024)
di: Zhong, Xi, et al.
Pubblicazione: (2024)
Survey of Distributed Algorithms for Resource Allocation over Multi-Agent Systems
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2024)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2024)
Distributed Allocation and Resource Scheduling Algorithms Resilient to Link Failure
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
Multi-Agentic AI for Fairness-Aware and Accelerated Multi-modal Large Model Inference in Real-world Mobile Edge Networks
di: Li, Haiyuan, et al.
Pubblicazione: (2026)
di: Li, Haiyuan, et al.
Pubblicazione: (2026)
Several Performance Bounds on Decentralized Online Optimization are Highly Conservative and Potentially Misleading
di: Meunier, Erwan, et al.
Pubblicazione: (2025)
di: Meunier, Erwan, et al.
Pubblicazione: (2025)
Green-LLM: Optimal Workload Allocation for Environmentally-Aware Distributed Inference
di: Cheng, Jiaming, et al.
Pubblicazione: (2025)
di: Cheng, Jiaming, et al.
Pubblicazione: (2025)
Wireless Distributed Matrix-Vector Multiplication using Over-the-Air Computation and Analog Coding
di: Choi, Jinho
Pubblicazione: (2024)
di: Choi, Jinho
Pubblicazione: (2024)
Impact of Clustering on the Observability and Controllability of Complex Networks
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2026)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2026)
Constraint Programming Models For Serial Batch Scheduling With Minimum Batch Size
di: Huertas, Jorge A., et al.
Pubblicazione: (2025)
di: Huertas, Jorge A., et al.
Pubblicazione: (2025)
Communication Efficient Distributed Training with Distributed Lion
di: Liu, Bo, et al.
Pubblicazione: (2024)
di: Liu, Bo, et al.
Pubblicazione: (2024)
Hierarchical Federated ADMM
di: Azimi-Abarghouyi, Seyed Mohammad, et al.
Pubblicazione: (2024)
di: Azimi-Abarghouyi, Seyed Mohammad, et al.
Pubblicazione: (2024)
A Communication and Computation Efficient Fully First-order Method for Decentralized Bilevel Optimization
di: Wen, Min, et al.
Pubblicazione: (2024)
di: Wen, Min, et al.
Pubblicazione: (2024)
MAST: Model-Agnostic Sparsified Training
di: Demidovich, Yury, et al.
Pubblicazione: (2023)
di: Demidovich, Yury, et al.
Pubblicazione: (2023)
Optimal Update Policy for the Monitoring of Distributed Sources
di: Graves, Eric, et al.
Pubblicazione: (2024)
di: Graves, Eric, et al.
Pubblicazione: (2024)
Accuracy-Delay Trade-Off in LLM Offloading via Token-Level Uncertainty
di: Kim, Yumin, et al.
Pubblicazione: (2026)
di: Kim, Yumin, et al.
Pubblicazione: (2026)
Input-Based Ensemble-Learning Method for Dynamic Memory Configuration of Serverless Computing Functions
di: Agarwal, Siddharth, et al.
Pubblicazione: (2024)
di: Agarwal, Siddharth, et al.
Pubblicazione: (2024)
On-demand Cold Start Frequency Reduction with Off-Policy Reinforcement Learning in Serverless Computing
di: Agarwal, Siddharth, et al.
Pubblicazione: (2023)
di: Agarwal, Siddharth, et al.
Pubblicazione: (2023)
Joint Network-and-Server Congestion in Multi-Source Traffic Allocation: A Convex Formulation and Price-Based Decentralization
di: Sarkar, Tamoghna, et al.
Pubblicazione: (2026)
di: Sarkar, Tamoghna, et al.
Pubblicazione: (2026)
Optimizing LLM Inference: Fluid-Guided Online Scheduling with Memory Constraints
di: Ao, Ruicheng, et al.
Pubblicazione: (2025)
di: Ao, Ruicheng, et al.
Pubblicazione: (2025)
A New Theoretical Perspective on Data Heterogeneity in Federated Optimization
di: Wang, Jiayi, et al.
Pubblicazione: (2024)
di: Wang, Jiayi, et al.
Pubblicazione: (2024)
A Lightweight Method for Tackling Unknown Participation Statistics in Federated Averaging
di: Wang, Shiqiang, et al.
Pubblicazione: (2023)
di: Wang, Shiqiang, et al.
Pubblicazione: (2023)
A Unified Analysis of Federated Learning with Arbitrary Client Participation
di: Wang, Shiqiang, et al.
Pubblicazione: (2022)
di: Wang, Shiqiang, et al.
Pubblicazione: (2022)
Demystifying Why Local Aggregation Helps: Convergence Analysis of Hierarchical SGD
di: Wang, Jiayi, et al.
Pubblicazione: (2020)
di: Wang, Jiayi, et al.
Pubblicazione: (2020)
VREM-FL: Mobility-Aware Computation-Scheduling Co-Design for Vehicular Federated Learning
di: Ballotta, Luca, et al.
Pubblicazione: (2023)
di: Ballotta, Luca, et al.
Pubblicazione: (2023)
ATA: Adaptive Task Allocation for Efficient Resource Management in Distributed Machine Learning
di: Maranjyan, Artavazd, et al.
Pubblicazione: (2025)
di: Maranjyan, Artavazd, et al.
Pubblicazione: (2025)
Momentum-based Distributed Resource Scheduling Optimization Subject to Sector-Bound Nonlinearity and Latency
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
Large-Scale LLM Inference with Heterogeneous Workloads: Prefill-Decode Contention and Asymptotically Optimal Control
di: Lin, Ruihan, et al.
Pubblicazione: (2026)
di: Lin, Ruihan, et al.
Pubblicazione: (2026)
Non-ergodic linear convergence property of the delayed gradient descent under the strongly convexity and the Polyak-Łojasiewicz condition
di: Choi, Hyung Jun, et al.
Pubblicazione: (2023)
di: Choi, Hyung Jun, et al.
Pubblicazione: (2023)
Machine Learning and CPU (Central Processing Unit) Scheduling Co-Optimization over a Network of Computing Centers
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
An Accelerated Distributed Stochastic Gradient Method with Momentum
di: Huang, Kun, et al.
Pubblicazione: (2024)
di: Huang, Kun, et al.
Pubblicazione: (2024)
Distributed Conjugate Gradient Method via Conjugate Direction Tracking
di: Shorinwa, Ola, et al.
Pubblicazione: (2023)
di: Shorinwa, Ola, et al.
Pubblicazione: (2023)
dHPR: A Distributed Halpern Peaceman--Rachford Method for Non-smooth Distributed Optimization Problems
di: Feng, Zhangcheng, et al.
Pubblicazione: (2025)
di: Feng, Zhangcheng, et al.
Pubblicazione: (2025)
Efficient Gradient Methods for Distributed Saddle Problems
di: Luo, Ruichen, et al.
Pubblicazione: (2026)
di: Luo, Ruichen, et al.
Pubblicazione: (2026)
Nonlinear Perturbation-based Non-Convex Optimization over Time-Varying Networks
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2024)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2024)
Distributed Automatic Generation Control subject to Ramp-Rate-Limits: Anytime Feasibility and Uniform Network-Connectivity
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2025)
From Barrier to Bridge: The Case for AI Data Center/Power Grid Co-Design
di: Bashir, Noman, et al.
Pubblicazione: (2026)
di: Bashir, Noman, et al.
Pubblicazione: (2026)
Systemic approach for modeling a generic smart grid
di: Amor, Sofiane Ben, et al.
Pubblicazione: (2025)
di: Amor, Sofiane Ben, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Resource Allocation in Large Language Model Integrated 6G Vehicular Networks
di: Liu, Chang, et al.
Pubblicazione: (2024) -
OptPipe: Memory- and Scheduling-Optimized Pipeline Parallelism for LLM Training
di: Li, Hongpei, et al.
Pubblicazione: (2025) -
Distributed Difference of Convex Optimization
di: Khatana, Vivek, et al.
Pubblicazione: (2024) -
Uncoded Storage Coded Transmission Elastic Computing with Straggler Tolerance in Heterogeneous Systems
di: Zhong, Xi, et al.
Pubblicazione: (2024) -
Survey of Distributed Algorithms for Resource Allocation over Multi-Agent Systems
di: Doostmohammadian, Mohammadreza, et al.
Pubblicazione: (2024)