Efficient Resource Scheduling for Distributed Infrastructures Using Negotiation Capabilities
Fuente:
arXiv
Saved in:
| Main Authors: | Chu, Junjie, Singh, Prashant, Toor, Salman |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Empowering Data Mesh with Federated Learning
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
Online Client Scheduling and Resource Allocation for Efficient Federated Edge Learning
by: Gao, Zhidong, et al.
Published: (2024)
by: Gao, Zhidong, et al.
Published: (2024)
Sustainable Carbon-Aware and Water-Efficient LLM Scheduling in Geo-Distributed Cloud Datacenters
by: Moore, Hayden, et al.
Published: (2025)
by: Moore, Hayden, et al.
Published: (2025)
Application of Machine Learning Optimization in Cloud Computing Resource Scheduling and Management
by: Zhang, Yifan, et al.
Published: (2024)
by: Zhang, Yifan, et al.
Published: (2024)
Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure
by: He, Jun, et al.
Published: (2026)
by: He, Jun, et al.
Published: (2026)
Research on Edge Computing and Cloud Collaborative Resource Scheduling Optimization Based on Deep Reinforcement Learning
by: Wang, Yuqing, et al.
Published: (2025)
by: Wang, Yuqing, et al.
Published: (2025)
ELIS: Efficient LLM Iterative Scheduling System with Response Length Predictor
by: Choi, Seungbeom, et al.
Published: (2025)
by: Choi, Seungbeom, et al.
Published: (2025)
Efficient Fine-Grained GPU Performance Modeling for Distributed Deep Learning of LLM
by: Zhang, Biyao, et al.
Published: (2025)
by: Zhang, Biyao, et al.
Published: (2025)
DPQuant: Efficient and Differentially-Private Model Training via Dynamic Quantization Scheduling
by: Gao, Yubo, et al.
Published: (2025)
by: Gao, Yubo, et al.
Published: (2025)
Justitia: Fair and Efficient Scheduling of Task-parallel LLM Agents with Selective Pampering
by: Yang, Mingyan, et al.
Published: (2025)
by: Yang, Mingyan, et al.
Published: (2025)
Byzantine-Robust Federated Learning Using Generative Adversarial Networks
by: Zafar, Usama, et al.
Published: (2025)
by: Zafar, Usama, et al.
Published: (2025)
InstGenIE: Generative Image Editing Made Efficient with Mask-aware Caching and Scheduling
by: Jiang, Xiaoxiao, et al.
Published: (2025)
by: Jiang, Xiaoxiao, et al.
Published: (2025)
Echo: Efficient Co-Scheduling of Hybrid Online-Offline Tasks for Large Language Model Serving
by: Wang, Zhibin, et al.
Published: (2025)
by: Wang, Zhibin, et al.
Published: (2025)
Semantic Parallelism: Redefining Efficient MoE Inference via Model-Data Co-Scheduling
by: Li, Yan, et al.
Published: (2025)
by: Li, Yan, et al.
Published: (2025)
Efficient Federated Finetuning of Tiny Transformers with Resource-Constrained Devices
by: Pfeiffer, Kilian, et al.
Published: (2024)
by: Pfeiffer, Kilian, et al.
Published: (2024)
Robust LLM Training Infrastructure at ByteDance
by: Wan, Borui, et al.
Published: (2025)
by: Wan, Borui, et al.
Published: (2025)
ZipMoE: Efficient On-Device MoE Serving via Lossless Compression and Cache-Affinity Scheduling
by: Yang, Yuchen, et al.
Published: (2026)
by: Yang, Yuchen, et al.
Published: (2026)
Adaptive Approach to Enhance Machine Learning Scheduling Algorithms During Runtime Using Reinforcement Learning in Metascheduling Applications
by: Alshaer, Samer, et al.
Published: (2025)
by: Alshaer, Samer, et al.
Published: (2025)
MinT: Managed Infrastructure for Training and Serving Millions of LLMs
by: Lab, Mind, et al.
Published: (2026)
by: Lab, Mind, et al.
Published: (2026)
Loss- and Reward-Weighting for Efficient Distributed Reinforcement Learning
by: Holen, Martin, et al.
Published: (2023)
by: Holen, Martin, et al.
Published: (2023)
DeepHYDRA: Resource-Efficient Time-Series Anomaly Detection in Dynamically-Configured Systems
by: Stehle, Franz Kevin, et al.
Published: (2024)
by: Stehle, Franz Kevin, et al.
Published: (2024)
Galvatron: An Automatic Distributed System for Efficient Foundation Model Training
by: Liu, Xinyi, et al.
Published: (2025)
by: Liu, Xinyi, et al.
Published: (2025)
RollArt: Scaling Agentic RL Training via Disaggregated Infrastructure
by: Gao, Wei, et al.
Published: (2025)
by: Gao, Wei, et al.
Published: (2025)
Tackling the Dynamicity in a Production LLM Serving System with SOTA Optimizations via Hybrid Prefill/Decode/Verify Scheduling on Efficient Meta-kernels
by: Song, Mingcong, et al.
Published: (2024)
by: Song, Mingcong, et al.
Published: (2024)
SatFed: A Resource-Efficient LEO Satellite-Assisted Heterogeneous Federated Learning Framework
by: Zhang, Yuxin, et al.
Published: (2024)
by: Zhang, Yuxin, et al.
Published: (2024)
Learning Like Humans: Resource-Efficient Federated Fine-Tuning through Cognitive Developmental Stages
by: Wu, Yebo, et al.
Published: (2025)
by: Wu, Yebo, et al.
Published: (2025)
Interpretable Modeling of Deep Reinforcement Learning Driven Scheduling
by: Li, Boyang, et al.
Published: (2024)
by: Li, Boyang, et al.
Published: (2024)
PGT-I: Scaling Spatiotemporal GNNs with Memory-Efficient Distributed Training
by: Ockerman, Seth, et al.
Published: (2025)
by: Ockerman, Seth, et al.
Published: (2025)
ATTENTION2D: Communication Efficient Distributed Self-Attention Mechanism
by: Elango, Venmugil
Published: (2025)
by: Elango, Venmugil
Published: (2025)
Trillion Parameter AI Serving Infrastructure for Scientific Discovery: A Survey and Vision
by: Hudson, Nathaniel, et al.
Published: (2024)
by: Hudson, Nathaniel, et al.
Published: (2024)
Loop Improvement: An Efficient Approach for Extracting Shared Features from Heterogeneous Data without Central Server
by: Li, Fei, et al.
Published: (2024)
by: Li, Fei, et al.
Published: (2024)
Piper: Efficient Large-Scale MoE Training via Resource Modeling and Pipelined Hybrid Parallelism
by: Dash, Sajal, et al.
Published: (2026)
by: Dash, Sajal, et al.
Published: (2026)
Fairness-Aware Job Scheduling for Multi-Job Federated Learning
by: Shi, Yuxin, et al.
Published: (2024)
by: Shi, Yuxin, et al.
Published: (2024)
AB-Training: A Communication-Efficient Approach for Distributed Low-Rank Learning
by: Coquelin, Daniel, et al.
Published: (2024)
by: Coquelin, Daniel, et al.
Published: (2024)
FedComLoc: Communication-Efficient Distributed Training of Sparse and Quantized Models
by: Yi, Kai, et al.
Published: (2024)
by: Yi, Kai, et al.
Published: (2024)
PubSub-VFL: Towards Efficient Two-Party Split Learning in Heterogeneous Environments via Publisher/Subscriber Architecture
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
TRAIL: Trust-Aware Client Scheduling for Semi-Decentralized Federated Learning
by: Hu, Gangqiang, et al.
Published: (2024)
by: Hu, Gangqiang, et al.
Published: (2024)
SFPrompt: Communication-Efficient Split Federated Fine-Tuning for Large Pre-Trained Models over Resource-Limited Devices
by: Cao, Linxiao, et al.
Published: (2024)
by: Cao, Linxiao, et al.
Published: (2024)
An Advanced Reinforcement Learning Framework for Online Scheduling of Deferrable Workloads in Cloud Computing
by: Dong, Hang, et al.
Published: (2024)
by: Dong, Hang, et al.
Published: (2024)
FedCGD: Collective Gradient Divergence Optimized Scheduling for Wireless Federated Learning
by: Chen, Tan, et al.
Published: (2025)
by: Chen, Tan, et al.
Published: (2025)
Similar Items
-
Empowering Data Mesh with Federated Learning
by: Li, Haoyuan, et al.
Published: (2024) -
Online Client Scheduling and Resource Allocation for Efficient Federated Edge Learning
by: Gao, Zhidong, et al.
Published: (2024) -
Sustainable Carbon-Aware and Water-Efficient LLM Scheduling in Geo-Distributed Cloud Datacenters
by: Moore, Hayden, et al.
Published: (2025) -
Application of Machine Learning Optimization in Cloud Computing Resource Scheduling and Management
by: Zhang, Yifan, et al.
Published: (2024) -
Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure
by: He, Jun, et al.
Published: (2026)