KnapsackLB: Enabling Performance-Aware Layer-4 Load Balancing
Fuente:
arXiv
Salvato in:
| Autori principali: | Gandhi, Rohan, Narayana, Srinivas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
di: Korndörfer, Jonas H. Müller, et al.
Pubblicazione: (2021)
di: Korndörfer, Jonas H. Müller, et al.
Pubblicazione: (2021)
ReaLB: Real-Time Load Balancing for Multimodal MoE Inference
di: Wang, Yingping, et al.
Pubblicazione: (2026)
di: Wang, Yingping, et al.
Pubblicazione: (2026)
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
di: Giannakopoulos, Panagiotis, et al.
Pubblicazione: (2025)
di: Giannakopoulos, Panagiotis, et al.
Pubblicazione: (2025)
A Communication- and Memory-Aware Model for Load Balancing Tasks
di: Lifflander, Jonathan, et al.
Pubblicazione: (2024)
di: Lifflander, Jonathan, et al.
Pubblicazione: (2024)
miniLB: A Performance Portability Study of Lattice-Boltzmann Simulations
di: Crisci, Luigi, et al.
Pubblicazione: (2024)
di: Crisci, Luigi, et al.
Pubblicazione: (2024)
QEdgeProxy: QoS-Aware Load Balancing for IoT Services in the Computing Continuum
di: Čilić, Ivan, et al.
Pubblicazione: (2024)
di: Čilić, Ivan, et al.
Pubblicazione: (2024)
DualMap: Enabling Both Cache Affinity and Load Balancing for Distributed LLM Serving
di: Yuan, Ying, et al.
Pubblicazione: (2026)
di: Yuan, Ying, et al.
Pubblicazione: (2026)
SkyWalker: A Locality-Aware Cross-Region Load Balancer for LLM Inference
di: Xia, Tian, et al.
Pubblicazione: (2025)
di: Xia, Tian, et al.
Pubblicazione: (2025)
CascadeInfer: Length-Aware Scheduling of LLM Serving with Low Latency and Load Balancing
di: Yuan, Yitao, et al.
Pubblicazione: (2025)
di: Yuan, Yitao, et al.
Pubblicazione: (2025)
Tetris: Efficient Intra-Datacenter Calls Packing for Large Conferencing Services
di: Gandhi, Rohan, et al.
Pubblicazione: (2025)
di: Gandhi, Rohan, et al.
Pubblicazione: (2025)
S-HPLB: Efficient LLM Attention Serving via Sparsity-Aware Head Parallelism Load Balance
di: Liu, Di, et al.
Pubblicazione: (2026)
di: Liu, Di, et al.
Pubblicazione: (2026)
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
di: Taylor, Maya, et al.
Pubblicazione: (2026)
di: Taylor, Maya, et al.
Pubblicazione: (2026)
Distributed Edge Analytics in Edge-Fog-Cloud Continuum
di: Srirama, Satish Narayana
Pubblicazione: (2024)
di: Srirama, Satish Narayana
Pubblicazione: (2024)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
di: Jain, Kunal, et al.
Pubblicazione: (2024)
di: Jain, Kunal, et al.
Pubblicazione: (2024)
Distributed Load Balancing with Workload-Dependent Service Rates
di: Zhang, Wenxin, et al.
Pubblicazione: (2024)
di: Zhang, Wenxin, et al.
Pubblicazione: (2024)
Review of Hybrid Load Balancing Algorithms in Cloud Computing Environment
di: Ijeoma, Chukwuneke Chiamaka, et al.
Pubblicazione: (2022)
di: Ijeoma, Chukwuneke Chiamaka, et al.
Pubblicazione: (2022)
Load Balanced Parallel Node Generation for Meshless Numerical Methods
di: Vehovar, Jon, et al.
Pubblicazione: (2026)
di: Vehovar, Jon, et al.
Pubblicazione: (2026)
Fine-grained MoE Load Balancing with Linear Programming
di: Zhao, Chenqi, et al.
Pubblicazione: (2025)
di: Zhao, Chenqi, et al.
Pubblicazione: (2025)
Fog enabled distributed training architecture for federated learning
di: Kumar, Aditya, et al.
Pubblicazione: (2024)
di: Kumar, Aditya, et al.
Pubblicazione: (2024)
Scalable and Performant Data Loading
di: Hira, Moto, et al.
Pubblicazione: (2025)
di: Hira, Moto, et al.
Pubblicazione: (2025)
Slice-Level Scheduling for High Throughput and Load Balanced LLM Serving
di: Cheng, Ke, et al.
Pubblicazione: (2024)
di: Cheng, Ke, et al.
Pubblicazione: (2024)
Load Balancing in Strongly Inhomogeneous Simulations -- a Vlasiator Case Study
di: Kotipalo, Leo, et al.
Pubblicazione: (2025)
di: Kotipalo, Leo, et al.
Pubblicazione: (2025)
Performance Cost Tradeoffs in Intelligent Load Balancing for Multi Data Center Cloud Systems: From Static Policies to Adaptive Resource Distribution
di: Najafabadi, Saeid Aghasoleymani, et al.
Pubblicazione: (2025)
di: Najafabadi, Saeid Aghasoleymani, et al.
Pubblicazione: (2025)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
di: Ramesh, Risshab Srinivas
Pubblicazione: (2024)
di: Ramesh, Risshab Srinivas
Pubblicazione: (2024)
TD-Orch: Scalable Load-Balancing for Distributed Systems with Applications to Graph Processing
di: Zhao, Yiwei, et al.
Pubblicazione: (2025)
di: Zhao, Yiwei, et al.
Pubblicazione: (2025)
Inference Load-Aware Orchestration for Hierarchical Federated Learning
di: Lackinger, Anna, et al.
Pubblicazione: (2024)
di: Lackinger, Anna, et al.
Pubblicazione: (2024)
AcceLLM: Accelerating LLM Inference using Redundancy for Load Balancing and Data Locality
di: Bournias, Ilias, et al.
Pubblicazione: (2024)
di: Bournias, Ilias, et al.
Pubblicazione: (2024)
NeutronTP: Load-Balanced Distributed Full-Graph GNN Training with Tensor Parallelism
di: Ai, Xin, et al.
Pubblicazione: (2024)
di: Ai, Xin, et al.
Pubblicazione: (2024)
An Analytical Overview Of Virtual Machine Load Balancing Scheduling Algorithms with their Comparative Case Study
di: Vaidya, Priyank, et al.
Pubblicazione: (2025)
di: Vaidya, Priyank, et al.
Pubblicazione: (2025)
(Almost) Perfect Discrete Iterative Load Balancing
di: Berenbrink, Petra, et al.
Pubblicazione: (2025)
di: Berenbrink, Petra, et al.
Pubblicazione: (2025)
Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML
di: Yoo, Jinsun, et al.
Pubblicazione: (2026)
di: Yoo, Jinsun, et al.
Pubblicazione: (2026)
Exploring Dynamic Load Balancing Algorithms for Block-Structured Mesh-and-Particle Simulations in AMReX
di: Nanda, Amitash, et al.
Pubblicazione: (2025)
di: Nanda, Amitash, et al.
Pubblicazione: (2025)
FEPLB: Exploiting Copy Engines for Nearly Free MoE Load Balancing in Distributed Training
di: Qi, Shuyao, et al.
Pubblicazione: (2026)
di: Qi, Shuyao, et al.
Pubblicazione: (2026)
A Dynamic Approach to Load Balancing in Cloud Infrastructure: Enhancing Energy Efficiency and Resource Utilization
di: Sakib, Shadman, et al.
Pubblicazione: (2025)
di: Sakib, Shadman, et al.
Pubblicazione: (2025)
Tackling the Data-Parallel Load Balancing Bottleneck in LLM Serving: Practical Online Routing at Scale
di: Bu, Tianci, et al.
Pubblicazione: (2026)
di: Bu, Tianci, et al.
Pubblicazione: (2026)
Sparse Checkpointing for Fast and Reliable MoE Training
di: Gandhi, Swapnil, et al.
Pubblicazione: (2024)
di: Gandhi, Swapnil, et al.
Pubblicazione: (2024)
EcoLife: Carbon-Aware Serverless Function Scheduling for Sustainable Computing
di: Jiang, Yankai, et al.
Pubblicazione: (2024)
di: Jiang, Yankai, et al.
Pubblicazione: (2024)
A Hierarchical Sharded Blockchain Balancing Performance and Availability
di: Jo, Yongrae, et al.
Pubblicazione: (2025)
di: Jo, Yongrae, et al.
Pubblicazione: (2025)
FedCostAware: Enabling Cost-Aware Federated Learning on the Cloud
di: Sinha, Aditya, et al.
Pubblicazione: (2025)
di: Sinha, Aditya, et al.
Pubblicazione: (2025)
Clock Distribution with Gradient TRIX
di: Lenzen, Christoph, et al.
Pubblicazione: (2023)
di: Lenzen, Christoph, et al.
Pubblicazione: (2023)
Documenti analoghi
-
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
di: Korndörfer, Jonas H. Müller, et al.
Pubblicazione: (2021) -
ReaLB: Real-Time Load Balancing for Multimodal MoE Inference
di: Wang, Yingping, et al.
Pubblicazione: (2026) -
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
di: Giannakopoulos, Panagiotis, et al.
Pubblicazione: (2025) -
A Communication- and Memory-Aware Model for Load Balancing Tasks
di: Lifflander, Jonathan, et al.
Pubblicazione: (2024) -
miniLB: A Performance Portability Study of Lattice-Boltzmann Simulations
di: Crisci, Luigi, et al.
Pubblicazione: (2024)