KnapsackLB: Enabling Performance-Aware Layer-4 Load Balancing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gandhi, Rohan, Narayana, Srinivas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2021)
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2021)
ReaLB: Real-Time Load Balancing for Multimodal MoE Inference
von: Wang, Yingping, et al.
Veröffentlicht: (2026)
von: Wang, Yingping, et al.
Veröffentlicht: (2026)
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
von: Giannakopoulos, Panagiotis, et al.
Veröffentlicht: (2025)
von: Giannakopoulos, Panagiotis, et al.
Veröffentlicht: (2025)
A Communication- and Memory-Aware Model for Load Balancing Tasks
von: Lifflander, Jonathan, et al.
Veröffentlicht: (2024)
von: Lifflander, Jonathan, et al.
Veröffentlicht: (2024)
miniLB: A Performance Portability Study of Lattice-Boltzmann Simulations
von: Crisci, Luigi, et al.
Veröffentlicht: (2024)
von: Crisci, Luigi, et al.
Veröffentlicht: (2024)
QEdgeProxy: QoS-Aware Load Balancing for IoT Services in the Computing Continuum
von: Čilić, Ivan, et al.
Veröffentlicht: (2024)
von: Čilić, Ivan, et al.
Veröffentlicht: (2024)
DualMap: Enabling Both Cache Affinity and Load Balancing for Distributed LLM Serving
von: Yuan, Ying, et al.
Veröffentlicht: (2026)
von: Yuan, Ying, et al.
Veröffentlicht: (2026)
SkyWalker: A Locality-Aware Cross-Region Load Balancer for LLM Inference
von: Xia, Tian, et al.
Veröffentlicht: (2025)
von: Xia, Tian, et al.
Veröffentlicht: (2025)
CascadeInfer: Length-Aware Scheduling of LLM Serving with Low Latency and Load Balancing
von: Yuan, Yitao, et al.
Veröffentlicht: (2025)
von: Yuan, Yitao, et al.
Veröffentlicht: (2025)
Tetris: Efficient Intra-Datacenter Calls Packing for Large Conferencing Services
von: Gandhi, Rohan, et al.
Veröffentlicht: (2025)
von: Gandhi, Rohan, et al.
Veröffentlicht: (2025)
S-HPLB: Efficient LLM Attention Serving via Sparsity-Aware Head Parallelism Load Balance
von: Liu, Di, et al.
Veröffentlicht: (2026)
von: Liu, Di, et al.
Veröffentlicht: (2026)
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
von: Taylor, Maya, et al.
Veröffentlicht: (2026)
von: Taylor, Maya, et al.
Veröffentlicht: (2026)
Distributed Edge Analytics in Edge-Fog-Cloud Continuum
von: Srirama, Satish Narayana
Veröffentlicht: (2024)
von: Srirama, Satish Narayana
Veröffentlicht: (2024)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
von: Jain, Kunal, et al.
Veröffentlicht: (2024)
von: Jain, Kunal, et al.
Veröffentlicht: (2024)
Distributed Load Balancing with Workload-Dependent Service Rates
von: Zhang, Wenxin, et al.
Veröffentlicht: (2024)
von: Zhang, Wenxin, et al.
Veröffentlicht: (2024)
Review of Hybrid Load Balancing Algorithms in Cloud Computing Environment
von: Ijeoma, Chukwuneke Chiamaka, et al.
Veröffentlicht: (2022)
von: Ijeoma, Chukwuneke Chiamaka, et al.
Veröffentlicht: (2022)
Load Balanced Parallel Node Generation for Meshless Numerical Methods
von: Vehovar, Jon, et al.
Veröffentlicht: (2026)
von: Vehovar, Jon, et al.
Veröffentlicht: (2026)
Fine-grained MoE Load Balancing with Linear Programming
von: Zhao, Chenqi, et al.
Veröffentlicht: (2025)
von: Zhao, Chenqi, et al.
Veröffentlicht: (2025)
Fog enabled distributed training architecture for federated learning
von: Kumar, Aditya, et al.
Veröffentlicht: (2024)
von: Kumar, Aditya, et al.
Veröffentlicht: (2024)
Scalable and Performant Data Loading
von: Hira, Moto, et al.
Veröffentlicht: (2025)
von: Hira, Moto, et al.
Veröffentlicht: (2025)
Slice-Level Scheduling for High Throughput and Load Balanced LLM Serving
von: Cheng, Ke, et al.
Veröffentlicht: (2024)
von: Cheng, Ke, et al.
Veröffentlicht: (2024)
Load Balancing in Strongly Inhomogeneous Simulations -- a Vlasiator Case Study
von: Kotipalo, Leo, et al.
Veröffentlicht: (2025)
von: Kotipalo, Leo, et al.
Veröffentlicht: (2025)
Performance Cost Tradeoffs in Intelligent Load Balancing for Multi Data Center Cloud Systems: From Static Policies to Adaptive Resource Distribution
von: Najafabadi, Saeid Aghasoleymani, et al.
Veröffentlicht: (2025)
von: Najafabadi, Saeid Aghasoleymani, et al.
Veröffentlicht: (2025)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
von: Ramesh, Risshab Srinivas
Veröffentlicht: (2024)
von: Ramesh, Risshab Srinivas
Veröffentlicht: (2024)
TD-Orch: Scalable Load-Balancing for Distributed Systems with Applications to Graph Processing
von: Zhao, Yiwei, et al.
Veröffentlicht: (2025)
von: Zhao, Yiwei, et al.
Veröffentlicht: (2025)
Inference Load-Aware Orchestration for Hierarchical Federated Learning
von: Lackinger, Anna, et al.
Veröffentlicht: (2024)
von: Lackinger, Anna, et al.
Veröffentlicht: (2024)
AcceLLM: Accelerating LLM Inference using Redundancy for Load Balancing and Data Locality
von: Bournias, Ilias, et al.
Veröffentlicht: (2024)
von: Bournias, Ilias, et al.
Veröffentlicht: (2024)
NeutronTP: Load-Balanced Distributed Full-Graph GNN Training with Tensor Parallelism
von: Ai, Xin, et al.
Veröffentlicht: (2024)
von: Ai, Xin, et al.
Veröffentlicht: (2024)
An Analytical Overview Of Virtual Machine Load Balancing Scheduling Algorithms with their Comparative Case Study
von: Vaidya, Priyank, et al.
Veröffentlicht: (2025)
von: Vaidya, Priyank, et al.
Veröffentlicht: (2025)
(Almost) Perfect Discrete Iterative Load Balancing
von: Berenbrink, Petra, et al.
Veröffentlicht: (2025)
von: Berenbrink, Petra, et al.
Veröffentlicht: (2025)
Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML
von: Yoo, Jinsun, et al.
Veröffentlicht: (2026)
von: Yoo, Jinsun, et al.
Veröffentlicht: (2026)
Exploring Dynamic Load Balancing Algorithms for Block-Structured Mesh-and-Particle Simulations in AMReX
von: Nanda, Amitash, et al.
Veröffentlicht: (2025)
von: Nanda, Amitash, et al.
Veröffentlicht: (2025)
FEPLB: Exploiting Copy Engines for Nearly Free MoE Load Balancing in Distributed Training
von: Qi, Shuyao, et al.
Veröffentlicht: (2026)
von: Qi, Shuyao, et al.
Veröffentlicht: (2026)
A Dynamic Approach to Load Balancing in Cloud Infrastructure: Enhancing Energy Efficiency and Resource Utilization
von: Sakib, Shadman, et al.
Veröffentlicht: (2025)
von: Sakib, Shadman, et al.
Veröffentlicht: (2025)
Tackling the Data-Parallel Load Balancing Bottleneck in LLM Serving: Practical Online Routing at Scale
von: Bu, Tianci, et al.
Veröffentlicht: (2026)
von: Bu, Tianci, et al.
Veröffentlicht: (2026)
Sparse Checkpointing for Fast and Reliable MoE Training
von: Gandhi, Swapnil, et al.
Veröffentlicht: (2024)
von: Gandhi, Swapnil, et al.
Veröffentlicht: (2024)
EcoLife: Carbon-Aware Serverless Function Scheduling for Sustainable Computing
von: Jiang, Yankai, et al.
Veröffentlicht: (2024)
von: Jiang, Yankai, et al.
Veröffentlicht: (2024)
A Hierarchical Sharded Blockchain Balancing Performance and Availability
von: Jo, Yongrae, et al.
Veröffentlicht: (2025)
von: Jo, Yongrae, et al.
Veröffentlicht: (2025)
FedCostAware: Enabling Cost-Aware Federated Learning on the Cloud
von: Sinha, Aditya, et al.
Veröffentlicht: (2025)
von: Sinha, Aditya, et al.
Veröffentlicht: (2025)
Clock Distribution with Gradient TRIX
von: Lenzen, Christoph, et al.
Veröffentlicht: (2023)
von: Lenzen, Christoph, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2021) -
ReaLB: Real-Time Load Balancing for Multimodal MoE Inference
von: Wang, Yingping, et al.
Veröffentlicht: (2026) -
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
von: Giannakopoulos, Panagiotis, et al.
Veröffentlicht: (2025) -
A Communication- and Memory-Aware Model for Load Balancing Tasks
von: Lifflander, Jonathan, et al.
Veröffentlicht: (2024) -
miniLB: A Performance Portability Study of Lattice-Boltzmann Simulations
von: Crisci, Luigi, et al.
Veröffentlicht: (2024)