A Communication- and Memory-Aware Model for Load Balancing Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Lifflander, Jonathan, Pebay, Philippe P., Slattengren, Nicole L., Pebay, Pierre L., Pfeiffer, Robert A., Kotulski, Joseph D., McGovern, Sean T. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
di: Taylor, Maya, et al.
Pubblicazione: (2026)
di: Taylor, Maya, et al.
Pubblicazione: (2026)
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
di: Giannakopoulos, Panagiotis, et al.
Pubblicazione: (2025)
di: Giannakopoulos, Panagiotis, et al.
Pubblicazione: (2025)
KnapsackLB: Enabling Performance-Aware Layer-4 Load Balancing
di: Gandhi, Rohan, et al.
Pubblicazione: (2024)
di: Gandhi, Rohan, et al.
Pubblicazione: (2024)
QEdgeProxy: QoS-Aware Load Balancing for IoT Services in the Computing Continuum
di: Čilić, Ivan, et al.
Pubblicazione: (2024)
di: Čilić, Ivan, et al.
Pubblicazione: (2024)
Distributed Load Balancing with Workload-Dependent Service Rates
di: Zhang, Wenxin, et al.
Pubblicazione: (2024)
di: Zhang, Wenxin, et al.
Pubblicazione: (2024)
SkyWalker: A Locality-Aware Cross-Region Load Balancer for LLM Inference
di: Xia, Tian, et al.
Pubblicazione: (2025)
di: Xia, Tian, et al.
Pubblicazione: (2025)
CascadeInfer: Length-Aware Scheduling of LLM Serving with Low Latency and Load Balancing
di: Yuan, Yitao, et al.
Pubblicazione: (2025)
di: Yuan, Yitao, et al.
Pubblicazione: (2025)
Sizey: Memory-Efficient Execution of Scientific Workflow Tasks
di: Bader, Jonathan, et al.
Pubblicazione: (2024)
di: Bader, Jonathan, et al.
Pubblicazione: (2024)
S-HPLB: Efficient LLM Attention Serving via Sparsity-Aware Head Parallelism Load Balance
di: Liu, Di, et al.
Pubblicazione: (2026)
di: Liu, Di, et al.
Pubblicazione: (2026)
Load Balancing Using Sparse Communication
di: Mendelson, Gal, et al.
Pubblicazione: (2022)
di: Mendelson, Gal, et al.
Pubblicazione: (2022)
Predicting Dynamic Memory Requirements for Scientific Workflow Tasks
di: Bader, Jonathan, et al.
Pubblicazione: (2023)
di: Bader, Jonathan, et al.
Pubblicazione: (2023)
KS+: Predicting Workflow Task Memory Usage Over Time
di: Bader, Jonathan, et al.
Pubblicazione: (2024)
di: Bader, Jonathan, et al.
Pubblicazione: (2024)
(Almost) Perfect Discrete Iterative Load Balancing
di: Berenbrink, Petra, et al.
Pubblicazione: (2025)
di: Berenbrink, Petra, et al.
Pubblicazione: (2025)
TD-Orch: Scalable Load-Balancing for Distributed Systems with Applications to Graph Processing
di: Zhao, Yiwei, et al.
Pubblicazione: (2025)
di: Zhao, Yiwei, et al.
Pubblicazione: (2025)
Review of Hybrid Load Balancing Algorithms in Cloud Computing Environment
di: Ijeoma, Chukwuneke Chiamaka, et al.
Pubblicazione: (2022)
di: Ijeoma, Chukwuneke Chiamaka, et al.
Pubblicazione: (2022)
Load Balanced Parallel Node Generation for Meshless Numerical Methods
di: Vehovar, Jon, et al.
Pubblicazione: (2026)
di: Vehovar, Jon, et al.
Pubblicazione: (2026)
Fine-grained MoE Load Balancing with Linear Programming
di: Zhao, Chenqi, et al.
Pubblicazione: (2025)
di: Zhao, Chenqi, et al.
Pubblicazione: (2025)
Ponder: Online Prediction of Task Memory Requirements for Scientific Workflows
di: Lehmann, Fabian, et al.
Pubblicazione: (2024)
di: Lehmann, Fabian, et al.
Pubblicazione: (2024)
Shared Memory-Aware Latency-Sensitive Message Aggregation for Fine-Grained Communication
di: Chandrasekar, Kavitha, et al.
Pubblicazione: (2024)
di: Chandrasekar, Kavitha, et al.
Pubblicazione: (2024)
Load Balancing in Strongly Inhomogeneous Simulations -- a Vlasiator Case Study
di: Kotipalo, Leo, et al.
Pubblicazione: (2025)
di: Kotipalo, Leo, et al.
Pubblicazione: (2025)
Slice-Level Scheduling for High Throughput and Load Balanced LLM Serving
di: Cheng, Ke, et al.
Pubblicazione: (2024)
di: Cheng, Ke, et al.
Pubblicazione: (2024)
ReaLB: Real-Time Load Balancing for Multimodal MoE Inference
di: Wang, Yingping, et al.
Pubblicazione: (2026)
di: Wang, Yingping, et al.
Pubblicazione: (2026)
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
di: Korndörfer, Jonas H. Müller, et al.
Pubblicazione: (2021)
di: Korndörfer, Jonas H. Müller, et al.
Pubblicazione: (2021)
Inference Load-Aware Orchestration for Hierarchical Federated Learning
di: Lackinger, Anna, et al.
Pubblicazione: (2024)
di: Lackinger, Anna, et al.
Pubblicazione: (2024)
DualMap: Enabling Both Cache Affinity and Load Balancing for Distributed LLM Serving
di: Yuan, Ying, et al.
Pubblicazione: (2026)
di: Yuan, Ying, et al.
Pubblicazione: (2026)
AcceLLM: Accelerating LLM Inference using Redundancy for Load Balancing and Data Locality
di: Bournias, Ilias, et al.
Pubblicazione: (2024)
di: Bournias, Ilias, et al.
Pubblicazione: (2024)
NeutronTP: Load-Balanced Distributed Full-Graph GNN Training with Tensor Parallelism
di: Ai, Xin, et al.
Pubblicazione: (2024)
di: Ai, Xin, et al.
Pubblicazione: (2024)
An Analytical Overview Of Virtual Machine Load Balancing Scheduling Algorithms with their Comparative Case Study
di: Vaidya, Priyank, et al.
Pubblicazione: (2025)
di: Vaidya, Priyank, et al.
Pubblicazione: (2025)
Exploring Dynamic Load Balancing Algorithms for Block-Structured Mesh-and-Particle Simulations in AMReX
di: Nanda, Amitash, et al.
Pubblicazione: (2025)
di: Nanda, Amitash, et al.
Pubblicazione: (2025)
FEPLB: Exploiting Copy Engines for Nearly Free MoE Load Balancing in Distributed Training
di: Qi, Shuyao, et al.
Pubblicazione: (2026)
di: Qi, Shuyao, et al.
Pubblicazione: (2026)
A Dynamic Approach to Load Balancing in Cloud Infrastructure: Enhancing Energy Efficiency and Resource Utilization
di: Sakib, Shadman, et al.
Pubblicazione: (2025)
di: Sakib, Shadman, et al.
Pubblicazione: (2025)
Tackling the Data-Parallel Load Balancing Bottleneck in LLM Serving: Practical Online Routing at Scale
di: Bu, Tianci, et al.
Pubblicazione: (2026)
di: Bu, Tianci, et al.
Pubblicazione: (2026)
WOW: Workflow-Aware Data Movement and Task Scheduling for Dynamic Scientific Workflows
di: Lehmann, Fabian, et al.
Pubblicazione: (2025)
di: Lehmann, Fabian, et al.
Pubblicazione: (2025)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
di: Jain, Kunal, et al.
Pubblicazione: (2024)
di: Jain, Kunal, et al.
Pubblicazione: (2024)
LOCO: Rethinking Objects for Network Memory
di: Hodgkins, George, et al.
Pubblicazione: (2025)
di: Hodgkins, George, et al.
Pubblicazione: (2025)
Pro-Prophet: A Systematic Load Balancing Method for Efficient Parallel Training of Large-scale MoE Models
di: Wang, Wei, et al.
Pubblicazione: (2024)
di: Wang, Wei, et al.
Pubblicazione: (2024)
Coherence-Aware Task Graph Modeling for Realistic Application
di: Xiong, Guochu, et al.
Pubblicazione: (2025)
di: Xiong, Guochu, et al.
Pubblicazione: (2025)
Scalable and Performant Data Loading
di: Hira, Moto, et al.
Pubblicazione: (2025)
di: Hira, Moto, et al.
Pubblicazione: (2025)
An Asynchronous Many-Task Algorithm for Unstructured $S_{N}$ Transport on Shared Memory Systems
di: Elwood, Alex, et al.
Pubblicazione: (2025)
di: Elwood, Alex, et al.
Pubblicazione: (2025)
SLO-Aware Task Offloading within Collaborative Vehicle Platoons
di: Sedlak, Boris, et al.
Pubblicazione: (2024)
di: Sedlak, Boris, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Communication-Aware Diffusion Load Balancing for Persistently Interacting Objects
di: Taylor, Maya, et al.
Pubblicazione: (2026) -
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
di: Giannakopoulos, Panagiotis, et al.
Pubblicazione: (2025) -
KnapsackLB: Enabling Performance-Aware Layer-4 Load Balancing
di: Gandhi, Rohan, et al.
Pubblicazione: (2024) -
QEdgeProxy: QoS-Aware Load Balancing for IoT Services in the Computing Continuum
di: Čilić, Ivan, et al.
Pubblicazione: (2024) -
Distributed Load Balancing with Workload-Dependent Service Rates
di: Zhang, Wenxin, et al.
Pubblicazione: (2024)