Enregistré dans:
| Auteurs principaux: | Cassagne, Adrien, Amiot, Noé, Bouyer, Manuel |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2508.10481 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Energy-Aware Scheduling Strategies for Partially-Replicable Task Chains on Heterogeneous Processors
par: Idouar, Yacine, et autres
Publié: (2025)
par: Idouar, Yacine, et autres
Publié: (2025)
Augur: Pre-Execution Energy Prediction for Workflow Tasks in Heterogeneous Clusters
par: West, Kathleen, et autres
Publié: (2026)
par: West, Kathleen, et autres
Publié: (2026)
Frenzy: A Memory-Aware Serverless LLM Training System for Heterogeneous GPU Clusters
par: Chang, Zihan, et autres
Publié: (2024)
par: Chang, Zihan, et autres
Publié: (2024)
Energy-Aware Workflow Execution: An Overview of Techniques for Saving Energy and Emissions in Scientific Compute Clusters
par: Thamsen, Lauritz, et autres
Publié: (2025)
par: Thamsen, Lauritz, et autres
Publié: (2025)
Optimal Resource Efficiency with Fairness in Heterogeneous GPU Clusters
par: Mo, Zizhao, et autres
Publié: (2024)
par: Mo, Zizhao, et autres
Publié: (2024)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
par: Liang, Antian, et autres
Publié: (2025)
par: Liang, Antian, et autres
Publié: (2025)
Zorse: Optimizing LLM Training Efficiency on Heterogeneous GPU Clusters
par: Guo, Runsheng Benson, et autres
Publié: (2025)
par: Guo, Runsheng Benson, et autres
Publié: (2025)
Training DNN Models over Heterogeneous Clusters with Optimal Performance
par: Nie, Chengyi, et autres
Publié: (2024)
par: Nie, Chengyi, et autres
Publié: (2024)
Hamava: Fault-tolerant Reconfigurable Geo-Replication on Heterogeneous Clusters
par: Mane, Tejas, et autres
Publié: (2024)
par: Mane, Tejas, et autres
Publié: (2024)
Cephalo: Harnessing Heterogeneous GPU Clusters for Training Transformer Models
par: Guo, Runsheng Benson, et autres
Publié: (2024)
par: Guo, Runsheng Benson, et autres
Publié: (2024)
ClusterLess: Deadline-Aware Serverless Workflow Orchestration on Federated Edge Clusters
par: Farahani, Reza, et autres
Publié: (2026)
par: Farahani, Reza, et autres
Publié: (2026)
Poplar: Efficient Scaling of Distributed DNN Training on Heterogeneous GPU Clusters
par: Zhang, WenZheng, et autres
Publié: (2024)
par: Zhang, WenZheng, et autres
Publié: (2024)
Toward Heterogeneous, Distributed, and Energy-Efficient Computing with SYCL
par: Cosenza, Biagio, et autres
Publié: (2025)
par: Cosenza, Biagio, et autres
Publié: (2025)
Checkpoint and Restart: An Energy Consumption Characterization in Clusters
par: Moran, Marina, et autres
Publié: (2024)
par: Moran, Marina, et autres
Publié: (2024)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
par: Mo, Zizhao, et autres
Publié: (2025)
par: Mo, Zizhao, et autres
Publié: (2025)
Sailor: Automating Distributed Training over Dynamic, Heterogeneous, and Geo-distributed Clusters
par: Strati, Foteini, et autres
Publié: (2025)
par: Strati, Foteini, et autres
Publié: (2025)
HAP: SPMD DNN Training on Heterogeneous GPU Clusters with Automated Program Synthesis
par: Zhang, Shiwei, et autres
Publié: (2024)
par: Zhang, Shiwei, et autres
Publié: (2024)
FATE: Future-State-Aware Scheduling for Heterogeneous LLM Workflows
par: Huang, Zirui, et autres
Publié: (2026)
par: Huang, Zirui, et autres
Publié: (2026)
Bandwidth-Aware LLM Inference on Heterogeneous Many-Core Supercomputers
par: Lu, Yao, et autres
Publié: (2026)
par: Lu, Yao, et autres
Publié: (2026)
Toward Sustainability-Aware LLM Inference on Edge Clusters
par: Rajashekar, Kolichala, et autres
Publié: (2025)
par: Rajashekar, Kolichala, et autres
Publié: (2025)
Hybrid Heterogeneous Clusters Can Lower the Energy Consumption of LLM Inference Workloads
par: Wilkins, Grant, et autres
Publié: (2024)
par: Wilkins, Grant, et autres
Publié: (2024)
Cronus: Efficient LLM inference on Heterogeneous GPU Clusters via Partially Disaggregated Prefill
par: Liu, Yunzhao, et autres
Publié: (2025)
par: Liu, Yunzhao, et autres
Publié: (2025)
Data Heterogeneity-Aware Client Selection for Federated Learning in Wireless Networks
par: Yang, Yanbing, et autres
Publié: (2025)
par: Yang, Yanbing, et autres
Publié: (2025)
Offline Energy-Optimal LLM Serving: Workload-Based Energy Models for LLM Inference on Heterogeneous Systems
par: Wilkins, Grant, et autres
Publié: (2024)
par: Wilkins, Grant, et autres
Publié: (2024)
Squeezing Edge Performance: A Sensitivity-Aware Container Management for Heterogeneous Tasks
par: Zhang, Yongmin, et autres
Publié: (2025)
par: Zhang, Yongmin, et autres
Publié: (2025)
D-Rex: Heterogeneity-Aware Reliability Framework and Adaptive Algorithms for Distributed Storage
par: Gonthier, Maxime, et autres
Publié: (2025)
par: Gonthier, Maxime, et autres
Publié: (2025)
Heterogeneity-Aware Memory Efficient Federated Learning via Progressive Layer Freezing
par: Yebo, Wu, et autres
Publié: (2024)
par: Yebo, Wu, et autres
Publié: (2024)
EcoShift: Performance-Aware Power Management for Power-Constrained Heterogeneous Systems
par: Zheng, Zhong, et autres
Publié: (2026)
par: Zheng, Zhong, et autres
Publié: (2026)
QONNECT: A QoS-Aware Orchestration System for Distributed Kubernetes Clusters
par: Aslan, Haci Ismail, et autres
Publié: (2025)
par: Aslan, Haci Ismail, et autres
Publié: (2025)
FREESH: Fair, Resource- and Energy-Efficient Scheduling for LLM Serving on Heterogeneous GPUs
par: He, Xuan, et autres
Publié: (2025)
par: He, Xuan, et autres
Publié: (2025)
Federated Learning within Global Energy Budget over Heterogeneous Edge Accelerators
par: Banerjee, Roopkatha, et autres
Publié: (2025)
par: Banerjee, Roopkatha, et autres
Publié: (2025)
PPipe: Efficient Video Analytics Serving on Heterogeneous GPU Clusters via Pool-Based Pipeline Parallelism
par: Kong, Z. Jonny, et autres
Publié: (2025)
par: Kong, Z. Jonny, et autres
Publié: (2025)
The Case for Time-Shared Computing Resources
par: Jacquet, Pierre, et autres
Publié: (2025)
par: Jacquet, Pierre, et autres
Publié: (2025)
HexAGenT: Efficient Agentic LLM Serving via Workflow- and Heterogeneity-Aware Scheduling
par: Peng, You, et autres
Publié: (2026)
par: Peng, You, et autres
Publié: (2026)
BandPilot: Towards Performance- and Contention-Aware GPU Dispatching in AI Clusters
par: Zhang, Kunming, et autres
Publié: (2025)
par: Zhang, Kunming, et autres
Publié: (2025)
PAL: A Variability-Aware Policy for Scheduling ML Workloads in GPU Clusters
par: Jain, Rutwik, et autres
Publié: (2024)
par: Jain, Rutwik, et autres
Publié: (2024)
Calibrating Microgrid Simulations for Energy-Aware Computing Systems
par: Steinke, Marvin
Publié: (2026)
par: Steinke, Marvin
Publié: (2026)
Scaling Up Throughput-oriented LLM Inference Applications on Heterogeneous Opportunistic GPU Clusters with Pervasive Context Management
par: Phung, Thanh Son, et autres
Publié: (2025)
par: Phung, Thanh Son, et autres
Publié: (2025)
H2:Towards Efficient Large-Scale LLM Training on Hyper-Heterogeneous Cluster over 1,000 Chips
par: Tang, Ding, et autres
Publié: (2025)
par: Tang, Ding, et autres
Publié: (2025)
Efficiently Executing High-throughput Lightweight LLM Inference Applications on Heterogeneous Opportunistic GPU Clusters with Pervasive Context Management
par: Phung, Thanh Son, et autres
Publié: (2025)
par: Phung, Thanh Son, et autres
Publié: (2025)
Documents similaires
-
Energy-Aware Scheduling Strategies for Partially-Replicable Task Chains on Heterogeneous Processors
par: Idouar, Yacine, et autres
Publié: (2025) -
Augur: Pre-Execution Energy Prediction for Workflow Tasks in Heterogeneous Clusters
par: West, Kathleen, et autres
Publié: (2026) -
Frenzy: A Memory-Aware Serverless LLM Training System for Heterogeneous GPU Clusters
par: Chang, Zihan, et autres
Publié: (2024) -
Energy-Aware Workflow Execution: An Overview of Techniques for Saving Energy and Emissions in Scientific Compute Clusters
par: Thamsen, Lauritz, et autres
Publié: (2025) -
Optimal Resource Efficiency with Fairness in Heterogeneous GPU Clusters
par: Mo, Zizhao, et autres
Publié: (2024)