Standardized Methods and Recommendations for Green Federated Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Tapp, Austin, Roth, Holger R., Xu, Ziyue, Parida, Abhijeet, Nisar, Hareem, Linguraru, Marius George |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
par: Sun, Bowen, et autres
Publié: (2026)
par: Sun, Bowen, et autres
Publié: (2026)
An Experimental Study of Different Aggregation Schemes in Semi-Asynchronous Federated Learning
par: Li, Yunbo, et autres
Publié: (2024)
par: Li, Yunbo, et autres
Publié: (2024)
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
par: Barron, Ryan, et autres
Publié: (2024)
par: Barron, Ryan, et autres
Publié: (2024)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
par: Lacey, Dane C., et autres
Publié: (2024)
par: Lacey, Dane C., et autres
Publié: (2024)
Rethinking Inference Placement for Deep Learning across Edge and Cloud Platforms: A Multi-Objective Optimization Perspective and Future Directions
par: Zhang, Zongshun, et autres
Publié: (2025)
par: Zhang, Zongshun, et autres
Publié: (2025)
Efficient GPU-Centered Singular Value Decomposition Using the Divide-and-Conquer Method
par: Liu, Shifang, et autres
Publié: (2025)
par: Liu, Shifang, et autres
Publié: (2025)
On the Partitioning of GPU Power among Multi-Instances
par: Vamja, Tirth, et autres
Publié: (2025)
par: Vamja, Tirth, et autres
Publié: (2025)
Architecture Specific Generation of Large Scale Lattice Boltzmann Methods for Sparse Complex Geometries
par: Suffa, Philipp, et autres
Publié: (2024)
par: Suffa, Philipp, et autres
Publié: (2024)
Shifting the Sweet Spot: High-Performance Matrix-Free Method for High-Order Elasticity
par: Chang, Dali, et autres
Publié: (2026)
par: Chang, Dali, et autres
Publié: (2026)
An Auto-tuning Method for Run-time Data Transformation for Sparse Matrix-Vector Multiplication
par: Katagiri, Takahiro, et autres
Publié: (2024)
par: Katagiri, Takahiro, et autres
Publié: (2024)
Optimizing Federated Learning in the Era of LLMs: Message Quantization and Streaming
par: Xu, Ziyue, et autres
Publié: (2025)
par: Xu, Ziyue, et autres
Publié: (2025)
FalconFS: Distributed File System for Large-Scale Deep Learning Pipeline
par: Xu, Jingwei, et autres
Publié: (2025)
par: Xu, Jingwei, et autres
Publié: (2025)
RAPID-LLM: Resilience-Aware Performance analysis of Infrastructure for Distributed LLM Training and Inference
par: Karfakis, George, et autres
Publié: (2025)
par: Karfakis, George, et autres
Publié: (2025)
Learning-Augmented Performance Model for Tensor Product Factorization in High-Order FEM
par: Ren, Xuanzhengbo, et autres
Publié: (2026)
par: Ren, Xuanzhengbo, et autres
Publié: (2026)
Can Large Language Models Predict Parallel Code Performance?
par: Bolet, Gregory, et autres
Publié: (2025)
par: Bolet, Gregory, et autres
Publié: (2025)
oneDAL Optimization for ARM Scalable Vector Extension: Maximizing Efficiency for High-Performance Data Science
par: Sharma, Chandan, et autres
Publié: (2025)
par: Sharma, Chandan, et autres
Publié: (2025)
AGOCS -- Accurate Google Cloud Simulator Framework
par: Sliwko, Leszek, et autres
Publié: (2025)
par: Sliwko, Leszek, et autres
Publié: (2025)
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
par: Zhu, Jianwei, et autres
Publié: (2024)
par: Zhu, Jianwei, et autres
Publié: (2024)
EdgeProfiler: A Fast Profiling Framework for Lightweight LLMs on Edge Using Analytical Model
par: Pinnock, Alyssa, et autres
Publié: (2025)
par: Pinnock, Alyssa, et autres
Publié: (2025)
GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving
par: Jayakody, Shakya, et autres
Publié: (2026)
par: Jayakody, Shakya, et autres
Publié: (2026)
DUAL-BLADE: Dual-Path NVMe-Direct KV-Cache Offloading for Edge LLM Inference
par: Jeong, Bodon, et autres
Publié: (2026)
par: Jeong, Bodon, et autres
Publié: (2026)
CoServe: Efficient Collaboration-of-Experts (CoE) Model Inference with Limited Memory
par: Suo, Jiashun, et autres
Publié: (2025)
par: Suo, Jiashun, et autres
Publié: (2025)
ExpertFlow: Adaptive Expert Scheduling and Memory Coordination for Efficient MoE Inference
par: Shen, Zixu, et autres
Publié: (2025)
par: Shen, Zixu, et autres
Publié: (2025)
Counting Without Running: Evaluating LLMs' Reasoning About Code Complexity
par: Bolet, Gregory, et autres
Publié: (2025)
par: Bolet, Gregory, et autres
Publié: (2025)
When AI Bends Metal: AI-Assisted Optimization of Design Parameters in Sheet Metal Forming
par: Tarraf, Ahmad, et autres
Publié: (2025)
par: Tarraf, Ahmad, et autres
Publié: (2025)
Shared Memory-contention-aware Concurrent DNN Execution for Diversely Heterogeneous System-on-Chips
par: Dagli, Ismet, et autres
Publié: (2023)
par: Dagli, Ismet, et autres
Publié: (2023)
Synthesizing Proxy Applications for MPI Programs
par: Luo, Jiyu, et autres
Publié: (2023)
par: Luo, Jiyu, et autres
Publié: (2023)
DIAL: Decentralized I/O AutoTuning via Learned Client-side Local Metrics for Parallel File System
par: Rashid, Md Hasanur, et autres
Publié: (2026)
par: Rashid, Md Hasanur, et autres
Publié: (2026)
CUTHERMO: Understanding GPU Memory Inefficiencies with Heat Map Profiling
par: Zhao, Yanbo, et autres
Publié: (2025)
par: Zhao, Yanbo, et autres
Publié: (2025)
Efficient Fault Localization in a Cloud Stack Using End-to-End Application Service Topology
par: Mathews, Dhanya R, et autres
Publié: (2025)
par: Mathews, Dhanya R, et autres
Publié: (2025)
A Performance Analysis of BFT Consensus for Blockchains
par: Chan, J. D., et autres
Publié: (2024)
par: Chan, J. D., et autres
Publié: (2024)
CARAT: Client-Side Adaptive RPC and Cache Co-Tuning for Parallel File Systems
par: Rashid, Md Hasanur, et autres
Publié: (2026)
par: Rashid, Md Hasanur, et autres
Publié: (2026)
Comments on "Federated Learning with Differential Privacy: Algorithms and Performance Analysis"
par: Talaei, Mahtab, et autres
Publié: (2024)
par: Talaei, Mahtab, et autres
Publié: (2024)
QoSFlow: Ensuring Service Quality of Distributed Workflows Using Interpretable Sensitivity Models
par: Rashid, Md Hasanur, et autres
Publié: (2026)
par: Rashid, Md Hasanur, et autres
Publié: (2026)
Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles
par: Gallici, Matteo, et autres
Publié: (2025)
par: Gallici, Matteo, et autres
Publié: (2025)
GPU Cluster Scheduling for Network-Sensitive Deep Learning
par: Sharma, Aakash, et autres
Publié: (2024)
par: Sharma, Aakash, et autres
Publié: (2024)
MLCommons Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces
par: Sridharan, Srinivas, et autres
Publié: (2026)
par: Sridharan, Srinivas, et autres
Publié: (2026)
Extracting Practical, Actionable Energy Insights from Supercomputer Telemetry and Logs
par: Cornelius, Melanie, et autres
Publié: (2025)
par: Cornelius, Melanie, et autres
Publié: (2025)
Scaling Large-scale GNN Training to Thousands of Processors on CPU-based Supercomputers
par: Zhuang, Chen, et autres
Publié: (2024)
par: Zhuang, Chen, et autres
Publié: (2024)
Profiling and optimization of multi-card GPU machine learning jobs
par: Lawenda, Marcin, et autres
Publié: (2025)
par: Lawenda, Marcin, et autres
Publié: (2025)
Documents similaires
-
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
par: Sun, Bowen, et autres
Publié: (2026) -
An Experimental Study of Different Aggregation Schemes in Semi-Asynchronous Federated Learning
par: Li, Yunbo, et autres
Publié: (2024) -
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
par: Barron, Ryan, et autres
Publié: (2024) -
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
par: Lacey, Dane C., et autres
Publié: (2024) -
Rethinking Inference Placement for Deep Learning across Edge and Cloud Platforms: A Multi-Objective Optimization Perspective and Future Directions
par: Zhang, Zongshun, et autres
Publié: (2025)