SCARIF: Towards Carbon Modeling of Cloud Servers with Accelerators
Fuente:
arXiv
Saved in:
| Main Authors: | Ji, Shixin, Yang, Zhuoping, Chen, Xingzhen, Cahoon, Stephen, Hu, Jingtong, Shi, Yiyu, Jones, Alex K., Zhou, Peipei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Advancing Environmental Sustainability in Data Centers via Carbon Depreciation Models
by: Ji, Shixin, et al.
Published: (2024)
by: Ji, Shixin, et al.
Published: (2024)
AGILE: Lightweight and Efficient Asynchronous GPU-SSD Integration
by: Yang, Zhuoping, et al.
Published: (2025)
by: Yang, Zhuoping, et al.
Published: (2025)
BSODiag: A Global Diagnosis Framework for Batch Servers Outage in Large-scale Cloud Infrastructure Systems
by: Duan, Tao, et al.
Published: (2025)
by: Duan, Tao, et al.
Published: (2025)
More is Different: Prototyping and Analyzing a New Form of Edge Server with Massive Mobile SoCs
by: Zhang, Li, et al.
Published: (2022)
by: Zhang, Li, et al.
Published: (2022)
EC2MoE: Adaptive End-Cloud Pipeline Collaboration Enabling Scalable Mixture-of-Experts Inference
by: Yang, Zheming, et al.
Published: (2025)
by: Yang, Zheming, et al.
Published: (2025)
XaaS: Acceleration as a Service to Enable Productive High-Performance Cloud Computing
by: Hoefler, Torsten, et al.
Published: (2024)
by: Hoefler, Torsten, et al.
Published: (2024)
Are Bus-Mounted Edge Servers Feasible?
by: Li, Xuezhi, et al.
Published: (2025)
by: Li, Xuezhi, et al.
Published: (2025)
WaterWise: Co-optimizing Carbon- and Water-Footprint Toward Environmentally Sustainable Cloud Computing
by: Jiang, Yankai, et al.
Published: (2025)
by: Jiang, Yankai, et al.
Published: (2025)
A Multi-Server Information-Sharing Environment for Cross-Party Collaboration on A Private Cloud
by: Zhang, Jianping, et al.
Published: (2024)
by: Zhang, Jianping, et al.
Published: (2024)
Practical Federated Learning without a Server
by: Dhasade, Akash, et al.
Published: (2025)
by: Dhasade, Akash, et al.
Published: (2025)
CarbonFlex: Enabling Carbon-aware Provisioning and Scheduling for Cloud Clusters
by: Hanafy, Walid A., et al.
Published: (2025)
by: Hanafy, Walid A., et al.
Published: (2025)
Experimental Analysis of Server-Side Caching for Web Performance
by: Umar, Mohammad, et al.
Published: (2026)
by: Umar, Mohammad, et al.
Published: (2026)
Analysis of Server Throughput For Managed Big Data Analytics Frameworks
by: Anagnostakis, Emmanouil, et al.
Published: (2025)
by: Anagnostakis, Emmanouil, et al.
Published: (2025)
MoA-Off: Adaptive Heterogeneous Modality-Aware Offloading with Edge-Cloud Collaboration for Efficient Multimodal LLM Inference
by: Yang, Zheming, et al.
Published: (2025)
by: Yang, Zheming, et al.
Published: (2025)
Accelerating Cloud-Based Transcriptomics: Performance Analysis and Optimization of the STAR Aligner Workflow
by: Kica, Piotr, et al.
Published: (2025)
by: Kica, Piotr, et al.
Published: (2025)
PipeMax: Enhancing Offline LLM Inference on Commodity GPU Servers
by: Zhang, Hongbin, et al.
Published: (2026)
by: Zhang, Hongbin, et al.
Published: (2026)
Experiences with Model Context Protocol Servers for Science and High Performance Computing
by: Pan, Haochen, et al.
Published: (2025)
by: Pan, Haochen, et al.
Published: (2025)
Towards an Adaptive Runtime System for Cloud-Native HPC
by: Bhosale, Aditya, et al.
Published: (2026)
by: Bhosale, Aditya, et al.
Published: (2026)
Towards Cloud Efficiency with Large-scale Workload Characterization
by: Parayil, Anjaly, et al.
Published: (2024)
by: Parayil, Anjaly, et al.
Published: (2024)
Accelerating End-Cloud Collaborative Inference via Near Bubble-free Pipeline Optimization
by: Gao, Luyao, et al.
Published: (2024)
by: Gao, Luyao, et al.
Published: (2024)
Towards Seamless Serverless Computing Across an Edge-Cloud Continuum
by: Simion, Emilian, et al.
Published: (2024)
by: Simion, Emilian, et al.
Published: (2024)
City-Scale Visibility Graph Analysis via GPU-Accelerated HyperBall
by: Hodge, Alex, et al.
Published: (2026)
by: Hodge, Alex, et al.
Published: (2026)
TailBench++: Flexible Multi-Client, Multi-Server Benchmarking for Latency-Critical Workloads
by: Li, Zhilin, et al.
Published: (2025)
by: Li, Zhilin, et al.
Published: (2025)
From Servers to Sites: Compositional Power Trace Generation of LLM Inference for Infrastructure Planning
by: Wilkins, Grant, et al.
Published: (2026)
by: Wilkins, Grant, et al.
Published: (2026)
Failure-Resilient and Carbon-Efficient Deployment of Microservices over the Cloud-Edge Continuum
by: Ponce, Francisco, et al.
Published: (2026)
by: Ponce, Francisco, et al.
Published: (2026)
Radar DataTree: A FAIR and Cloud-Native Framework for Scalable Weather Radar Archives
by: Ladino-Rincon, Alfonso, et al.
Published: (2025)
by: Ladino-Rincon, Alfonso, et al.
Published: (2025)
Huawei Cloud Model-as-a-Service on the CloudMatrix384 SuperPod
by: Xiao, Ao, et al.
Published: (2025)
by: Xiao, Ao, et al.
Published: (2025)
Spatio-Temporal Shifting to Reduce Carbon, Water, and Land-Use Footprints of Cloud Workloads
by: Attenni, Giulio, et al.
Published: (2025)
by: Attenni, Giulio, et al.
Published: (2025)
Aging-aware CPU Core Management for Embodied Carbon Amortization in Cloud LLM Inference
by: Hewage, Tharindu B., et al.
Published: (2025)
by: Hewage, Tharindu B., et al.
Published: (2025)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
by: Qiao, Tong, et al.
Published: (2025)
by: Qiao, Tong, et al.
Published: (2025)
Oases: Efficient Large-Scale Model Training on Commodity Servers via Overlapped and Automated Tensor Model Parallelism
by: Li, Shengwei, et al.
Published: (2023)
by: Li, Shengwei, et al.
Published: (2023)
A Framework for Carbon-aware Real-Time Workload Management in Clouds using Renewables-driven Cores
by: Hewage, Tharindu B., et al.
Published: (2024)
by: Hewage, Tharindu B., et al.
Published: (2024)
PALM: A Efficient Performance Simulator for Tiled Accelerators with Large-scale Model Training
by: Fang, Jiahao, et al.
Published: (2024)
by: Fang, Jiahao, et al.
Published: (2024)
Towards Designing an Energy Aware Data Replication Strategy for Cloud Systems Using Reinforcement Learning
by: Najjar, Amir, et al.
Published: (2025)
by: Najjar, Amir, et al.
Published: (2025)
MSAO: Adaptive Modality Sparsity-Aware Offloading with Edge-Cloud Collaboration for Efficient Multimodal LLM Inference
by: Yang, Zheming, et al.
Published: (2026)
by: Yang, Zheming, et al.
Published: (2026)
Towards singular optimality in the presence of local initial knowledge
by: Ji, Hongyan, et al.
Published: (2024)
by: Ji, Hongyan, et al.
Published: (2024)
FASTEN: Towards a FAult-tolerant and STorage EfficieNt Cloud: Balancing Between Replication and Deduplication
by: Ahmed, Sabbir, et al.
Published: (2023)
by: Ahmed, Sabbir, et al.
Published: (2023)
Performance Characterization of Distributed Deep Learning Strategies: A Quantitative Evaluation of DDP, FSDP, and Parameter Server Architectures on GPU Clusters
by: Ovi, Md Sultanul Islam
Published: (2025)
by: Ovi, Md Sultanul Islam
Published: (2025)
Kairos: Low-latency Multi-Agent Serving with Shared LLMs and Excessive Loads in the Public Cloud
by: Chen, Jinyuan, et al.
Published: (2025)
by: Chen, Jinyuan, et al.
Published: (2025)
Cloud Revolution: Tracing the Origins and Rise of Cloud Computing
by: Gurung, Deepa, et al.
Published: (2025)
by: Gurung, Deepa, et al.
Published: (2025)
Similar Items
-
Advancing Environmental Sustainability in Data Centers via Carbon Depreciation Models
by: Ji, Shixin, et al.
Published: (2024) -
AGILE: Lightweight and Efficient Asynchronous GPU-SSD Integration
by: Yang, Zhuoping, et al.
Published: (2025) -
BSODiag: A Global Diagnosis Framework for Batch Servers Outage in Large-scale Cloud Infrastructure Systems
by: Duan, Tao, et al.
Published: (2025) -
More is Different: Prototyping and Analyzing a New Form of Edge Server with Massive Mobile SoCs
by: Zhang, Li, et al.
Published: (2022) -
EC2MoE: Adaptive End-Cloud Pipeline Collaboration Enabling Scalable Mixture-of-Experts Inference
by: Yang, Zheming, et al.
Published: (2025)