A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers
Fuente:
arXiv
Guardado en:
| Autores principales: | Saleh, Mohammad AlShaikh, Chawla, Sanjay, Bayhan, Sertac, Abu-Rub, Haitham, Ghrayeb, Ali |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GPUVM: GPU-driven Unified Virtual Memory
por: Nazaraliyev, Nurlan, et al.
Publicado: (2024)
por: Nazaraliyev, Nurlan, et al.
Publicado: (2024)
Validation of GPU Computation in Decentralized, Trustless Networks
por: Boniardi, Eric, et al.
Publicado: (2025)
por: Boniardi, Eric, et al.
Publicado: (2025)
CCCL: Node-Spanning GPU Collectives with CXL Memory Pooling
por: Xu, Dong, et al.
Publicado: (2026)
por: Xu, Dong, et al.
Publicado: (2026)
Towards Efficient and Practical GPU Multitasking in the Era of LLM
por: Xing, Jiarong, et al.
Publicado: (2025)
por: Xing, Jiarong, et al.
Publicado: (2025)
Peformance Isolation for Inference Processes in Edge GPU Systems
por: Martín, Juan José, et al.
Publicado: (2026)
por: Martín, Juan José, et al.
Publicado: (2026)
Performance Isolation and Semantic Determinism in Efficient GPU Spatial Sharing
por: Yang, Zhenyuan, et al.
Publicado: (2026)
por: Yang, Zhenyuan, et al.
Publicado: (2026)
NCCLbpf: Verified, Composable Policy Execution for GPU Collective Communication
por: Zheng, Yusheng
Publicado: (2026)
por: Zheng, Yusheng
Publicado: (2026)
GPUOS: A GPU Operating System Primitive for Transparent Operation Fusion
por: Yang, Yiwei, et al.
Publicado: (2026)
por: Yang, Yiwei, et al.
Publicado: (2026)
Night-Window Batching versus Carbon-Aware Scheduling for Clinical AI GPU Workloads
por: Doshi, Nishi, et al.
Publicado: (2026)
por: Doshi, Nishi, et al.
Publicado: (2026)
RAGDoll: Efficient Offloading-based Online RAG System on a Single GPU
por: Yu, Weiping, et al.
Publicado: (2025)
por: Yu, Weiping, et al.
Publicado: (2025)
PhoenixOS: Concurrent OS-level GPU Checkpoint and Restore with Validated Speculation
por: Wei, Xingda, et al.
Publicado: (2024)
por: Wei, Xingda, et al.
Publicado: (2024)
Scaling Sample-Based Quantum Diagonalization on GPU-Accelerated Systems using OpenMP Offload
por: Walkup, Robert, et al.
Publicado: (2026)
por: Walkup, Robert, et al.
Publicado: (2026)
The Landscape of GPU-Centric Communication
por: Unat, Didem, et al.
Publicado: (2024)
por: Unat, Didem, et al.
Publicado: (2024)
Long-Term or Temporary? Hybrid Worker Recruitment for Mobile Crowd Sensing and Computing
por: Liwang, Minghui, et al.
Publicado: (2022)
por: Liwang, Minghui, et al.
Publicado: (2022)
Optimizing Task Scheduling in Heterogeneous Computing Environments: A Comparative Analysis of CPU, GPU, and ASIC Platforms Using E2C Simulator
por: Mohammadjafari, Ali, et al.
Publicado: (2024)
por: Mohammadjafari, Ali, et al.
Publicado: (2024)
Mitigating GIL Bottlenecks in Edge AI Systems
por: Mandal, Mridankan, et al.
Publicado: (2026)
por: Mandal, Mridankan, et al.
Publicado: (2026)
Mitigating context switching in densely packed Linux clusters with Latency-Aware Group Scheduling
por: Isstaif, Al Amjad Tawfiq, et al.
Publicado: (2025)
por: Isstaif, Al Amjad Tawfiq, et al.
Publicado: (2025)
Data Trading and Monetization: Challenges and Open Research Directions
por: Ramadan, Qusai, et al.
Publicado: (2024)
por: Ramadan, Qusai, et al.
Publicado: (2024)
Mercury: QoS-Aware Tiered Memory System
por: Lu, Jiaheng, et al.
Publicado: (2024)
por: Lu, Jiaheng, et al.
Publicado: (2024)
MegaTrain: Full Precision Training of 100B+ Parameter Large Language Models on a Single GPU
por: Yuan, Zhengqing, et al.
Publicado: (2026)
por: Yuan, Zhengqing, et al.
Publicado: (2026)
GPU-Accelerated Quantum Simulation: Empirical Backend Selection, Gate Fusion, and Adaptive Precision
por: Kumaresan, Poornima, et al.
Publicado: (2026)
por: Kumaresan, Poornima, et al.
Publicado: (2026)
Demystifying Serverless Costs on Public Platforms: Bridging Billing, Architecture, and OS Scheduling
por: Lin, Changyuan, et al.
Publicado: (2025)
por: Lin, Changyuan, et al.
Publicado: (2025)
VUDA: Breaking CUDA-Vulkan Isolation for Spatial Sharing of Compute and Graphics on the Same GPU
por: Xu, Bin, et al.
Publicado: (2026)
por: Xu, Bin, et al.
Publicado: (2026)
Blockchain Transaction Conflicts: A Historical Perspective
por: Anjana, Parwat Singh, et al.
Publicado: (2025)
por: Anjana, Parwat Singh, et al.
Publicado: (2025)
Stencil Computations on Tenstorrent Wormhole
por: Piarulli, Lorenzo, et al.
Publicado: (2026)
por: Piarulli, Lorenzo, et al.
Publicado: (2026)
Single Bridge Formation in Self-Organizing Particle Systems
por: Oh, Shunhao, et al.
Publicado: (2024)
por: Oh, Shunhao, et al.
Publicado: (2024)
Scaling to 32 GPUs on a Novel Composable System Architecture
por: Ihnotic, John
Publicado: (2024)
por: Ihnotic, John
Publicado: (2024)
GrapheonRL: A Graph Neural Network and Reinforcement Learning Framework for Constraint and Data-Aware Workflow Mapping and Scheduling in Heterogeneous HPC Systems
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
Fork is All You Need in Heterogeneous Systems
por: Wang, Zixuan, et al.
Publicado: (2024)
por: Wang, Zixuan, et al.
Publicado: (2024)
Raft Distributed System for Multi-access Edge Computing Sharing Resources
por: Khaliq, Zain, et al.
Publicado: (2024)
por: Khaliq, Zain, et al.
Publicado: (2024)
A Scored Non-Deterministic Finite Automata Processor for Sequence Alignment
por: Karakchi, Ryan Karbowniczak Rasha
Publicado: (2024)
por: Karakchi, Ryan Karbowniczak Rasha
Publicado: (2024)
3D System Design: A Case for Building Customized Modular Systems in 3D
por: Emma, Philip, et al.
Publicado: (2024)
por: Emma, Philip, et al.
Publicado: (2024)
Instant Resonance: Dual Strategy Enhances the Data Consensus Success Rate of Blockchain Threshold Signature Oracles
por: Xian, Youquan, et al.
Publicado: (2024)
por: Xian, Youquan, et al.
Publicado: (2024)
LLMChain: Blockchain-based Reputation System for Sharing and Evaluating Large Language Models
por: Bouchiha, Mouhamed Amine, et al.
Publicado: (2024)
por: Bouchiha, Mouhamed Amine, et al.
Publicado: (2024)
Quantum Cloud Computing: A Review, Open Problems, and Future Directions
por: Nguyen, Hoa T., et al.
Publicado: (2024)
por: Nguyen, Hoa T., et al.
Publicado: (2024)
Is RISC-V Ready for Machine Learning? Portable Gaussian Processes Using Asynchronous Tasks
por: Strack, Alexander, et al.
Publicado: (2026)
por: Strack, Alexander, et al.
Publicado: (2026)
DRLQ: A Deep Reinforcement Learning-based Task Placement for Quantum Cloud Computing
por: Nguyen, Hoa T., et al.
Publicado: (2024)
por: Nguyen, Hoa T., et al.
Publicado: (2024)
Uniting the World by Dividing it: Federated Maps to Enable Spatial Applications
por: Bharadwaj, Sagar, et al.
Publicado: (2025)
por: Bharadwaj, Sagar, et al.
Publicado: (2025)
Parallel FFTW on RISC-V: A Comparative Study including OpenMP, MPI, and HPX
por: Strack, Alexander, et al.
Publicado: (2025)
por: Strack, Alexander, et al.
Publicado: (2025)
SPAARC: Spatial Proximity and Association based prefetching for Augmented Reality in edge Cache
por: Sreekumar, Nikhil, et al.
Publicado: (2025)
por: Sreekumar, Nikhil, et al.
Publicado: (2025)
Ejemplares similares
-
GPUVM: GPU-driven Unified Virtual Memory
por: Nazaraliyev, Nurlan, et al.
Publicado: (2024) -
Validation of GPU Computation in Decentralized, Trustless Networks
por: Boniardi, Eric, et al.
Publicado: (2025) -
CCCL: Node-Spanning GPU Collectives with CXL Memory Pooling
por: Xu, Dong, et al.
Publicado: (2026) -
Towards Efficient and Practical GPU Multitasking in the Era of LLM
por: Xing, Jiarong, et al.
Publicado: (2025) -
Peformance Isolation for Inference Processes in Edge GPU Systems
por: Martín, Juan José, et al.
Publicado: (2026)