GridPilot: Real-Time Grid-Responsive Control for AI Supercomputers
Fuente:
arXiv
Salvato in:
| Autori principali: | Constantinescu, Denisa-Andreea, Atienza, David |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Flex-MIG: Enabling Distributed Execution on MIG
di: Kim, Myeongsu, et al.
Pubblicazione: (2025)
di: Kim, Myeongsu, et al.
Pubblicazione: (2025)
ZenFlow: Enabling Stall-Free Offloading Training via Asynchronous Updates
di: Lan, Tingfeng, et al.
Pubblicazione: (2025)
di: Lan, Tingfeng, et al.
Pubblicazione: (2025)
Scheduler-Driven Job Atomization
di: Konopa, Michal, et al.
Pubblicazione: (2025)
di: Konopa, Michal, et al.
Pubblicazione: (2025)
JASDA: Introducing Job-Aware Scheduling in Scheduler-Driven Job Atomization
di: Konopa, Michal, et al.
Pubblicazione: (2025)
di: Konopa, Michal, et al.
Pubblicazione: (2025)
Serverless Cold Starts and Where to Find Them
di: Joosen, Artjom, et al.
Pubblicazione: (2024)
di: Joosen, Artjom, et al.
Pubblicazione: (2024)
Efficiently Scheduling Parallel DAG Tasks on Identical Multiprocessors
di: Lendve, Shardul, et al.
Pubblicazione: (2024)
di: Lendve, Shardul, et al.
Pubblicazione: (2024)
ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving
di: Li, Xiangchen, et al.
Pubblicazione: (2026)
di: Li, Xiangchen, et al.
Pubblicazione: (2026)
WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
di: Li, Xiangchen, et al.
Pubblicazione: (2026)
di: Li, Xiangchen, et al.
Pubblicazione: (2026)
Laminar: A Probe-First Scheduling Paradigm with Deterministic Runtime Survival
di: Chu, Zhengyan
Pubblicazione: (2026)
di: Chu, Zhengyan
Pubblicazione: (2026)
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
di: Xie, Tianfang
Pubblicazione: (2026)
di: Xie, Tianfang
Pubblicazione: (2026)
astroCAMP: A Community Benchmark and Co-Design Framework for Sustainable SKA-Scale Radio Imaging
di: Constantinescu, Denisa-Andreea, et al.
Pubblicazione: (2025)
di: Constantinescu, Denisa-Andreea, et al.
Pubblicazione: (2025)
push0: Scalable and Fault-Tolerant Orchestration for Zero-Knowledge Proof Generation
di: Ahmadvand, Mohsen, et al.
Pubblicazione: (2026)
di: Ahmadvand, Mohsen, et al.
Pubblicazione: (2026)
Evaluating Large Language Models for Workload Mapping and Scheduling in Heterogeneous HPC Systems
di: Sharma, Aasish Kumar, et al.
Pubblicazione: (2025)
di: Sharma, Aasish Kumar, et al.
Pubblicazione: (2025)
PoCL-R: An Open Standard Based Offloading Layer for Heterogeneous Multi-Access Edge Computing with Server Side Scalability
di: Solanti, Jan, et al.
Pubblicazione: (2023)
di: Solanti, Jan, et al.
Pubblicazione: (2023)
Toward a Universal GPU Instruction Set Architecture: A Cross-Vendor Analysis of Hardware-Invariant Computational Primitives in Parallel Processors
di: Abraham, Ojima, et al.
Pubblicazione: (2026)
di: Abraham, Ojima, et al.
Pubblicazione: (2026)
Parallelization Strategies for Dense LLM Deployment: Navigating Through Application-Specific Tradeoffs and Bottlenecks
di: Topcu, Burak, et al.
Pubblicazione: (2026)
di: Topcu, Burak, et al.
Pubblicazione: (2026)
Semaphores Augmented with a Waiting Array
di: Dice, Dave, et al.
Pubblicazione: (2025)
di: Dice, Dave, et al.
Pubblicazione: (2025)
Reciprocating Locks
di: Dice, Dave, et al.
Pubblicazione: (2025)
di: Dice, Dave, et al.
Pubblicazione: (2025)
Hapax Locks : Value-Based Mutual Exclusion
di: Dice, Dave, et al.
Pubblicazione: (2025)
di: Dice, Dave, et al.
Pubblicazione: (2025)
Intelligent Cloud Orchestration: A Hybrid Predictive and Heuristic Framework for Cost Optimization
di: Nagoriya, Heet, et al.
Pubblicazione: (2026)
di: Nagoriya, Heet, et al.
Pubblicazione: (2026)
CEO-DC: Driving Decarbonization in HPC Data Centers with Actionable Insights
di: Álvarez, Rubén Rodríguez, et al.
Pubblicazione: (2025)
di: Álvarez, Rubén Rodríguez, et al.
Pubblicazione: (2025)
SAGA: Workflow-Atomic Scheduling for AI Agent Inference on GPU Clusters
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
DPDPU: Data Processing with DPUs
di: Hu, Jiasheng, et al.
Pubblicazione: (2024)
di: Hu, Jiasheng, et al.
Pubblicazione: (2024)
LAMMPS-KOKKOS: Performance Portable Molecular Dynamics Across Exascale Architectures
di: Johansson, Anders, et al.
Pubblicazione: (2025)
di: Johansson, Anders, et al.
Pubblicazione: (2025)
Unlocking Python's Cores: Hardware Usage and Energy Implications of Removing the GIL
di: Salazar, José Daniel Montoya
Pubblicazione: (2026)
di: Salazar, José Daniel Montoya
Pubblicazione: (2026)
Reexamining Paradigms of End-to-End Data Movement
di: Fang, Chin, et al.
Pubblicazione: (2025)
di: Fang, Chin, et al.
Pubblicazione: (2025)
Big Data Workload Profiling for Energy-Aware Cloud Resource Management
di: Parikh, Milan, et al.
Pubblicazione: (2026)
di: Parikh, Milan, et al.
Pubblicazione: (2026)
Rotary GPU: Exploring Local Execution Paths for Large Mixture-of-Experts Models Under Limited GPU Memory
di: Jo, Myeong Jun
Pubblicazione: (2026)
di: Jo, Myeong Jun
Pubblicazione: (2026)
A Methodology to Assess Power Modeling in Energy-Aware Federated Learning on Heterogeneous Mobile Devices
di: Jallouli, Chaimae, et al.
Pubblicazione: (2026)
di: Jallouli, Chaimae, et al.
Pubblicazione: (2026)
Boosting Cross-Architectural Emulation Performance by Foregoing the Intermediate Representation Model
di: Parker, Amy Iris
Pubblicazione: (2025)
di: Parker, Amy Iris
Pubblicazione: (2025)
Lincoln AI Computing Survey (LAICS) and Trends
di: Reuther, Albert, et al.
Pubblicazione: (2025)
di: Reuther, Albert, et al.
Pubblicazione: (2025)
Guess-Verify-Refine: Data-Aware Top-K for Sparse-Attention Decoding on Blackwell via Temporal Correlation
di: Cheng, Long, et al.
Pubblicazione: (2026)
di: Cheng, Long, et al.
Pubblicazione: (2026)
nvidia-pcm: A D-Bus-Driven Platform Configuration Manager for OpenBMC Environments
di: Singh, Harinder
Pubblicazione: (2026)
di: Singh, Harinder
Pubblicazione: (2026)
NM-SpMM: Accelerating Matrix Multiplication Using N:M Sparsity with GPGPU
di: Ma, Cong, et al.
Pubblicazione: (2025)
di: Ma, Cong, et al.
Pubblicazione: (2025)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
di: Putra, Jody Almaida
Pubblicazione: (2026)
di: Putra, Jody Almaida
Pubblicazione: (2026)
A Virtual Processor brings back the Free Lunch
di: Kutschbach, Haymo
Pubblicazione: (2026)
di: Kutschbach, Haymo
Pubblicazione: (2026)
Next-Generation Event-Driven Architectures: Performance, Scalability, and Intelligent Orchestration Across Messaging Frameworks
di: Arafat, Jahidul, et al.
Pubblicazione: (2025)
di: Arafat, Jahidul, et al.
Pubblicazione: (2025)
Implementation and Evaluation of Fast Raft for Hierarchical Consensus
di: Melnychuk, Anton, et al.
Pubblicazione: (2025)
di: Melnychuk, Anton, et al.
Pubblicazione: (2025)
Rhizomes and Diffusions for Processing Highly Skewed Graphs on Fine-Grain Message-Driven Systems
di: Chandio, Bibrak Qamar, et al.
Pubblicazione: (2024)
di: Chandio, Bibrak Qamar, et al.
Pubblicazione: (2024)
Predictive Multi-Tier Memory Management for KV Cache in Large-Scale GPU Inference
di: Ganjihal, Sanjeev Rao
Pubblicazione: (2026)
di: Ganjihal, Sanjeev Rao
Pubblicazione: (2026)
Documenti analoghi
-
Flex-MIG: Enabling Distributed Execution on MIG
di: Kim, Myeongsu, et al.
Pubblicazione: (2025) -
ZenFlow: Enabling Stall-Free Offloading Training via Asynchronous Updates
di: Lan, Tingfeng, et al.
Pubblicazione: (2025) -
Scheduler-Driven Job Atomization
di: Konopa, Michal, et al.
Pubblicazione: (2025) -
JASDA: Introducing Job-Aware Scheduling in Scheduler-Driven Job Atomization
di: Konopa, Michal, et al.
Pubblicazione: (2025) -
Serverless Cold Starts and Where to Find Them
di: Joosen, Artjom, et al.
Pubblicazione: (2024)