Intent-driven scheduling of backup jobs
Fuente:
arXiv
Guardado en:
| Autores principales: | Dutta, Souvik, Brahmaroutu, Suri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
eScope: A Fine-Grained Power Prediction Mechanism for Mobile Applications
por: Mukherjee, Dipayan, et al.
Publicado: (2024)
por: Mukherjee, Dipayan, et al.
Publicado: (2024)
FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines
por: He, Jiaao, et al.
Publicado: (2024)
por: He, Jiaao, et al.
Publicado: (2024)
Comprehensive Plugin-Based Monitoring of Nexflow Workflow Executions
por: Kharma, Sami, et al.
Publicado: (2026)
por: Kharma, Sami, et al.
Publicado: (2026)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
por: Putra, Jody Almaida
Publicado: (2026)
por: Putra, Jody Almaida
Publicado: (2026)
OPTIMUMP2P: Fast and Reliable Gossiping in P2P Networks
por: Nicolaou, Nicolas, et al.
Publicado: (2025)
por: Nicolaou, Nicolas, et al.
Publicado: (2025)
The World's Fastest Matching Engine Algorithm
por: Yoon, Jake
Publicado: (2026)
por: Yoon, Jake
Publicado: (2026)
Profiling and optimization of multi-card GPU machine learning jobs
por: Lawenda, Marcin, et al.
Publicado: (2025)
por: Lawenda, Marcin, et al.
Publicado: (2025)
Efficient Construction of Large Search Spaces for Auto-Tuning
por: Willemsen, Floris-Jan, et al.
Publicado: (2025)
por: Willemsen, Floris-Jan, et al.
Publicado: (2025)
Comparative Analysis of Large Language Model Inference Serving Systems: A Performance Study of vLLM and HuggingFace TGI
por: Kolluru, Saicharan
Publicado: (2025)
por: Kolluru, Saicharan
Publicado: (2025)
Bine Trees: Enhancing Collective Operations by Optimizing Communication Locality
por: De Sensi, Daniele, et al.
Publicado: (2025)
por: De Sensi, Daniele, et al.
Publicado: (2025)
Studying the Effect of Schedule Preemption on Dynamic Task Graph Scheduling
por: Khodabandehlou, Mohammadali, et al.
Publicado: (2026)
por: Khodabandehlou, Mohammadali, et al.
Publicado: (2026)
LLAMP: Assessing Network Latency Tolerance of HPC Applications with Linear Programming
por: Shen, Siyuan, et al.
Publicado: (2024)
por: Shen, Siyuan, et al.
Publicado: (2024)
Exploiting Spot Instances for Time-Critical Cloud Workloads Using Optimal Randomized Strategies
por: Bhuyan, Neelkamal, et al.
Publicado: (2026)
por: Bhuyan, Neelkamal, et al.
Publicado: (2026)
Opportunistic Scheduling for Optimal Spot Instance Savings in the Cloud
por: Bhuyan, Neelkamal, et al.
Publicado: (2026)
por: Bhuyan, Neelkamal, et al.
Publicado: (2026)
Accelerating Causal Algorithms for Industrial-scale Data: A Distributed Computing Approach with Ray Framework
por: Verma, Vishal, et al.
Publicado: (2024)
por: Verma, Vishal, et al.
Publicado: (2024)
Rotary GPU: Exploring Local Execution Paths for Large Mixture-of-Experts Models Under Limited GPU Memory
por: Jo, Myeong Jun
Publicado: (2026)
por: Jo, Myeong Jun
Publicado: (2026)
Libra: Unleashing GPU Heterogeneity for High-Performance Sparse Matrix Multiplication
por: Shi, Jinliang, et al.
Publicado: (2025)
por: Shi, Jinliang, et al.
Publicado: (2025)
Serverless Cold Starts and Where to Find Them
por: Joosen, Artjom, et al.
Publicado: (2024)
por: Joosen, Artjom, et al.
Publicado: (2024)
A Methodology to Assess Power Modeling in Energy-Aware Federated Learning on Heterogeneous Mobile Devices
por: Jallouli, Chaimae, et al.
Publicado: (2026)
por: Jallouli, Chaimae, et al.
Publicado: (2026)
Cross-Platform Fused MoE Dispatch in Triton: Portable Expert Routing Without CUDA
por: Mitra, Subhadip
Publicado: (2026)
por: Mitra, Subhadip
Publicado: (2026)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
por: Yuan, Renzhong, et al.
Publicado: (2026)
por: Yuan, Renzhong, et al.
Publicado: (2026)
Optimization of a Radiofrequency Ablation FEM Application Using Parallel Sparse Solvers
por: Miletto, Marcelo Cogo, et al.
Publicado: (2024)
por: Miletto, Marcelo Cogo, et al.
Publicado: (2024)
Evaluating Emerging AI/ML Accelerators: IPU, RDU, and NVIDIA/AMD GPUs
por: Peng, Hongwu, et al.
Publicado: (2023)
por: Peng, Hongwu, et al.
Publicado: (2023)
GREEN-CODE: Learning to Optimize Energy Efficiency in LLM-based Code Generation
por: Ilager, Shashikant, et al.
Publicado: (2025)
por: Ilager, Shashikant, et al.
Publicado: (2025)
CGSim: A Simulation Framework for Large Scale Distributed Computing Environment
por: Vatsavai, Sairam Sri, et al.
Publicado: (2025)
por: Vatsavai, Sairam Sri, et al.
Publicado: (2025)
Swing: Short-cutting Rings for Higher Bandwidth Allreduce
por: De Sensi, Daniele, et al.
Publicado: (2024)
por: De Sensi, Daniele, et al.
Publicado: (2024)
MIREncoder: Multi-modal IR-based Pretrained Embeddings for Performance Optimizations
por: Dutta, Akash, et al.
Publicado: (2024)
por: Dutta, Akash, et al.
Publicado: (2024)
Relaxation for Efficient Asynchronous Queues
por: Baldwin, Samuel, et al.
Publicado: (2025)
por: Baldwin, Samuel, et al.
Publicado: (2025)
Federated Fine-Tuning of LLMs on the Very Edge: The Good, the Bad, the Ugly
por: Woisetschläger, Herbert, et al.
Publicado: (2023)
por: Woisetschläger, Herbert, et al.
Publicado: (2023)
Unlocking Python's Cores: Hardware Usage and Energy Implications of Removing the GIL
por: Salazar, José Daniel Montoya
Publicado: (2026)
por: Salazar, José Daniel Montoya
Publicado: (2026)
Serinv: A Scalable Library for the Selected Inversion of Block-Tridiagonal with Arrowhead Matrices
por: Maillou, Vincent, et al.
Publicado: (2025)
por: Maillou, Vincent, et al.
Publicado: (2025)
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
por: Guan, Bo
Publicado: (2026)
por: Guan, Bo
Publicado: (2026)
Scaling Large-scale GNN Training to Thousands of Processors on CPU-based Supercomputers
por: Zhuang, Chen, et al.
Publicado: (2024)
por: Zhuang, Chen, et al.
Publicado: (2024)
Staging Blocked Evaluation over Structured Sparse Matrices
por: Das, Pratyush, et al.
Publicado: (2024)
por: Das, Pratyush, et al.
Publicado: (2024)
Preliminary report: Initial evaluation of StdPar implementations on AMD GPUs for HPC
por: Lin, Wei-Chen, et al.
Publicado: (2024)
por: Lin, Wei-Chen, et al.
Publicado: (2024)
Reducing Tail Latencies Through Environment- and Neighbour-aware Thread Management
por: Jeffery, Andrew, et al.
Publicado: (2024)
por: Jeffery, Andrew, et al.
Publicado: (2024)
Dissecting the software-based measurement of CPU energy consumption: a comparative analysis
por: Raffin, Guillaume, et al.
Publicado: (2024)
por: Raffin, Guillaume, et al.
Publicado: (2024)
Asymptotically Optimal Scheduling of Multiple Parallelizable Job Classes
por: Berg, Benjamin, et al.
Publicado: (2024)
por: Berg, Benjamin, et al.
Publicado: (2024)
Performance Debugging through Microarchitectural Sensitivity and Causality Analysis
por: Dutilleul, Alban, et al.
Publicado: (2024)
por: Dutilleul, Alban, et al.
Publicado: (2024)
Matryoshka: Optimization of Dynamic Diverse Quantum Chemistry Systems via Elastic Parallelism Transformation
por: Wang, Tuowei, et al.
Publicado: (2024)
por: Wang, Tuowei, et al.
Publicado: (2024)
Ejemplares similares
-
eScope: A Fine-Grained Power Prediction Mechanism for Mobile Applications
por: Mukherjee, Dipayan, et al.
Publicado: (2024) -
FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines
por: He, Jiaao, et al.
Publicado: (2024) -
Comprehensive Plugin-Based Monitoring of Nexflow Workflow Executions
por: Kharma, Sami, et al.
Publicado: (2026) -
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
por: Putra, Jody Almaida
Publicado: (2026) -
OPTIMUMP2P: Fast and Reliable Gossiping in P2P Networks
por: Nicolaou, Nicolas, et al.
Publicado: (2025)