LookAhead: The Optimal Non-decreasing Index Policy for a Time-Varying Holding Cost problem
Fuente:
arXiv
Guardado en:
| Autores principales: | Gurushankar, Keerthana, Li, Zhouzi, Harchol-Balter, Mor, Scheller-Wolf, Alan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improving Upon the generalized c-mu rule: a Whittle approach
por: Li, Zhouzi, et al.
Publicado: (2025)
por: Li, Zhouzi, et al.
Publicado: (2025)
SPLIT: SymPathy for Large jobs Improves Tail latency
por: Li, Zhouzi, et al.
Publicado: (2026)
por: Li, Zhouzi, et al.
Publicado: (2026)
How to Rent GPUs on a Budget
por: Li, Zhouzi, et al.
Publicado: (2024)
por: Li, Zhouzi, et al.
Publicado: (2024)
An Upper Bound on the M/M/k Queue With Deterministic Setup Times
por: Williams, Jalani, et al.
Publicado: (2025)
por: Williams, Jalani, et al.
Publicado: (2025)
Can Increasing the Hit Ratio Hurt Cache Throughput? (Long Version)
por: Qiu, Ziyue, et al.
Publicado: (2024)
por: Qiu, Ziyue, et al.
Publicado: (2024)
Analysis of Markovian Arrivals and Service with Applications to Intermittent Overload
por: Grosof, Isaac, et al.
Publicado: (2024)
por: Grosof, Isaac, et al.
Publicado: (2024)
Asymptotically Optimal Scheduling of Multiple Parallelizable Job Classes
por: Berg, Benjamin, et al.
Publicado: (2024)
por: Berg, Benjamin, et al.
Publicado: (2024)
Mean field optimal Core Allocation across Malleable jobs
por: Li, Zhouzi, et al.
Publicado: (2026)
por: Li, Zhouzi, et al.
Publicado: (2026)
BOA Constrictor: Squeezing Performance out of GPUs in the Cloud via Budget-Optimal Allocation
por: Li, Zhouzi, et al.
Publicado: (2026)
por: Li, Zhouzi, et al.
Publicado: (2026)
When Does the Gittins Policy Have Asymptotically Optimal Response Time Tail?
por: Scully, Ziv, et al.
Publicado: (2021)
por: Scully, Ziv, et al.
Publicado: (2021)
LookAhead Tuning: Safer Language Models via Partial Answer Previews
por: Liu, Kangwei, et al.
Publicado: (2025)
por: Liu, Kangwei, et al.
Publicado: (2025)
LookAhead: Preventing DeFi Attacks via Unveiling Adversarial Contracts
por: Ren, Shoupeng, et al.
Publicado: (2024)
por: Ren, Shoupeng, et al.
Publicado: (2024)
Network Calculus Characterization of Congestion Control for Time-Varying Traffic
por: Lehal, Harvinder, et al.
Publicado: (2024)
por: Lehal, Harvinder, et al.
Publicado: (2024)
Scheduling with Uncertain Holding Costs and its Application to Content Moderation
por: Gocmen, Caner, et al.
Publicado: (2025)
por: Gocmen, Caner, et al.
Publicado: (2025)
Hold Onto That Thought: Assessing KV Cache Compression On Reasoning
por: Liu, Minghui, et al.
Publicado: (2025)
por: Liu, Minghui, et al.
Publicado: (2025)
Stability and Heavy-traffic Delay Optimality of General Load Balancing Policies in Heterogeneous Service Systems
por: Luo, Yishun, et al.
Publicado: (2025)
por: Luo, Yishun, et al.
Publicado: (2025)
On Resolving Non-Preemptivity in Multitask Scheduling: An Optimal Algorithm in Deterministic and Stochastic Worlds
por: Li, Wenxin
Publicado: (2024)
por: Li, Wenxin
Publicado: (2024)
WritePolicyBench: Benchmarking Memory Write Policies under Byte Budgets
por: Cham, Edgard El
Publicado: (2026)
por: Cham, Edgard El
Publicado: (2026)
Decision-Epoch Matters: Unveiling its Impact on the Stability of Scheduling with Randomly Varying Connectivity
por: Soprano-Loto, Nahuel, et al.
Publicado: (2024)
por: Soprano-Loto, Nahuel, et al.
Publicado: (2024)
CAMP: A Cost Adaptive Multi-Queue Eviction Policy for Key-Value Stores
por: Ghandeharizadeh, Shahram, et al.
Publicado: (2024)
por: Ghandeharizadeh, Shahram, et al.
Publicado: (2024)
Heavy-Traffic Optimal Size- and State-Aware Dispatching
por: Xie, Runhan, et al.
Publicado: (2023)
por: Xie, Runhan, et al.
Publicado: (2023)
On Combining Two Server Control Policies for Energy Efficiency
por: Dai, Jingze, et al.
Publicado: (2025)
por: Dai, Jingze, et al.
Publicado: (2025)
Improving Multiresource Job Scheduling with Markovian Service Rate Policies
por: Chen, Zhongrui, et al.
Publicado: (2025)
por: Chen, Zhongrui, et al.
Publicado: (2025)
Improving Multiresource Job Scheduling with Markovian Service Rate Policies
por: Chen, Zhongrui, et al.
Publicado: (2024)
por: Chen, Zhongrui, et al.
Publicado: (2024)
CEBench: A Benchmarking Toolkit for the Cost-Effectiveness of LLM Pipelines
por: Sun, Wenbo, et al.
Publicado: (2024)
por: Sun, Wenbo, et al.
Publicado: (2024)
Sharp bounds on $p$-norms for sums of independent uniform random variables, $0 < p < 1$
por: Chasapis, Giorgos, et al.
Publicado: (2021)
por: Chasapis, Giorgos, et al.
Publicado: (2021)
Exact Persistent Stochastic Non-Interference
por: Piazza, Carla, et al.
Publicado: (2025)
por: Piazza, Carla, et al.
Publicado: (2025)
Tabular and Deep Reinforcement Learning for Gittins Index
por: Dhankhar, Harshit, et al.
Publicado: (2024)
por: Dhankhar, Harshit, et al.
Publicado: (2024)
Spatiotemporal Non-Uniformity-Aware Online Task Scheduling in Collaborative Edge Computing for Industrial Internet of Things
por: Li, Yang, et al.
Publicado: (2025)
por: Li, Yang, et al.
Publicado: (2025)
PARD: Accelerating LLM Inference with Low-Cost PARallel Draft Model Adaptation
por: An, Zihao, et al.
Publicado: (2025)
por: An, Zihao, et al.
Publicado: (2025)
Shortest-Path FFT: Optimal SIMD Instruction Scheduling via Graph Search
por: Bergach, Mohamed Amine
Publicado: (2026)
por: Bergach, Mohamed Amine
Publicado: (2026)
Looking Forward: Challenges and Opportunities in Agentic AI Reliability
por: Xing, Liudong, et al.
Publicado: (2025)
por: Xing, Liudong, et al.
Publicado: (2025)
Starlink on the Road: A First Look at Mobile Starlink Performance in Central Europe
por: Laniewski, Dominic, et al.
Publicado: (2024)
por: Laniewski, Dominic, et al.
Publicado: (2024)
Computational Algorithms for the Product Form Solution of Closed Queuing Networks with Finite Buffers and Skip-Over Policy
por: Balbo, Gianfranco, et al.
Publicado: (2024)
por: Balbo, Gianfranco, et al.
Publicado: (2024)
Impact of AI-Triage on Radiologist Report Turnaround Time: Real-World Time-Savings and Insights from Model Predictions
por: Thompson, Yee Lam Elim, et al.
Publicado: (2025)
por: Thompson, Yee Lam Elim, et al.
Publicado: (2025)
TINA: Acceleration of Non-NN Signal Processing Algorithms Using NN Accelerators
por: Boerkamp, Christiaan, et al.
Publicado: (2024)
por: Boerkamp, Christiaan, et al.
Publicado: (2024)
Tail Optimality and Performance Analysis of the Nudge*(M) Scheduling Algorithm
por: Charlet, Nils, et al.
Publicado: (2024)
por: Charlet, Nils, et al.
Publicado: (2024)
Non-Asymptotic Performance Analysis of DOA Estimation Based on Real-Valued Root-MUSIC
por: Liu, Junyang, et al.
Publicado: (2025)
por: Liu, Junyang, et al.
Publicado: (2025)
Faster LLM Inference using DBMS-Inspired Preemption and Cache Replacement Policies
por: Kim, Kyoungmin, et al.
Publicado: (2024)
por: Kim, Kyoungmin, et al.
Publicado: (2024)
Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using Dask
por: Abraham, Ashley N., et al.
Publicado: (2026)
por: Abraham, Ashley N., et al.
Publicado: (2026)
Ejemplares similares
-
Improving Upon the generalized c-mu rule: a Whittle approach
por: Li, Zhouzi, et al.
Publicado: (2025) -
SPLIT: SymPathy for Large jobs Improves Tail latency
por: Li, Zhouzi, et al.
Publicado: (2026) -
How to Rent GPUs on a Budget
por: Li, Zhouzi, et al.
Publicado: (2024) -
An Upper Bound on the M/M/k Queue With Deterministic Setup Times
por: Williams, Jalani, et al.
Publicado: (2025) -
Can Increasing the Hit Ratio Hurt Cache Throughput? (Long Version)
por: Qiu, Ziyue, et al.
Publicado: (2024)