Bodega: Serving Linearizable Reads Locally from Anywhere at Anytime via Roster Leases
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Guanzhou, Arpaci-Dusseau, Andrea, Arpaci-Dusseau, Remzi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Unified, Practical, and Understandable Model of Non-transactional Consistency Levels in Distributed Replication
by: Hu, Guanzhou, et al.
Published: (2024)
by: Hu, Guanzhou, et al.
Published: (2024)
Crossword: Adaptive Consensus for Dynamic Data-Heavy Workloads
by: Hu, Guanzhou, et al.
Published: (2025)
by: Hu, Guanzhou, et al.
Published: (2025)
TensorHub: Scalable and Elastic Weight Transfer for LLM RL Training
by: Ye, Chenhao, et al.
Published: (2026)
by: Ye, Chenhao, et al.
Published: (2026)
Foreactor: Exploiting Storage I/O Parallelism with Explicit Speculation
by: Hu, Guanzhou, et al.
Published: (2024)
by: Hu, Guanzhou, et al.
Published: (2024)
Towards Reconfigurable Linearizable Reads
by: Thiessen, Myles, et al.
Published: (2024)
by: Thiessen, Myles, et al.
Published: (2024)
LeaseGuard: Raft Leases Done Right
by: Davis, A. Jesse Jiryu, et al.
Published: (2025)
by: Davis, A. Jesse Jiryu, et al.
Published: (2025)
Efficient Wait-Free Linearizable Implementations of Approximate Bounded Counters Using Read-Write Registers
by: Johnen, Colette, et al.
Published: (2024)
by: Johnen, Colette, et al.
Published: (2024)
Communication Requirements for Linearizable Registers
by: Nataf, Raïssa, et al.
Published: (2026)
by: Nataf, Raïssa, et al.
Published: (2026)
The Power of Strong Linearizability: the Difficulty of Consistent Refereeing
by: Attiya, Hagit, et al.
Published: (2025)
by: Attiya, Hagit, et al.
Published: (2025)
Linearizability and State-Machine Replication: Is it a match?
by: Hauck, Franz J., et al.
Published: (2024)
by: Hauck, Franz J., et al.
Published: (2024)
Strong Linearizability using Primitives with Consensus Number 2
by: Attiya, Hagit, et al.
Published: (2024)
by: Attiya, Hagit, et al.
Published: (2024)
Asynchronous Wait-Free Runtime Verification and Enforcement of Linearizability
by: Castañeda, Armando, et al.
Published: (2023)
by: Castañeda, Armando, et al.
Published: (2023)
Tame the Wild with Byzantine Linearizability: Reliable Broadcast, Snapshots, and Asset Transfer
by: Cohen, Shir, et al.
Published: (2021)
by: Cohen, Shir, et al.
Published: (2021)
Anywhere: A Web Crawler Automation Management Interface
by: Lin, Jinwei
Published: (2024)
by: Lin, Jinwei
Published: (2024)
LARK -- Linearizability Algorithms for Replicated Keys in Aerospike
by: Goodng, Andrew, et al.
Published: (2025)
by: Goodng, Andrew, et al.
Published: (2025)
Getting the MOST out of your Storage Hierarchy with Mirror-Optimized Storage Tiering
by: Tu, Kaiwei, et al.
Published: (2025)
by: Tu, Kaiwei, et al.
Published: (2025)
MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
by: Hu, Cunchen, et al.
Published: (2024)
by: Hu, Cunchen, et al.
Published: (2024)
Enabling Efficient Batch Serving for LMaaS via Generation Length Prediction
by: Cheng, Ke, et al.
Published: (2024)
by: Cheng, Ke, et al.
Published: (2024)
BrownoutServe: SLO-Aware Inference Serving under Bursty Workloads for MoE-based LLMs
by: Hu, Jianmin, et al.
Published: (2025)
by: Hu, Jianmin, et al.
Published: (2025)
DistServe: Disaggregating Prefill and Decoding for Goodput-optimized Large Language Model Serving
by: Zhong, Yinmin, et al.
Published: (2024)
by: Zhong, Yinmin, et al.
Published: (2024)
DeepServe: Serverless Large Language Model Serving at Scale
by: Hu, Junhao, et al.
Published: (2025)
by: Hu, Junhao, et al.
Published: (2025)
Harpagon: Minimizing DNN Serving Cost via Efficient Dispatching, Scheduling and Splitting
by: Zhao, Zhixin, et al.
Published: (2024)
by: Zhao, Zhixin, et al.
Published: (2024)
DualScale: Energy-Efficient Disaggregated LLM Serving via Phase-Aware Placement and DVFS
by: Basit, Omar, et al.
Published: (2026)
by: Basit, Omar, et al.
Published: (2026)
PPipe: Efficient Video Analytics Serving on Heterogeneous GPU Clusters via Pool-Based Pipeline Parallelism
by: Kong, Z. Jonny, et al.
Published: (2025)
by: Kong, Z. Jonny, et al.
Published: (2025)
Angelfish: Leader, DAG, or Anywhere in Between
by: Yu, Qianyu, et al.
Published: (2025)
by: Yu, Qianyu, et al.
Published: (2025)
BanaServe: Unified KV Cache and Dynamic Module Migration for Balancing Disaggregated LLM Serving in AI Infrastructure
by: He, Yiyuan, et al.
Published: (2025)
by: He, Yiyuan, et al.
Published: (2025)
EdgeServing: Deadline-Aware Multi-DNN Serving at the Edge
by: Cao, Jiahe, et al.
Published: (2026)
by: Cao, Jiahe, et al.
Published: (2026)
OTAS: An Elastic Transformer Serving System via Token Adaptation
by: Chen, Jinyu, et al.
Published: (2024)
by: Chen, Jinyu, et al.
Published: (2024)
DynaServe: Unified and Elastic Execution for Dynamic Disaggregated LLM Serving
by: Ruan, Chaoyi, et al.
Published: (2025)
by: Ruan, Chaoyi, et al.
Published: (2025)
TridentServe: A Stage-level Serving System for Diffusion Pipelines
by: Xia, Yifei, et al.
Published: (2025)
by: Xia, Yifei, et al.
Published: (2025)
MuxServe: Flexible Spatial-Temporal Multiplexing for Multiple LLM Serving
by: Duan, Jiangfei, et al.
Published: (2024)
by: Duan, Jiangfei, et al.
Published: (2024)
ThunderServe: High-performance and Cost-efficient LLM Serving in Cloud Environments
by: Jiang, Youhe, et al.
Published: (2025)
by: Jiang, Youhe, et al.
Published: (2025)
ServeGen: Workload Characterization and Generation of Large Language Model Serving in Production
by: Xiang, Yuxing, et al.
Published: (2025)
by: Xiang, Yuxing, et al.
Published: (2025)
EconoServe: Maximizing Multi-Resource Utilization with SLO Guarantees in LLM Serving
by: Shen, Haiying, et al.
Published: (2024)
by: Shen, Haiying, et al.
Published: (2024)
EcoServe: Designing Carbon-Aware AI Inference Systems
by: Li, Yueying, et al.
Published: (2025)
by: Li, Yueying, et al.
Published: (2025)
ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL
by: Gao, Wei, et al.
Published: (2026)
by: Gao, Wei, et al.
Published: (2026)
OServe: Accelerating LLM Serving via Spatial-Temporal Workload Orchestration
by: Jiang, Youhe, et al.
Published: (2026)
by: Jiang, Youhe, et al.
Published: (2026)
HydraServe: Minimizing Cold Start Latency for Serverless LLM Serving in Public Clouds
by: Lou, Chiheng, et al.
Published: (2025)
by: Lou, Chiheng, et al.
Published: (2025)
SparseServe: Unlocking Parallelism for Dynamic Sparse Attention in Long-Context LLM Serving
by: Zhou, Qihui, et al.
Published: (2025)
by: Zhou, Qihui, et al.
Published: (2025)
Slice-Level Scheduling for High Throughput and Load Balanced LLM Serving
by: Cheng, Ke, et al.
Published: (2024)
by: Cheng, Ke, et al.
Published: (2024)
Similar Items
-
A Unified, Practical, and Understandable Model of Non-transactional Consistency Levels in Distributed Replication
by: Hu, Guanzhou, et al.
Published: (2024) -
Crossword: Adaptive Consensus for Dynamic Data-Heavy Workloads
by: Hu, Guanzhou, et al.
Published: (2025) -
TensorHub: Scalable and Elastic Weight Transfer for LLM RL Training
by: Ye, Chenhao, et al.
Published: (2026) -
Foreactor: Exploiting Storage I/O Parallelism with Explicit Speculation
by: Hu, Guanzhou, et al.
Published: (2024) -
Towards Reconfigurable Linearizable Reads
by: Thiessen, Myles, et al.
Published: (2024)