Predictive Bayesian Arbitration: A Scalable Noisy-OR Model with Service Criticality Awareness
Fuente:
arXiv
Saved in:
| Main Authors: | Jangam, Anil, Rajendran, Ganesh Karthick, Kantharajah, Roy |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Arbitration-Free Consistency is Available (and Vice Versa)
by: Attiya, Hagit, et al.
Published: (2025)
by: Attiya, Hagit, et al.
Published: (2025)
Fast Topology-Aware Lossy Data Compression with Full Preservation of Critical Points and Local Order
by: Fallin, Alex, et al.
Published: (2026)
by: Fallin, Alex, et al.
Published: (2026)
ERA: Epoch-Resolved Arbitration for Duelling Admins in Group Management CRDTs
by: Dougal, Kegan
Published: (2026)
by: Dougal, Kegan
Published: (2026)
Fast and Robust Information Spreading in the Noisy PULL Model
by: D'Archivio, Niccolò, et al.
Published: (2024)
by: D'Archivio, Niccolò, et al.
Published: (2024)
Kairos: A Scalable Serving System for Physical AI
by: Dai, Yinwei, et al.
Published: (2026)
by: Dai, Yinwei, et al.
Published: (2026)
SCAREY: Location-Aware Service Lifecycle Management
by: Horvath, Kurt, et al.
Published: (2025)
by: Horvath, Kurt, et al.
Published: (2025)
SIMT/GPU Data Race Verification using ISCC and Intermediary Code Representations: A Case Study
by: Osterhout, Andrew, et al.
Published: (2025)
by: Osterhout, Andrew, et al.
Published: (2025)
SeaLLM: Service-Aware and Latency-Optimized Resource Sharing for Large Language Model Inference
by: Zhao, Yihao, et al.
Published: (2025)
by: Zhao, Yihao, et al.
Published: (2025)
FLARE: A Dataflow-Aware and Scalable Hardware Architecture for Neural-Hybrid Scientific Lossy Compression
by: Jia, Wenqi, et al.
Published: (2025)
by: Jia, Wenqi, et al.
Published: (2025)
Hardware-Aware Reformulation of Convolutions for Efficient Execution on Specialized AI Hardware: A Case Study on NVIDIA Tensor Cores
by: Bikshandi, Ganesh
Published: (2026)
by: Bikshandi, Ganesh
Published: (2026)
A Cloud in the Sky: Geo-Aware On-board Data Services for LEO Satellites
by: Sandholm, Thomas, et al.
Published: (2024)
by: Sandholm, Thomas, et al.
Published: (2024)
Performance and Security Aware Distributed Service Placement in Fog Computing
by: Goudarzi, Mohammad, et al.
Published: (2026)
by: Goudarzi, Mohammad, et al.
Published: (2026)
Expert-as-a-Service: Towards Efficient, Scalable, and Robust Large-scale MoE Serving
by: Liu, Ziming, et al.
Published: (2025)
by: Liu, Ziming, et al.
Published: (2025)
AI-driven Predictive Shard Allocation for Scalable Next Generation Blockchains
by: Haider, M. Zeeshan, et al.
Published: (2025)
by: Haider, M. Zeeshan, et al.
Published: (2025)
A Risk-Aware UAV-Edge Service Framework for Wildfire Monitoring and Emergency Response
by: Huang, Yulun, et al.
Published: (2026)
by: Huang, Yulun, et al.
Published: (2026)
TopoSZp: Lightweight Topology-Aware Error-controlled Compression for Scientific Data
by: Agarwal, Tripti, et al.
Published: (2026)
by: Agarwal, Tripti, et al.
Published: (2026)
SWARM+: Scalable and Resilient Multi-Agent Consensus for Fully-Decentralized Data-Aware Workload Management
by: Thareja, Komal, et al.
Published: (2026)
by: Thareja, Komal, et al.
Published: (2026)
QEdgeProxy: QoS-Aware Load Balancing for IoT Services in the Computing Continuum
by: Čilić, Ivan, et al.
Published: (2024)
by: Čilić, Ivan, et al.
Published: (2024)
EcoLife: Carbon-Aware Serverless Function Scheduling for Sustainable Computing
by: Jiang, Yankai, et al.
Published: (2024)
by: Jiang, Yankai, et al.
Published: (2024)
Predictive Sectorization and Bayesian Optimized Consensus for Admission Control in Autonomous Airspace Operations
by: Dhodapkar, Aditya, et al.
Published: (2026)
by: Dhodapkar, Aditya, et al.
Published: (2026)
e112: A Context-Aware Mobile Emergency Communication Platform Leveraging Smartphone Sensing and Cloud Services
by: Ioannidou, Katerina, et al.
Published: (2026)
by: Ioannidou, Katerina, et al.
Published: (2026)
Morpheus: Lightweight RTT Prediction for Performance-Aware Load Balancing
by: Giannakopoulos, Panagiotis, et al.
Published: (2025)
by: Giannakopoulos, Panagiotis, et al.
Published: (2025)
Predictive-LoRA: A Proactive and Fragmentation-Aware Serverless Inference System for LLMs
by: Ni, Yinan, et al.
Published: (2025)
by: Ni, Yinan, et al.
Published: (2025)
Seer: Proactive Revenue-Aware Scheduling for Live Streaming Services in Crowdsourced Cloud-Edge Platforms
by: Huang, Shaoyuan, et al.
Published: (2024)
by: Huang, Shaoyuan, et al.
Published: (2024)
HiRace: Accurate and Fast Source-Level Race Checking of GPU Programs
by: Jacobson, John, et al.
Published: (2024)
by: Jacobson, John, et al.
Published: (2024)
Aragog: Just-in-Time Model Routing for Scalable Serving of Agentic Workflows
by: Dai, Yinwei, et al.
Published: (2025)
by: Dai, Yinwei, et al.
Published: (2025)
Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda
by: Xu, Minxian, et al.
Published: (2026)
by: Xu, Minxian, et al.
Published: (2026)
EES-CND: Collaborative Neural Decision-Making for Drift-Aware Fault-Tolerant Edge-Cloud Service Placement
by: Herabad, Mohammadsadeq Garshasbi, et al.
Published: (2026)
by: Herabad, Mohammadsadeq Garshasbi, et al.
Published: (2026)
AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
Towards Privacy-, Budget-, and Deadline-Aware Service Optimization for Large Medical Image Processing across Hybrid Clouds
by: Wang, Yuandou, et al.
Published: (2024)
by: Wang, Yuandou, et al.
Published: (2024)
A Communication- and Memory-Aware Model for Load Balancing Tasks
by: Lifflander, Jonathan, et al.
Published: (2024)
by: Lifflander, Jonathan, et al.
Published: (2024)
Unleashing Scalable Context Parallelism for Foundation Models Pre-Training via FCP
by: Zhao, Yilong, et al.
Published: (2026)
by: Zhao, Yilong, et al.
Published: (2026)
A GPU accelerated mixed-precision Smoothed Particle Hydrodynamics framework with cell-based relative coordinates
by: Mao, Zirui, et al.
Published: (2023)
by: Mao, Zirui, et al.
Published: (2023)
A More Scalable Sparse Dynamic Data Exchange
by: Geyko, Andrew, et al.
Published: (2023)
by: Geyko, Andrew, et al.
Published: (2023)
A Scalable Recipe on SuperMUC-NG Phase 2: Efficient Large-Scale Training of Language Models
by: Rajgopal, Ajay Navilarekal, et al.
Published: (2026)
by: Rajgopal, Ajay Navilarekal, et al.
Published: (2026)
Scalable and Performant Data Loading
by: Hira, Moto, et al.
Published: (2025)
by: Hira, Moto, et al.
Published: (2025)
PRISM: Probabilistic Runtime Insights and Scalable Performance Modeling for Large-Scale Distributed Training
by: Golden, Alicia, et al.
Published: (2025)
by: Golden, Alicia, et al.
Published: (2025)
Sharding Distributed Databases: A Critical Review
by: Solat, Siamak
Published: (2024)
by: Solat, Siamak
Published: (2024)
MOPAR: A Model Partitioning Framework for Deep Learning Inference Services on Serverless Platforms
by: Duan, Jiaang, et al.
Published: (2024)
by: Duan, Jiaang, et al.
Published: (2024)
Service-Level Energy Modeling and Experimentation for Cloud-Native Microservices
by: Legler, Julian, et al.
Published: (2025)
by: Legler, Julian, et al.
Published: (2025)
Similar Items
-
Arbitration-Free Consistency is Available (and Vice Versa)
by: Attiya, Hagit, et al.
Published: (2025) -
Fast Topology-Aware Lossy Data Compression with Full Preservation of Critical Points and Local Order
by: Fallin, Alex, et al.
Published: (2026) -
ERA: Epoch-Resolved Arbitration for Duelling Admins in Group Management CRDTs
by: Dougal, Kegan
Published: (2026) -
Fast and Robust Information Spreading in the Noisy PULL Model
by: D'Archivio, Niccolò, et al.
Published: (2024) -
Kairos: A Scalable Serving System for Physical AI
by: Dai, Yinwei, et al.
Published: (2026)