Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
Fuente:
arXiv
Saved in:
| Main Authors: | Pathak, Prashant Kumar, Mouleeswaran, Chandra Biksheswaran, Repaka, Rama Teja |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
dpBento: Benchmarking DPUs for Data Processing
by: Hu, Jiasheng, et al.
Published: (2025)
by: Hu, Jiasheng, et al.
Published: (2025)
Theseus: A Distributed and Scalable GPU-Accelerated Query Processing Platform Optimized for Efficient Data Movement
by: Aramburú, Felipe, et al.
Published: (2025)
by: Aramburú, Felipe, et al.
Published: (2025)
DDS: DPU-optimized Disaggregated Storage [Extended Report]
by: Zhang, Qizhen, et al.
Published: (2024)
by: Zhang, Qizhen, et al.
Published: (2024)
Fifty Years of Transaction Processing Research (extended)
by: Bernstein, Philip A.
Published: (2026)
by: Bernstein, Philip A.
Published: (2026)
Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
DPDPU: Data Processing with DPUs
by: Hu, Jiasheng, et al.
Published: (2024)
by: Hu, Jiasheng, et al.
Published: (2024)
Characterizing and Fixing Silent Data Loss in Spark-on-AWS-Lambda with Open Table Formats
by: Gandla, Srujan Kumar
Published: (2026)
by: Gandla, Srujan Kumar
Published: (2026)
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
by: Guan, Bo
Published: (2026)
by: Guan, Bo
Published: (2026)
A Survey on Transactional Stream Processing
by: Zhang, Shuhao, et al.
Published: (2022)
by: Zhang, Shuhao, et al.
Published: (2022)
GPU-Augmented OLAP Execution Engine: GPU Offloading
by: Chang, Ilsun
Published: (2025)
by: Chang, Ilsun
Published: (2025)
Passing the Baton: High Throughput Distributed Disk-Based Vector Search with BatANN
by: Dang, Nam Anh, et al.
Published: (2025)
by: Dang, Nam Anh, et al.
Published: (2025)
Combining Serverless and High-Performance Computing Paradigms to support ML Data-Intensive Applications
by: Staylor, Mills, et al.
Published: (2025)
by: Staylor, Mills, et al.
Published: (2025)
Deep RC: A Scalable Data Engineering and Deep Learning Pipeline
by: Sarker, Arup Kumar, et al.
Published: (2025)
by: Sarker, Arup Kumar, et al.
Published: (2025)
Design and Implementation of an Analysis Pipeline for Heterogeneous Data
by: Sarker, Arup Kumar, et al.
Published: (2024)
by: Sarker, Arup Kumar, et al.
Published: (2024)
Shipwright: Proving liveness of distributed systems with Byzantine participants
by: Leung, Derek, et al.
Published: (2025)
by: Leung, Derek, et al.
Published: (2025)
Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems
by: Sharma, Aasish Kumar, et al.
Published: (2026)
by: Sharma, Aasish Kumar, et al.
Published: (2026)
Valori: A Deterministic Memory Substrate for AI Systems
by: Gudur, Varshith
Published: (2025)
by: Gudur, Varshith
Published: (2025)
Operational Memory Architecture for Kubernetes:Preserving Causal Context Across the Evidence Horizon
by: Khan, Shamsher
Published: (2026)
by: Khan, Shamsher
Published: (2026)
Verifying In-Network Computing Systems for Design Risks
by: Bai, Tianyu, et al.
Published: (2026)
by: Bai, Tianyu, et al.
Published: (2026)
Self-Adaptive Probabilistic Skyline Query Processing in Distributed Edge Computing via Deep Reinforcement Learning
by: Lai, Chuan-Chi
Published: (2026)
by: Lai, Chuan-Chi
Published: (2026)
Analysis of Design Patterns and Benchmark Practices in Apache Kafka Event-Streaming Systems
by: Mohammad, Muzeeb
Published: (2025)
by: Mohammad, Muzeeb
Published: (2025)
Deadline-Aware Joint Task Scheduling and Offloading in Mobile Edge Computing Systems
by: Nguyen, Ngoc Hung, et al.
Published: (2025)
by: Nguyen, Ngoc Hung, et al.
Published: (2025)
Experimentally Evaluating the Resource Efficiency of Big Data Autoscaling
by: Will, Jonathan, et al.
Published: (2025)
by: Will, Jonathan, et al.
Published: (2025)
Light Cone Consistency: Toward a Unified Theory of Consistency in Message-Passing Systems
by: Landers, Rob, et al.
Published: (2026)
by: Landers, Rob, et al.
Published: (2026)
Pioplat: A Scalable, Low-Cost Framework for Latency Reduction in Ethereum Blockchain
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
by: Talluri, Sacheendra, et al.
Published: (2025)
by: Talluri, Sacheendra, et al.
Published: (2025)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
by: Li, Zhifei, et al.
Published: (2026)
by: Li, Zhifei, et al.
Published: (2026)
Distributed Recoverable Sketches (Extended Version)
by: Cohen, Diana, et al.
Published: (2025)
by: Cohen, Diana, et al.
Published: (2025)
Intersections of Web3 and AI -- View in 2024
by: Hyland-Wood, David, et al.
Published: (2024)
by: Hyland-Wood, David, et al.
Published: (2024)
Artifact Evaluation for Distributed Systems: Current Practices and Beyond
by: Sedghpour, Mohammad Reza Saleh, et al.
Published: (2024)
by: Sedghpour, Mohammad Reza Saleh, et al.
Published: (2024)
NotebookOS: A Replicated Notebook Platform for Interactive Training with On-Demand GPUs
by: Carver, Benjamin, et al.
Published: (2025)
by: Carver, Benjamin, et al.
Published: (2025)
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
Data Race Satisfiability on Array Elements
by: Shim, Junhyung, et al.
Published: (2025)
by: Shim, Junhyung, et al.
Published: (2025)
Dodoor: Efficient Randomized Decentralized Scheduling with Load Caching for Heterogeneous Tasks and Clusters
by: Da, Wei, et al.
Published: (2025)
by: Da, Wei, et al.
Published: (2025)
Generic Multicast (Extended Version)
by: Bolina, José Augusto, et al.
Published: (2024)
by: Bolina, José Augusto, et al.
Published: (2024)
Play like a Vertex: A Stackelberg Game Approach for Streaming Graph Partitioning
by: Ding, Zezhong, et al.
Published: (2024)
by: Ding, Zezhong, et al.
Published: (2024)
Serverless GPU Architecture for Enterprise HR Analytics: A Production-Scale BDaaS Implementation
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
Shaved Ice: Optimal Compute Resource Commitments for Dynamic Multi-Cloud Workloads
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
OPTIMUMP2P: Fast and Reliable Gossiping in P2P Networks
by: Nicolaou, Nicolas, et al.
Published: (2025)
by: Nicolaou, Nicolas, et al.
Published: (2025)
FCDP: Fully Cached Data Parallel for Communication-Avoiding Large-Scale Training
by: Park, Gyeongseo, et al.
Published: (2026)
by: Park, Gyeongseo, et al.
Published: (2026)
Similar Items
-
dpBento: Benchmarking DPUs for Data Processing
by: Hu, Jiasheng, et al.
Published: (2025) -
Theseus: A Distributed and Scalable GPU-Accelerated Query Processing Platform Optimized for Efficient Data Movement
by: Aramburú, Felipe, et al.
Published: (2025) -
DDS: DPU-optimized Disaggregated Storage [Extended Report]
by: Zhang, Qizhen, et al.
Published: (2024) -
Fifty Years of Transaction Processing Research (extended)
by: Bernstein, Philip A.
Published: (2026) -
Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)