Ksurf-Drone: Attention Kalman Filter for Contextual Bandit Optimization in Cloud Resource Allocation
Fuente:
arXiv
Saved in:
| Main Authors: | Dang'ana, Michael, Zhang, Yuqiu, Jacobsen, Hans-Arno |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ksurf: Attention Kalman Filter and Principal Component Analysis for Prediction under Highly Variable Cloud Workloads
by: Dang'ana, Michael, et al.
Published: (2024)
by: Dang'ana, Michael, et al.
Published: (2024)
FractalSortCPU: Bandwidth-Efficient Compressed Radix Sort on CPU
by: Dang'ana, Michael
Published: (2026)
by: Dang'ana, Michael
Published: (2026)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
by: Erben, Alexander, et al.
Published: (2023)
by: Erben, Alexander, et al.
Published: (2023)
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
by: Woisetschläger, Herbert, et al.
Published: (2023)
by: Woisetschläger, Herbert, et al.
Published: (2023)
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
by: Xu, Youxuan, et al.
Published: (2025)
by: Xu, Youxuan, et al.
Published: (2025)
Scalability Evaluation of HPC Multi-GPU Training for ECG-based LLMs
by: Mileski, Dimitar, et al.
Published: (2025)
by: Mileski, Dimitar, et al.
Published: (2025)
Federated Fine-Tuning of LLMs on the Very Edge: The Good, the Bad, the Ugly
by: Woisetschläger, Herbert, et al.
Published: (2023)
by: Woisetschläger, Herbert, et al.
Published: (2023)
Cabinet: Dynamically Weighted Consensus Made Fast
by: Zhang, Gengrui, et al.
Published: (2025)
by: Zhang, Gengrui, et al.
Published: (2025)
Efficient and Reuseable Cloud Configuration Search Using Discovery Spaces
by: Johnston, Michael, et al.
Published: (2025)
by: Johnston, Michael, et al.
Published: (2025)
Knowledge Graphs-Driven Intelligence for Distributed Decision Systems
by: Napoli, Rosario, et al.
Published: (2026)
by: Napoli, Rosario, et al.
Published: (2026)
Experimentally Evaluating the Resource Efficiency of Big Data Autoscaling
by: Will, Jonathan, et al.
Published: (2025)
by: Will, Jonathan, et al.
Published: (2025)
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
MAS-Attention: Memory-Aware Stream Processing for Attention Acceleration on Resource-Constrained Edge Devices
by: Shakerdargah, Mohammadali, et al.
Published: (2024)
by: Shakerdargah, Mohammadali, et al.
Published: (2024)
Optimization of a Radiofrequency Ablation FEM Application Using Parallel Sparse Solvers
by: Miletto, Marcelo Cogo, et al.
Published: (2024)
by: Miletto, Marcelo Cogo, et al.
Published: (2024)
AutoDDL: Automatic Distributed Deep Learning with Near-Optimal Bandwidth Cost
by: Chen, Jinfan, et al.
Published: (2023)
by: Chen, Jinfan, et al.
Published: (2023)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
by: Murimi, Almond Kiruthu
Published: (2025)
by: Murimi, Almond Kiruthu
Published: (2025)
Cloud Resource Allocation with Convex Optimization
by: Boghani, Shayan, et al.
Published: (2025)
by: Boghani, Shayan, et al.
Published: (2025)
POD-Attention: Unlocking Full Prefill-Decode Overlap for Faster LLM Inference
by: Kamath, Aditya K, et al.
Published: (2024)
by: Kamath, Aditya K, et al.
Published: (2024)
Efficient Construction of Large Search Spaces for Auto-Tuning
by: Willemsen, Floris-Jan, et al.
Published: (2025)
by: Willemsen, Floris-Jan, et al.
Published: (2025)
A Survey on Efficient Federated Learning Methods for Foundation Model Training
by: Woisetschläger, Herbert, et al.
Published: (2024)
by: Woisetschläger, Herbert, et al.
Published: (2024)
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
by: Talluri, Sacheendra, et al.
Published: (2025)
by: Talluri, Sacheendra, et al.
Published: (2025)
TokenCake: A KV-Cache-centric Serving Framework for LLM-based Multi-Agent Applications
by: Bian, Zhuohang, et al.
Published: (2025)
by: Bian, Zhuohang, et al.
Published: (2025)
Epoch-based Optimistic Concurrency Control in Geo-replicated Databases
by: Mao, Yunhao, et al.
Published: (2026)
by: Mao, Yunhao, et al.
Published: (2026)
Accelerating Causal Algorithms for Industrial-scale Data: A Distributed Computing Approach with Ray Framework
by: Verma, Vishal, et al.
Published: (2024)
by: Verma, Vishal, et al.
Published: (2024)
AI/ML Model Cards in Edge AI Cyberinfrastructure: towards Agentic AI
by: Plale, Beth, et al.
Published: (2025)
by: Plale, Beth, et al.
Published: (2025)
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
Shaved Ice: Optimal Compute Resource Commitments for Dynamic Multi-Cloud Workloads
by: Stokely, Murray, et al.
Published: (2025)
by: Stokely, Murray, et al.
Published: (2025)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
by: Zhang, Guilin, et al.
Published: (2025)
by: Zhang, Guilin, et al.
Published: (2025)
Scaling Point-based Differentiable Rendering for Large-scale Reconstruction
by: Zhao, Hexu, et al.
Published: (2025)
by: Zhao, Hexu, et al.
Published: (2025)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
by: Polyakov, Igor, et al.
Published: (2025)
by: Polyakov, Igor, et al.
Published: (2025)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
by: Bhattacharya, Siddhartha, et al.
Published: (2024)
by: Bhattacharya, Siddhartha, et al.
Published: (2024)
Feature-Aware Task-to-Core Allocation in Embedded Multi-core Platforms via Statistical Learning
by: Pivezhandi, Mohammad, et al.
Published: (2025)
by: Pivezhandi, Mohammad, et al.
Published: (2025)
ATTNChecker: Highly-Optimized Fault Tolerant Attention for Large Language Model Training
by: Liang, Yuhang, et al.
Published: (2024)
by: Liang, Yuhang, et al.
Published: (2024)
Optimizing Foundation Model Inference on a Many-tiny-core Open-source RISC-V Platform
by: Potocnik, Viviane, et al.
Published: (2024)
by: Potocnik, Viviane, et al.
Published: (2024)
How Does Stake Distribution Influence Consensus? Analyzing Blockchain Decentralization
by: Motepalli, Shashank, et al.
Published: (2023)
by: Motepalli, Shashank, et al.
Published: (2023)
3D Point Cloud Object Detection on Edge Devices for Split Computing
by: Noguchi, Taisuke, et al.
Published: (2025)
by: Noguchi, Taisuke, et al.
Published: (2025)
LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management
by: Xiong, Yi, et al.
Published: (2024)
by: Xiong, Yi, et al.
Published: (2024)
Limitless FaaS: Overcoming serverless functions execution time limits with invoke driven architecture and memory checkpoints
by: Andraca, Rodrigo Landa, et al.
Published: (2024)
by: Andraca, Rodrigo Landa, et al.
Published: (2024)
Architecture-Aware LLM Inference Optimization on AMD Instinct GPUs: A Comprehensive Benchmark and Deployment Study
by: Georgiou, Athos
Published: (2026)
by: Georgiou, Athos
Published: (2026)
Parallelization Strategies for Dense LLM Deployment: Navigating Through Application-Specific Tradeoffs and Bottlenecks
by: Topcu, Burak, et al.
Published: (2026)
by: Topcu, Burak, et al.
Published: (2026)
Similar Items
-
Ksurf: Attention Kalman Filter and Principal Component Analysis for Prediction under Highly Variable Cloud Workloads
by: Dang'ana, Michael, et al.
Published: (2024) -
FractalSortCPU: Bandwidth-Efficient Compressed Radix Sort on CPU
by: Dang'ana, Michael
Published: (2026) -
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
by: Erben, Alexander, et al.
Published: (2023) -
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
by: Woisetschläger, Herbert, et al.
Published: (2023) -
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
by: Xu, Youxuan, et al.
Published: (2025)