Multi-Dimensional Autoscaling of Stream Processing Services on Edge Devices
Fuente:
arXiv
Saved in:
| Main Authors: | Sedlak, Boris, Raith, Philipp, Morichetta, Andrea, Pujol, Víctor Casamayor, Dustdar, Schahram |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Insights into Agentic Optimization of Pervasive Stream Processing Services
by: Sedlak, Boris, et al.
Published: (2026)
by: Sedlak, Boris, et al.
Published: (2026)
Adaptive Stream Processing on Edge Devices through Active Inference
by: Sedlak, Boris, et al.
Published: (2024)
by: Sedlak, Boris, et al.
Published: (2024)
Multi-dimensional Autoscaling of Processing Services: A Comparison of Agent-based Methods
by: Sedlak, Boris, et al.
Published: (2025)
by: Sedlak, Boris, et al.
Published: (2025)
Equilibrium in the Computing Continuum through Active Inference
by: Sedlak, Boris, et al.
Published: (2023)
by: Sedlak, Boris, et al.
Published: (2023)
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services
by: Wang, Zihang, et al.
Published: (2026)
by: Wang, Zihang, et al.
Published: (2026)
Towards Multi-dimensional Elasticity for Pervasive Stream Processing Services
by: Sedlak, Boris, et al.
Published: (2025)
by: Sedlak, Boris, et al.
Published: (2025)
Formal and Empirical Study of Metadata-Based Profiling for Resource Management in the Computing Continuum
by: Morichetta, Andrea, et al.
Published: (2025)
by: Morichetta, Andrea, et al.
Published: (2025)
Benchmarking Dynamic SLO Compliance in Distributed Computing Continuum Systems
by: Lapkovskis, Alfreds, et al.
Published: (2025)
by: Lapkovskis, Alfreds, et al.
Published: (2025)
FrankenSplit: Efficient Neural Feature Compression with Shallow Variational Bottleneck Injection for Mobile Edge Computing
by: Furutanpey, Alireza, et al.
Published: (2023)
by: Furutanpey, Alireza, et al.
Published: (2023)
Adaptive Active Inference Agents for Heterogeneous and Lifelong Federated Learning
by: Danilenka, Anastasiya, et al.
Published: (2024)
by: Danilenka, Anastasiya, et al.
Published: (2024)
Distributed Intelligence in the Computing Continuum with Active Inference
by: Pujol, Victor Casamayor, et al.
Published: (2025)
by: Pujol, Victor Casamayor, et al.
Published: (2025)
SLO-Aware Task Offloading within Collaborative Vehicle Platoons
by: Sedlak, Boris, et al.
Published: (2024)
by: Sedlak, Boris, et al.
Published: (2024)
Service Orchestration in the Computing Continuum: Structural Challenges and Vision
by: Sedlak, Boris, et al.
Published: (2026)
by: Sedlak, Boris, et al.
Published: (2026)
RainCloud: Decentralized Coordination and Communication in Heterogeneous IoT Swarms
by: Loisel, Filip, et al.
Published: (2024)
by: Loisel, Filip, et al.
Published: (2024)
Leveraging Neural Graph Compilers in Machine Learning Research for Edge-Cloud Systems
by: Furutanpey, Alireza, et al.
Published: (2025)
by: Furutanpey, Alireza, et al.
Published: (2025)
EdgeProfiler: A Fast Profiling Framework for Lightweight LLMs on Edge Using Analytical Model
by: Pinnock, Alyssa, et al.
Published: (2025)
by: Pinnock, Alyssa, et al.
Published: (2025)
Rethinking Inference Placement for Deep Learning across Edge and Cloud Platforms: A Multi-Objective Optimization Perspective and Future Directions
by: Zhang, Zongshun, et al.
Published: (2025)
by: Zhang, Zongshun, et al.
Published: (2025)
BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services
by: Lackinger, Anna, et al.
Published: (2025)
by: Lackinger, Anna, et al.
Published: (2025)
DUAL-BLADE: Dual-Path NVMe-Direct KV-Cache Offloading for Edge LLM Inference
by: Jeong, Bodon, et al.
Published: (2026)
by: Jeong, Bodon, et al.
Published: (2026)
CASA: A Framework for SLO and Carbon-Aware Autoscaling and Scheduling in Serverless Cloud Computing
by: Qi, S., et al.
Published: (2024)
by: Qi, S., et al.
Published: (2024)
Collaborative Processing for Multi-Tenant Inference on Memory-Constrained Edge TPUs
by: Ng, Nathan, et al.
Published: (2026)
by: Ng, Nathan, et al.
Published: (2026)
When AI Bends Metal: AI-Assisted Optimization of Design Parameters in Sheet Metal Forming
by: Tarraf, Ahmad, et al.
Published: (2025)
by: Tarraf, Ahmad, et al.
Published: (2025)
Performance Optimization in Stream Processing Systems: Experiment-Driven Configuration Tuning for Kafka Streams
by: Chen, David, et al.
Published: (2026)
by: Chen, David, et al.
Published: (2026)
Towards a Proactive Autoscaling Framework for Data Stream Processing at the Edge using GRU and Transfer Learning
by: Armah, Eugene, et al.
Published: (2025)
by: Armah, Eugene, et al.
Published: (2025)
CommunityAI: Towards Community-based Federated Learning
by: Murturi, Ilir, et al.
Published: (2023)
by: Murturi, Ilir, et al.
Published: (2023)
Performance Implications of Multi-Chiplet Neural Processing Units on Autonomous Driving Perception
by: Odema, Mohanad, et al.
Published: (2024)
by: Odema, Mohanad, et al.
Published: (2024)
SProBench: Stream Processing Benchmark for High Performance Computing Infrastructure
by: Kulkarni, Apurv Deepak, et al.
Published: (2025)
by: Kulkarni, Apurv Deepak, et al.
Published: (2025)
KPI2KVI: A Multi Agent Workflow for Calculating Key Value Indicators from Service Descriptions
by: Shokrnezhad, Masoud, et al.
Published: (2026)
by: Shokrnezhad, Masoud, et al.
Published: (2026)
Can Large Language Models Predict Parallel Code Performance?
by: Bolet, Gregory, et al.
Published: (2025)
by: Bolet, Gregory, et al.
Published: (2025)
oneDAL Optimization for ARM Scalable Vector Extension: Maximizing Efficiency for High-Performance Data Science
by: Sharma, Chandan, et al.
Published: (2025)
by: Sharma, Chandan, et al.
Published: (2025)
AGOCS -- Accurate Google Cloud Simulator Framework
by: Sliwko, Leszek, et al.
Published: (2025)
by: Sliwko, Leszek, et al.
Published: (2025)
Standardized Methods and Recommendations for Green Federated Learning
by: Tapp, Austin, et al.
Published: (2026)
by: Tapp, Austin, et al.
Published: (2026)
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
by: Zhu, Jianwei, et al.
Published: (2024)
by: Zhu, Jianwei, et al.
Published: (2024)
GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving
by: Jayakody, Shakya, et al.
Published: (2026)
by: Jayakody, Shakya, et al.
Published: (2026)
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
by: Barron, Ryan, et al.
Published: (2024)
by: Barron, Ryan, et al.
Published: (2024)
CoServe: Efficient Collaboration-of-Experts (CoE) Model Inference with Limited Memory
by: Suo, Jiashun, et al.
Published: (2025)
by: Suo, Jiashun, et al.
Published: (2025)
Green or Fast? Learning to Balance Cold Starts and Idle Carbon in Serverless Computing
by: Sun, Bowen, et al.
Published: (2026)
by: Sun, Bowen, et al.
Published: (2026)
ExpertFlow: Adaptive Expert Scheduling and Memory Coordination for Efficient MoE Inference
by: Shen, Zixu, et al.
Published: (2025)
by: Shen, Zixu, et al.
Published: (2025)
Counting Without Running: Evaluating LLMs' Reasoning About Code Complexity
by: Bolet, Gregory, et al.
Published: (2025)
by: Bolet, Gregory, et al.
Published: (2025)
Shared Memory-contention-aware Concurrent DNN Execution for Diversely Heterogeneous System-on-Chips
by: Dagli, Ismet, et al.
Published: (2023)
by: Dagli, Ismet, et al.
Published: (2023)
Similar Items
-
Visual Insights into Agentic Optimization of Pervasive Stream Processing Services
by: Sedlak, Boris, et al.
Published: (2026) -
Adaptive Stream Processing on Edge Devices through Active Inference
by: Sedlak, Boris, et al.
Published: (2024) -
Multi-dimensional Autoscaling of Processing Services: A Comparison of Agent-based Methods
by: Sedlak, Boris, et al.
Published: (2025) -
Equilibrium in the Computing Continuum through Active Inference
by: Sedlak, Boris, et al.
Published: (2023) -
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services
by: Wang, Zihang, et al.
Published: (2026)