SLO-Aware Task Offloading within Collaborative Vehicle Platoons
Fuente:
arXiv
Salvato in:
| Autori principali: | Sedlak, Boris, Morichetta, Andrea, Wang, Yuhao, Fei, Yang, Wang, Liang, Dustdar, Schahram, Qu, Xiaobo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Adaptive Stream Processing on Edge Devices through Active Inference
di: Sedlak, Boris, et al.
Pubblicazione: (2024)
di: Sedlak, Boris, et al.
Pubblicazione: (2024)
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services
di: Wang, Zihang, et al.
Pubblicazione: (2026)
di: Wang, Zihang, et al.
Pubblicazione: (2026)
Multi-Dimensional Autoscaling of Stream Processing Services on Edge Devices
di: Sedlak, Boris, et al.
Pubblicazione: (2025)
di: Sedlak, Boris, et al.
Pubblicazione: (2025)
Formal and Empirical Study of Metadata-Based Profiling for Resource Management in the Computing Continuum
di: Morichetta, Andrea, et al.
Pubblicazione: (2025)
di: Morichetta, Andrea, et al.
Pubblicazione: (2025)
RainCloud: Decentralized Coordination and Communication in Heterogeneous IoT Swarms
di: Loisel, Filip, et al.
Pubblicazione: (2024)
di: Loisel, Filip, et al.
Pubblicazione: (2024)
Visual Insights into Agentic Optimization of Pervasive Stream Processing Services
di: Sedlak, Boris, et al.
Pubblicazione: (2026)
di: Sedlak, Boris, et al.
Pubblicazione: (2026)
BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services
di: Lackinger, Anna, et al.
Pubblicazione: (2025)
di: Lackinger, Anna, et al.
Pubblicazione: (2025)
Benchmarking Dynamic SLO Compliance in Distributed Computing Continuum Systems
di: Lapkovskis, Alfreds, et al.
Pubblicazione: (2025)
di: Lapkovskis, Alfreds, et al.
Pubblicazione: (2025)
Collaborative Inference in DNN-based Satellite Systems with Dynamic Task Streams
di: Guan, Jinglong, et al.
Pubblicazione: (2023)
di: Guan, Jinglong, et al.
Pubblicazione: (2023)
Distributed Intelligence in the Computing Continuum with Active Inference
di: Pujol, Victor Casamayor, et al.
Pubblicazione: (2025)
di: Pujol, Victor Casamayor, et al.
Pubblicazione: (2025)
Memory Offloading for Large Language Model Inference with Latency SLO Guarantees
di: Ma, Chenxiang, et al.
Pubblicazione: (2025)
di: Ma, Chenxiang, et al.
Pubblicazione: (2025)
A Conflict-Aware Resource Management Framework for the Computing Continuum
di: Popescu-Vifor, Vlad, et al.
Pubblicazione: (2025)
di: Popescu-Vifor, Vlad, et al.
Pubblicazione: (2025)
Multi-dimensional Autoscaling of Processing Services: A Comparison of Agent-based Methods
di: Sedlak, Boris, et al.
Pubblicazione: (2025)
di: Sedlak, Boris, et al.
Pubblicazione: (2025)
QEdgeProxy: QoS-Aware Load Balancing for IoT Services in the Computing Continuum
di: Čilić, Ivan, et al.
Pubblicazione: (2024)
di: Čilić, Ivan, et al.
Pubblicazione: (2024)
Equilibrium in the Computing Continuum through Active Inference
di: Sedlak, Boris, et al.
Pubblicazione: (2023)
di: Sedlak, Boris, et al.
Pubblicazione: (2023)
Orchestrating Serverless Applications in the Edge Cloud Space Continuum: What Breaks and What is Next?
di: Malazi, Hadi Tabatabaee, et al.
Pubblicazione: (2026)
di: Malazi, Hadi Tabatabaee, et al.
Pubblicazione: (2026)
ClusterLess: Deadline-Aware Serverless Workflow Orchestration on Federated Edge Clusters
di: Farahani, Reza, et al.
Pubblicazione: (2026)
di: Farahani, Reza, et al.
Pubblicazione: (2026)
Collaborative Inference for Large Models with Task Offloading and Early Exiting
di: Xie, Zuan, et al.
Pubblicazione: (2024)
di: Xie, Zuan, et al.
Pubblicazione: (2024)
Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference
di: Wang, Qipeng
Pubblicazione: (2026)
di: Wang, Qipeng
Pubblicazione: (2026)
Inference Load-Aware Orchestration for Hierarchical Federated Learning
di: Lackinger, Anna, et al.
Pubblicazione: (2024)
di: Lackinger, Anna, et al.
Pubblicazione: (2024)
LLM-Enhanced Deep Reinforcement Learning for Task Offloading in Collaborative Edge Computing
di: Guo, Hao, et al.
Pubblicazione: (2026)
di: Guo, Hao, et al.
Pubblicazione: (2026)
Synergizing Monetization, Orchestration, and Semantics in Computing Continuum
di: Dehury, Chinmaya Kumar, et al.
Pubblicazione: (2025)
di: Dehury, Chinmaya Kumar, et al.
Pubblicazione: (2025)
QoS-Aware Load Balancing in the Computing Continuum via Multi-Player Bandits
di: Čilić, Ivan, et al.
Pubblicazione: (2025)
di: Čilić, Ivan, et al.
Pubblicazione: (2025)
SLO-Aware Scheduling for Large Language Model Inferences
di: Huang, Jinqi, et al.
Pubblicazione: (2025)
di: Huang, Jinqi, et al.
Pubblicazione: (2025)
Collaborative Satellite Computing through Adaptive DNN Task Splitting and Offloading
di: Peng, Shifeng, et al.
Pubblicazione: (2024)
di: Peng, Shifeng, et al.
Pubblicazione: (2024)
Adaptive Active Inference Agents for Heterogeneous and Lifelong Federated Learning
di: Danilenka, Anastasiya, et al.
Pubblicazione: (2024)
di: Danilenka, Anastasiya, et al.
Pubblicazione: (2024)
Service Orchestration in the Computing Continuum: Structural Challenges and Vision
di: Sedlak, Boris, et al.
Pubblicazione: (2026)
di: Sedlak, Boris, et al.
Pubblicazione: (2026)
Towards Adaptive Asynchronous Federated Learning for Human Activity Recognition
di: Gajanin, Rastko, et al.
Pubblicazione: (2024)
di: Gajanin, Rastko, et al.
Pubblicazione: (2024)
Aladdin: Joint Placement and Scaling for SLO-Aware LLM Serving
di: Nie, Chengyi, et al.
Pubblicazione: (2024)
di: Nie, Chengyi, et al.
Pubblicazione: (2024)
A Multidimensional Elasticity Framework for Adaptive Data Analytics Management in the Computing Continuum
di: Laso, Sergio, et al.
Pubblicazione: (2025)
di: Laso, Sergio, et al.
Pubblicazione: (2025)
PromptTuner: SLO-Aware Elastic System for LLM Prompt Tuning
di: Gao, Wei, et al.
Pubblicazione: (2026)
di: Gao, Wei, et al.
Pubblicazione: (2026)
SCOOT: SLO-Oriented Performance Tuning for LLM Inference Engines
di: Cheng, Ke, et al.
Pubblicazione: (2024)
di: Cheng, Ke, et al.
Pubblicazione: (2024)
MSAO: Adaptive Modality Sparsity-Aware Offloading with Edge-Cloud Collaboration for Efficient Multimodal LLM Inference
di: Yang, Zheming, et al.
Pubblicazione: (2026)
di: Yang, Zheming, et al.
Pubblicazione: (2026)
A House United Within Itself: SLO-Awareness for On-Premises Containerized ML Inference Clusters via Faro
di: Jeon, Beomyeol, et al.
Pubblicazione: (2024)
di: Jeon, Beomyeol, et al.
Pubblicazione: (2024)
CommunityAI: Towards Community-based Federated Learning
di: Murturi, Ilir, et al.
Pubblicazione: (2023)
di: Murturi, Ilir, et al.
Pubblicazione: (2023)
HarmonyBatch: Batching multi-SLO DNN Inference with Heterogeneous Serverless Functions
di: Chen, Jiabin, et al.
Pubblicazione: (2024)
di: Chen, Jiabin, et al.
Pubblicazione: (2024)
MoA-Off: Adaptive Heterogeneous Modality-Aware Offloading with Edge-Cloud Collaboration for Efficient Multimodal LLM Inference
di: Yang, Zheming, et al.
Pubblicazione: (2025)
di: Yang, Zheming, et al.
Pubblicazione: (2025)
Hummingbird: SLO-Oriented GPU Preemption at Microsecond-scale
di: Hu, Tiancheng, et al.
Pubblicazione: (2026)
di: Hu, Tiancheng, et al.
Pubblicazione: (2026)
Distributed Massive MIMO-Aided Task Offloading in Satellite-Terrestrial Integrated Multi-Tier VEC Networks
di: Liu, Yixin, et al.
Pubblicazione: (2024)
di: Liu, Yixin, et al.
Pubblicazione: (2024)
Preemption Aware Task Scheduling for Priority and Deadline Constrained DNN Inference Task Offloading in Homogeneous Mobile-Edge Networks
di: Cotter, Jamie, et al.
Pubblicazione: (2025)
di: Cotter, Jamie, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Adaptive Stream Processing on Edge Devices through Active Inference
di: Sedlak, Boris, et al.
Pubblicazione: (2024) -
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services
di: Wang, Zihang, et al.
Pubblicazione: (2026) -
Multi-Dimensional Autoscaling of Stream Processing Services on Edge Devices
di: Sedlak, Boris, et al.
Pubblicazione: (2025) -
Formal and Empirical Study of Metadata-Based Profiling for Resource Management in the Computing Continuum
di: Morichetta, Andrea, et al.
Pubblicazione: (2025) -
RainCloud: Decentralized Coordination and Communication in Heterogeneous IoT Swarms
di: Loisel, Filip, et al.
Pubblicazione: (2024)