Offloading Artificial Intelligence Workloads across the Computing Continuum by means of Active Storage Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Barceló, Alex, Ordoñez, Sebastián A. Cajas, Samanta, Jaydeep, Suárez-Cetrulo, Andrés L., Ghosh, Romila, Carbajo, Ricardo Simón, Queralt, Anna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OASIS: Object-based Analytics Storage for Intelligent SQL Query Offloading in Scientific Tabular Workloads
von: Hwang, Soon, et al.
Veröffentlicht: (2025)
von: Hwang, Soon, et al.
Veröffentlicht: (2025)
OffloadFS: Leveraging Disaggregated Storage for Computation Offloading
von: Moon, Sungho, et al.
Veröffentlicht: (2026)
von: Moon, Sungho, et al.
Veröffentlicht: (2026)
Workload Distribution with Rateless Encoding: A Low-Latency Computation Offloading Method within Edge Networks
von: Guo, Zhongfu, et al.
Veröffentlicht: (2023)
von: Guo, Zhongfu, et al.
Veröffentlicht: (2023)
Workflow-Driven Modeling for the Compute Continuum: An Optimization Approach to Automated System and Workload Scheduling
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025)
Workload Intelligence: Punching Holes Through the Cloud Abstraction
von: Huang, Lexiang, et al.
Veröffentlicht: (2024)
von: Huang, Lexiang, et al.
Veröffentlicht: (2024)
Saarthi: An End-to-End Intelligent Platform for Optimising Distributed Serverless Workloads
von: Agarwal, Siddharth, et al.
Veröffentlicht: (2025)
von: Agarwal, Siddharth, et al.
Veröffentlicht: (2025)
GriNNder: Breaking the Memory Capacity Wall in Full-Graph GNN Training with Storage Offloading
von: Song, Jaeyong, et al.
Veröffentlicht: (2026)
von: Song, Jaeyong, et al.
Veröffentlicht: (2026)
Collaborative Resource Management and Workloads Scheduling in Cloud-Assisted Mobile Edge Computing across Timescales
von: Tang, Lujie, et al.
Veröffentlicht: (2024)
von: Tang, Lujie, et al.
Veröffentlicht: (2024)
Proposal of Automatic Offloading Method in Mixed Offloading Destination Environment
von: Yamato, Yoji
Veröffentlicht: (2020)
von: Yamato, Yoji
Veröffentlicht: (2020)
DALI: A Workload-Aware Offloading Framework for Efficient MoE Inference on Local PCs
von: Zhu, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhu, Zeyu, et al.
Veröffentlicht: (2026)
To Offload or Not To Offload: Model-driven Comparison of Edge-native and On-device Processing In the Era of Accelerators
von: Ng, Nathan, et al.
Veröffentlicht: (2025)
von: Ng, Nathan, et al.
Veröffentlicht: (2025)
Agentic AI Workload Characteristics
von: Yuan, Yichao, et al.
Veröffentlicht: (2026)
von: Yuan, Yichao, et al.
Veröffentlicht: (2026)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
von: Jain, Kunal, et al.
Veröffentlicht: (2024)
von: Jain, Kunal, et al.
Veröffentlicht: (2024)
Distributed Intelligence in the Computing Continuum with Active Inference
von: Pujol, Victor Casamayor, et al.
Veröffentlicht: (2025)
von: Pujol, Victor Casamayor, et al.
Veröffentlicht: (2025)
Generative Artificial Intelligence Reproducibility and Consensus
von: Kim, Edward, et al.
Veröffentlicht: (2023)
von: Kim, Edward, et al.
Veröffentlicht: (2023)
Plug & Offload: Transparently Offloading TCP Stack onto Off-path SmartNIC with PnO-TCP
von: Nan, Hailong, et al.
Veröffentlicht: (2025)
von: Nan, Hailong, et al.
Veröffentlicht: (2025)
Resource Slicing through Intelligent Orchestration of Energy-aware IoT services in Edge-Cloud Continuum
von: Shahid, Hafiz Faheem, et al.
Veröffentlicht: (2024)
von: Shahid, Hafiz Faheem, et al.
Veröffentlicht: (2024)
ContinuumConductor : Decentralized Process Mining on the Edge-Cloud Continuum
von: Reiter, Hendrik, et al.
Veröffentlicht: (2025)
von: Reiter, Hendrik, et al.
Veröffentlicht: (2025)
Cosmos: A Cost Model for Serverless Workflows in the 3D Compute Continuum
von: Marcelino, Cynthia, et al.
Veröffentlicht: (2025)
von: Marcelino, Cynthia, et al.
Veröffentlicht: (2025)
A Survey of Computation Offloading with Task Types
von: Zhang, Siqi, et al.
Veröffentlicht: (2023)
von: Zhang, Siqi, et al.
Veröffentlicht: (2023)
CRIUgpu: Transparent Checkpointing of GPU-Accelerated Workloads
von: Stoyanov, Radostin, et al.
Veröffentlicht: (2025)
von: Stoyanov, Radostin, et al.
Veröffentlicht: (2025)
AI Surrogate Model for Distributed Computing Workloads
von: Park, David K., et al.
Veröffentlicht: (2024)
von: Park, David K., et al.
Veröffentlicht: (2024)
Accelerating Compound LLM Training Workloads with Maestro
von: Yuan, Xiulong, et al.
Veröffentlicht: (2026)
von: Yuan, Xiulong, et al.
Veröffentlicht: (2026)
Optimal Workload Placement on Multi-Instance GPUs
von: Turkkan, Bekir, et al.
Veröffentlicht: (2024)
von: Turkkan, Bekir, et al.
Veröffentlicht: (2024)
YAIFS: Yet (not) Another Intelligent Fog Simulator: A Framework for Agent-Driven Computing Continuum Modeling & Simulation
von: Lera, Isaac, et al.
Veröffentlicht: (2026)
von: Lera, Isaac, et al.
Veröffentlicht: (2026)
Union: An Automatic Workload Manager for Accelerating Network Simulation
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Crossword: Adaptive Consensus for Dynamic Data-Heavy Workloads
von: Hu, Guanzhou, et al.
Veröffentlicht: (2025)
von: Hu, Guanzhou, et al.
Veröffentlicht: (2025)
Distributed Load Balancing with Workload-Dependent Service Rates
von: Zhang, Wenxin, et al.
Veröffentlicht: (2024)
von: Zhang, Wenxin, et al.
Veröffentlicht: (2024)
Eventually-Consistent Federated Scheduling for Data Center Workloads
von: Thiyyakat, Meghana, et al.
Veröffentlicht: (2023)
von: Thiyyakat, Meghana, et al.
Veröffentlicht: (2023)
Kub: Enabling Elastic HPC Workloads on Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
Data Management System Analysis for Distributed Computing Workloads
von: Hsu, Kuan-Chieh, et al.
Veröffentlicht: (2025)
von: Hsu, Kuan-Chieh, et al.
Veröffentlicht: (2025)
SYMPHONY: Improving Memory Management for LLM Inference Workloads
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
Towards Cloud Efficiency with Large-scale Workload Characterization
von: Parayil, Anjaly, et al.
Veröffentlicht: (2024)
von: Parayil, Anjaly, et al.
Veröffentlicht: (2024)
IntentContinuum: Using LLMs to Support Intent-Based Computing Across the Compute Continuum
von: Akbari, Negin, et al.
Veröffentlicht: (2025)
von: Akbari, Negin, et al.
Veröffentlicht: (2025)
Communication Offloading on SmartNIC DPUs: A Quantitative Approach
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2026)
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2026)
Joint Optimization of Offloading, Batching and DVFS for Multiuser Co-Inference
von: Xu, Yaodan, et al.
Veröffentlicht: (2025)
von: Xu, Yaodan, et al.
Veröffentlicht: (2025)
Network-Offloaded Bandwidth-Optimal Broadcast and Allgather for Distributed AI
von: Khalilov, Mikhail, et al.
Veröffentlicht: (2024)
von: Khalilov, Mikhail, et al.
Veröffentlicht: (2024)
Static Generation of Efficient OpenMP Offload Data Mappings
von: Marzen, Luke, et al.
Veröffentlicht: (2024)
von: Marzen, Luke, et al.
Veröffentlicht: (2024)
TURNIP: A "Nondeterministic" GPU Runtime with CPU RAM Offload
von: Ding, Zhimin, et al.
Veröffentlicht: (2024)
von: Ding, Zhimin, et al.
Veröffentlicht: (2024)
Egret: Reinforcement Mechanism for Sequential Computation Offloading in Edge Computing
von: Peng, Haosong, et al.
Veröffentlicht: (2024)
von: Peng, Haosong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
OASIS: Object-based Analytics Storage for Intelligent SQL Query Offloading in Scientific Tabular Workloads
von: Hwang, Soon, et al.
Veröffentlicht: (2025) -
OffloadFS: Leveraging Disaggregated Storage for Computation Offloading
von: Moon, Sungho, et al.
Veröffentlicht: (2026) -
Workload Distribution with Rateless Encoding: A Low-Latency Computation Offloading Method within Edge Networks
von: Guo, Zhongfu, et al.
Veröffentlicht: (2023) -
Workflow-Driven Modeling for the Compute Continuum: An Optimization Approach to Automated System and Workload Scheduling
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2025) -
Workload Intelligence: Punching Holes Through the Cloud Abstraction
von: Huang, Lexiang, et al.
Veröffentlicht: (2024)