Joint Partitioning and Placement of Foundation Models for Real-Time Edge AI
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Djuhera, Aladin, Koch, Fernando, Binotto, Alecio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Intelligent Orchestration of Distributed Large Foundation Model Inference at the Edge
von: Koch, Fernando, et al.
Veröffentlicht: (2025)
von: Koch, Fernando, et al.
Veröffentlicht: (2025)
SplitLLM: Collaborative Inference of LLMs for Model Placement and Throughput Optimization
von: Mudvari, Akrit, et al.
Veröffentlicht: (2024)
von: Mudvari, Akrit, et al.
Veröffentlicht: (2024)
Joint Optimization of Training and Inference in Federated Edge Learning via Constrained Multi-Objective Deep Reinforcement Learning
von: Li, Zhen, et al.
Veröffentlicht: (2026)
von: Li, Zhen, et al.
Veröffentlicht: (2026)
Service Placement in Small Cell Networks Using Distributed Best Arm Identification in Linear Bandits
von: Yahya, Mariam, et al.
Veröffentlicht: (2025)
von: Yahya, Mariam, et al.
Veröffentlicht: (2025)
DeepEdge: A Deep Reinforcement Learning based Task Orchestrator for Edge Computing
von: Yamansavascilar, Baris, et al.
Veröffentlicht: (2021)
von: Yamansavascilar, Baris, et al.
Veröffentlicht: (2021)
Satellite Federated Fine-Tuning for Foundation Models in Space Computing Power Networks
von: Zhu, Yan, et al.
Veröffentlicht: (2025)
von: Zhu, Yan, et al.
Veröffentlicht: (2025)
Optimizing Server Placement for Vertical Federated Learning in Dynamic Edge/Fog Networks
von: Wang, Su, et al.
Veröffentlicht: (2026)
von: Wang, Su, et al.
Veröffentlicht: (2026)
Optimizing Edge Offloading Decisions for Object Detection
von: Qiu, Jiaming, et al.
Veröffentlicht: (2024)
von: Qiu, Jiaming, et al.
Veröffentlicht: (2024)
Split Learning in 6G Edge Networks
von: Lin, Zheng, et al.
Veröffentlicht: (2023)
von: Lin, Zheng, et al.
Veröffentlicht: (2023)
SlimCaching: Edge Caching of Mixture-of-Experts for Distributed Inference
von: Chen, Qian, et al.
Veröffentlicht: (2025)
von: Chen, Qian, et al.
Veröffentlicht: (2025)
A Learning-Based Caching Mechanism for Edge Content Delivery
von: Torabi, Hoda, et al.
Veröffentlicht: (2024)
von: Torabi, Hoda, et al.
Veröffentlicht: (2024)
ScaDLES: Scalable Deep Learning over Streaming data at the Edge
von: Tyagi, Sahil, et al.
Veröffentlicht: (2023)
von: Tyagi, Sahil, et al.
Veröffentlicht: (2023)
Device Sampling and Resource Optimization for Federated Learning in Cooperative Edge Networks
von: Wang, Su, et al.
Veröffentlicht: (2023)
von: Wang, Su, et al.
Veröffentlicht: (2023)
On-demand Quantization for Green Federated Generative Diffusion in Mobile Edge Networks
von: Lai, Bingkun, et al.
Veröffentlicht: (2024)
von: Lai, Bingkun, et al.
Veröffentlicht: (2024)
Solving AI Foundational Model Latency with Telco Infrastructure
von: Barros, Sebastian
Veröffentlicht: (2025)
von: Barros, Sebastian
Veröffentlicht: (2025)
DRL-Based Federated Self-Supervised Learning for Task Offloading and Resource Allocation in ISAC-Enabled Vehicle Edge Computing
von: Gu, Xueying, et al.
Veröffentlicht: (2024)
von: Gu, Xueying, et al.
Veröffentlicht: (2024)
EdgeLoc: A Communication-Adaptive Parallel System for Real-Time Localization in Infrastructure-Assisted Autonomous Driving
von: Liu, Boyi, et al.
Veröffentlicht: (2024)
von: Liu, Boyi, et al.
Veröffentlicht: (2024)
Implementation of Big AI Models for Wireless Networks with Collaborative Edge Computing
von: Zeng, Liekang, et al.
Veröffentlicht: (2024)
von: Zeng, Liekang, et al.
Veröffentlicht: (2024)
CRAFT: Latency and Cost-Aware Genetic-Based Framework for Node Placement in Edge-Fog Environments
von: Mahdizadeh, Soheil, et al.
Veröffentlicht: (2025)
von: Mahdizadeh, Soheil, et al.
Veröffentlicht: (2025)
The Internet of Things in the Era of Generative AI: Vision and Challenges
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Adaptive DNN Partitioning and Offloading in Heterogeneous Edge-Cloud Continuum
von: Deng, Akuen Akoi, et al.
Veröffentlicht: (2026)
von: Deng, Akuen Akoi, et al.
Veröffentlicht: (2026)
Placing Timely Refreshing Services at the Network Edge
von: Li, Xishuo, et al.
Veröffentlicht: (2024)
von: Li, Xishuo, et al.
Veröffentlicht: (2024)
Towards Timely Video Analytics Services at the Network Edge
von: Li, Xishuo, et al.
Veröffentlicht: (2024)
von: Li, Xishuo, et al.
Veröffentlicht: (2024)
FedAQ: Communication-Efficient Federated Edge Learning via Joint Uplink and Downlink Adaptive Quantization
von: Qu, Linping, et al.
Veröffentlicht: (2024)
von: Qu, Linping, et al.
Veröffentlicht: (2024)
Federated Continual Learning for Edge-AI: A Comprehensive Survey
von: Wang, Zi, et al.
Veröffentlicht: (2024)
von: Wang, Zi, et al.
Veröffentlicht: (2024)
Federated Learning framework for LoRaWAN-enabled IIoT communication: A case study
von: Sanchez, Oscar Torres, et al.
Veröffentlicht: (2024)
von: Sanchez, Oscar Torres, et al.
Veröffentlicht: (2024)
Reconfigurable Intelligent Surface Aided Vehicular Edge Computing: Joint Phase-shift Optimization and Multi-User Power Allocation
von: Qi, Kangwei, et al.
Veröffentlicht: (2024)
von: Qi, Kangwei, et al.
Veröffentlicht: (2024)
EMA: Efficient Model Adaptation for Learning-based Systems
von: Yu, Daiyang, et al.
Veröffentlicht: (2026)
von: Yu, Daiyang, et al.
Veröffentlicht: (2026)
On Efficiently Partitioning a Topic in Apache Kafka
von: Raptis, Theofanis P., et al.
Veröffentlicht: (2022)
von: Raptis, Theofanis P., et al.
Veröffentlicht: (2022)
Federated Split Learning with Model Pruning and Gradient Quantization in Wireless Networks
von: Zhang, Junhe, et al.
Veröffentlicht: (2024)
von: Zhang, Junhe, et al.
Veröffentlicht: (2024)
Diffusion Models on the Edge: Challenges, Optimizations, and Applications
von: Zheng, Dongqi
Veröffentlicht: (2025)
von: Zheng, Dongqi
Veröffentlicht: (2025)
Galaxy: A Resource-Efficient Collaborative Edge AI System for In-situ Transformer Inference
von: Ye, Shengyuan, et al.
Veröffentlicht: (2024)
von: Ye, Shengyuan, et al.
Veröffentlicht: (2024)
Efficient Fog Node Placement using Nature-Inspired Metaheuristic for IoT Applications
von: Naouri, Abdenacer, et al.
Veröffentlicht: (2023)
von: Naouri, Abdenacer, et al.
Veröffentlicht: (2023)
Dynamic Edge Server Selection in Time-Varying Environments: A Reliability-Aware Predictive Approach
von: Burbano, Jaime Sebastian, et al.
Veröffentlicht: (2025)
von: Burbano, Jaime Sebastian, et al.
Veröffentlicht: (2025)
LOAM: Low-latency Communication, Caching, and Computation Placement in Data-Intensive Computing Networks
von: Zhang, Jinkun, et al.
Veröffentlicht: (2024)
von: Zhang, Jinkun, et al.
Veröffentlicht: (2024)
QECO: A QoE-Oriented Computation Offloading Algorithm based on Deep Reinforcement Learning for Mobile Edge Computing
von: Rahmaty, Iman, et al.
Veröffentlicht: (2023)
von: Rahmaty, Iman, et al.
Veröffentlicht: (2023)
Satellite Edge Artificial Intelligence with Large Models: Architectures and Technologies
von: Shi, Yuanming, et al.
Veröffentlicht: (2025)
von: Shi, Yuanming, et al.
Veröffentlicht: (2025)
Early-Exit meets Model-Distributed Inference at Edge Networks
von: Colocrese, Marco, et al.
Veröffentlicht: (2024)
von: Colocrese, Marco, et al.
Veröffentlicht: (2024)
Urgent Edge Computing
von: Dazzi, Patrizio, et al.
Veröffentlicht: (2024)
von: Dazzi, Patrizio, et al.
Veröffentlicht: (2024)
Edge Graph Intelligence: Reciprocally Empowering Edge Networks with Graph Intelligence
von: Zeng, Liekang, et al.
Veröffentlicht: (2024)
von: Zeng, Liekang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Intelligent Orchestration of Distributed Large Foundation Model Inference at the Edge
von: Koch, Fernando, et al.
Veröffentlicht: (2025) -
SplitLLM: Collaborative Inference of LLMs for Model Placement and Throughput Optimization
von: Mudvari, Akrit, et al.
Veröffentlicht: (2024) -
Joint Optimization of Training and Inference in Federated Edge Learning via Constrained Multi-Objective Deep Reinforcement Learning
von: Li, Zhen, et al.
Veröffentlicht: (2026) -
Service Placement in Small Cell Networks Using Distributed Best Arm Identification in Linear Bandits
von: Yahya, Mariam, et al.
Veröffentlicht: (2025) -
DeepEdge: A Deep Reinforcement Learning based Task Orchestrator for Edge Computing
von: Yamansavascilar, Baris, et al.
Veröffentlicht: (2021)