Saved in:
| Main Authors: | Li, Muqing, Li, Ning, Yuan, Xin, Xu, Wenchao, Chen, Quan, Guo, Song, Zhang, Haijun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.09208 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The MoE-Empowered Edge LLMs Deployment: Architecture, Challenges, and Opportunities
by: Li, Ning, et al.
Published: (2025)
by: Li, Ning, et al.
Published: (2025)
MoE$^2$: Optimizing Collaborative Inference for Edge Large Language Models
by: Jin, Lyudong, et al.
Published: (2025)
by: Jin, Lyudong, et al.
Published: (2025)
SiftMoE: Similarity-Aware Energy-Efficient Expert Selection for Wireless Distributed MoE Inference
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
UAV-Assisted Cooperative Edge Inference for Low-Altitude Economy via MoE-based Hierarchical Deep Reinforcement Learning
by: Zhuang, Wenhao, et al.
Published: (2026)
by: Zhuang, Wenhao, et al.
Published: (2026)
Cluster Topology-Driven Placement of Experts Reduces Network Traffic in MoE Inference
by: Sivtsov, Danil, et al.
Published: (2025)
by: Sivtsov, Danil, et al.
Published: (2025)
A Dynamic Service Offloading Algorithm Based on Lyapunov Optimization in Edge Computing
by: Yuan, Peiyan, et al.
Published: (2025)
by: Yuan, Peiyan, et al.
Published: (2025)
SpaceMoE: Towards Orbital General Intelligence with Distributed Mixture-of-Experts Inference
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
Rotatable RIS-Assisted Edge Computing: Orientation, Task Offloading, and Resource Optimization
by: Li, Bin, et al.
Published: (2025)
by: Li, Bin, et al.
Published: (2025)
Modeling Edge-to-Cloud Offloading Workloads for Autonomous Vehicles
by: Li, Longkun, et al.
Published: (2026)
by: Li, Longkun, et al.
Published: (2026)
Online Collaborative Resource Allocation and Task Offloading for Multi-access Edge Computing
by: Sun, Geng, et al.
Published: (2025)
by: Sun, Geng, et al.
Published: (2025)
Cooperative Edge Caching with Large Language Model in Wireless Networks
by: Yang, Ning, et al.
Published: (2026)
by: Yang, Ning, et al.
Published: (2026)
On the Optimization of Model Aggregation for Federated Learning at the Network Edge
by: Li, Mengyao, et al.
Published: (2025)
by: Li, Mengyao, et al.
Published: (2025)
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading
by: Xu, Minrui, et al.
Published: (2025)
by: Xu, Minrui, et al.
Published: (2025)
HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts
by: Fang, Zihan, et al.
Published: (2026)
by: Fang, Zihan, et al.
Published: (2026)
Beyond the Edge: An Advanced Exploration of Reinforcement Learning for Mobile Edge Computing, its Applications, and Future Research Trajectories
by: Yang, Ning, et al.
Published: (2024)
by: Yang, Ning, et al.
Published: (2024)
MalMoE: Mixture-of-Experts Enhanced Encrypted Malicious Traffic Detection Under Graph Drift
by: Tan, Yunpeng, et al.
Published: (2026)
by: Tan, Yunpeng, et al.
Published: (2026)
Compact LLM Deployment and World Model Assisted Offloading in Mobile Edge Computing
by: Zhang, Ruichen, et al.
Published: (2026)
by: Zhang, Ruichen, et al.
Published: (2026)
QoE-Driven Multi-Task Offloading for Semantic-Aware Edge Computing Systems
by: Chen, Xuyang, et al.
Published: (2024)
by: Chen, Xuyang, et al.
Published: (2024)
TrafficMoE: Heterogeneity-aware Mixture of Experts for Encrypted Traffic Classification
by: He, Qing, et al.
Published: (2026)
by: He, Qing, et al.
Published: (2026)
Joint Optimization of Completion Ratio and Latency of Offloaded Tasks with Multiple Priority Levels in 5G Edge
by: Moshiri, Parisa Fard, et al.
Published: (2024)
by: Moshiri, Parisa Fard, et al.
Published: (2024)
Hierarchical Edge-Cloud Task Offloading in NTN for Remote Healthcare
by: Flores, Alejandro, et al.
Published: (2026)
by: Flores, Alejandro, et al.
Published: (2026)
RRTO: A High-Performance Transparent Offloading System for Model Inference in Mobile Edge Computing
by: Sun, Zekai, et al.
Published: (2025)
by: Sun, Zekai, et al.
Published: (2025)
Toward Resource-Efficient Collaboration of Large AI Models in Mobile Edge Networks
by: Li, Peichun, et al.
Published: (2026)
by: Li, Peichun, et al.
Published: (2026)
Hybrid Reinforcement Learning-based Sustainable Multi-User Computation Offloading for Mobile Edge-Quantum Computing
by: Xu, Minrui, et al.
Published: (2025)
by: Xu, Minrui, et al.
Published: (2025)
Deterministic Task Offloading and Resource Allocation in the IoT-Edge-Cloud Continuum
by: Aghababaiyan, Keyvan, et al.
Published: (2026)
by: Aghababaiyan, Keyvan, et al.
Published: (2026)
Autonomous Task Offloading of Vehicular Edge Computing with Parallel Computation Queues
by: Cho, Sungho, et al.
Published: (2025)
by: Cho, Sungho, et al.
Published: (2025)
Entropy-Aware Task Offloading in Mobile Edge Computing
by: Ardakani, Mohsen Sahraei, et al.
Published: (2026)
by: Ardakani, Mohsen Sahraei, et al.
Published: (2026)
A Centrality Approach to Select Offloading Data Aggregation Points in Vehicular Sensor Networks
by: Moura, Douglas, et al.
Published: (2024)
by: Moura, Douglas, et al.
Published: (2024)
A 5G-Edge Architecture for Computational Offloading of Computer Vision Applications
by: da Silva, Marcelo V. B., et al.
Published: (2025)
by: da Silva, Marcelo V. B., et al.
Published: (2025)
Scalable Deterministic Task Offloading and Resource Allocation in the IoT-Edge-Cloud Continuum
by: Aghababaiyan, Keyvan, et al.
Published: (2026)
by: Aghababaiyan, Keyvan, et al.
Published: (2026)
Multi-Agent RL-Based Industrial AIGC Service Offloading over Wireless Edge Networks
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
Edge Offloading in Smart Grid
by: Arcas, Gabriel Ioan, et al.
Published: (2024)
by: Arcas, Gabriel Ioan, et al.
Published: (2024)
SpaceMoE: Realizing Distributed Mixture-of-Experts Inference over Space Networks
by: Wang, Zhanwei, et al.
Published: (2026)
by: Wang, Zhanwei, et al.
Published: (2026)
Optimizing Edge Offloading Decisions for Object Detection
by: Qiu, Jiaming, et al.
Published: (2024)
by: Qiu, Jiaming, et al.
Published: (2024)
Zephyrus: Scaling Gateways Beyond the Petabit-Era with DPU-Augmented Hierarchical Co-Offloading
by: Xu, Yuemeng, et al.
Published: (2025)
by: Xu, Yuemeng, et al.
Published: (2025)
MEC Task Offloading in AIoT: A User-Centric DRL Model Splitting Inference Scheme
by: Li, Weixi, et al.
Published: (2025)
by: Li, Weixi, et al.
Published: (2025)
MoE-GPS: Guidlines for Prediction Strategy for Dynamic Expert Duplication in MoE Load Balancing
by: Ma, Haiyue, et al.
Published: (2025)
by: Ma, Haiyue, et al.
Published: (2025)
Joint Edge Server Deployment and Computation Offloading: A Multi-Timescale Stochastic Programming Framework
by: Liu, Huaizhe, et al.
Published: (2025)
by: Liu, Huaizhe, et al.
Published: (2025)
Renewables Power the Orbit? Achieving Sustainable Space Edge Computing via QoS-Aware Offloading
by: Fan, Xiaoyi, et al.
Published: (2026)
by: Fan, Xiaoyi, et al.
Published: (2026)
Meeting Deadlines in Motion: Deep RL for Real-Time Task Offloading in Vehicular Edge Networks
by: Paknejad, Mahsa, et al.
Published: (2025)
by: Paknejad, Mahsa, et al.
Published: (2025)
Similar Items
-
The MoE-Empowered Edge LLMs Deployment: Architecture, Challenges, and Opportunities
by: Li, Ning, et al.
Published: (2025) -
MoE$^2$: Optimizing Collaborative Inference for Edge Large Language Models
by: Jin, Lyudong, et al.
Published: (2025) -
SiftMoE: Similarity-Aware Energy-Efficient Expert Selection for Wireless Distributed MoE Inference
by: Chen, Qian, et al.
Published: (2026) -
UAV-Assisted Cooperative Edge Inference for Low-Altitude Economy via MoE-based Hierarchical Deep Reinforcement Learning
by: Zhuang, Wenhao, et al.
Published: (2026) -
Cluster Topology-Driven Placement of Experts Reduces Network Traffic in MoE Inference
by: Sivtsov, Danil, et al.
Published: (2025)