EdgeRL: Reinforcement Learning-driven Deep Learning Model Inference Optimization at Edge
Fuente:
arXiv
Saved in:
| Main Authors: | Mounesan, Motahare, Zhang, Xiaojie, Debroy, Saptarshi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Infer-EDGE: Dynamic DNN Inference Optimization in 'Just-in-time' Edge-AI Implementations
by: Mounesan, Motahare, et al.
Published: (2025)
by: Mounesan, Motahare, et al.
Published: (2025)
Reinforcement Learning-driven Data-intensive Workflow Scheduling for Volunteer Edge-Cloud
by: Mounesan, Motahare, et al.
Published: (2024)
by: Mounesan, Motahare, et al.
Published: (2024)
Reinforcement Learning-Driven Edge Management for Reliable Multi-view 3D Reconstruction
by: Mounesan, Motahare, et al.
Published: (2025)
by: Mounesan, Motahare, et al.
Published: (2025)
Variational Autoencoder-Based Black-Box Adversarial Attack on Collaborative DNN Inference
by: Yousefi, Shima, et al.
Published: (2025)
by: Yousefi, Shima, et al.
Published: (2025)
Research on Edge Computing and Cloud Collaborative Resource Scheduling Optimization Based on Deep Reinforcement Learning
by: Wang, Yuqing, et al.
Published: (2025)
by: Wang, Yuqing, et al.
Published: (2025)
Hermes: Memory-Efficient Pipeline Inference for Large Models on Edge Devices
by: Han, Xueyuan, et al.
Published: (2024)
by: Han, Xueyuan, et al.
Published: (2024)
DistrEE: Distributed Early Exit of Deep Neural Network Inference on Edge Devices
by: Peng, Xian, et al.
Published: (2025)
by: Peng, Xian, et al.
Published: (2025)
Inference Offloading for Cost-Sensitive Binary Classification at the Edge
by: Moothedath, Vishnu Narayanan, et al.
Published: (2025)
by: Moothedath, Vishnu Narayanan, et al.
Published: (2025)
Online Client Scheduling and Resource Allocation for Efficient Federated Edge Learning
by: Gao, Zhidong, et al.
Published: (2024)
by: Gao, Zhidong, et al.
Published: (2024)
Heterogeneity-Aware Resource Allocation and Topology Design for Hierarchical Federated Edge Learning
by: Gao, Zhidong, et al.
Published: (2024)
by: Gao, Zhidong, et al.
Published: (2024)
Intelligent Offloading in Vehicular Edge Computing: A Comprehensive Review of Deep Reinforcement Learning Approaches and Architectures
by: Uddin, Ashab, et al.
Published: (2025)
by: Uddin, Ashab, et al.
Published: (2025)
Edge-Cloud Collaborative Computing on Distributed Intelligence and Model Optimization: A Survey
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Deep Reinforcement Learning for Optimizing Energy Consumption in Smart Grid Systems
by: Alsheikhi, Abeer, et al.
Published: (2026)
by: Alsheikhi, Abeer, et al.
Published: (2026)
Personalizing Federated Learning for Hierarchical Edge Networks with Non-IID Data
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
HiDP: Hierarchical DNN Partitioning for Distributed Inference on Heterogeneous Edge Platforms
by: Taufique, Zain, et al.
Published: (2024)
by: Taufique, Zain, et al.
Published: (2024)
Distributed Inference on Mobile Edge and Cloud: An Early Exit based Clustering Approach
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
Interpretable Modeling of Deep Reinforcement Learning Driven Scheduling
by: Li, Boyang, et al.
Published: (2024)
by: Li, Boyang, et al.
Published: (2024)
EdgeLoRA: An Efficient Multi-Tenant LLM Serving System on Edge Devices
by: Shen, Zheyu, et al.
Published: (2025)
by: Shen, Zheyu, et al.
Published: (2025)
AdaptSFL: Adaptive Split Federated Learning in Resource-constrained Edge Networks
by: Lin, Zheng, et al.
Published: (2024)
by: Lin, Zheng, et al.
Published: (2024)
HASFL: Heterogeneity-aware Split Federated Learning over Edge Computing Systems
by: Lin, Zheng, et al.
Published: (2025)
by: Lin, Zheng, et al.
Published: (2025)
Federated Attention: A Distributed Paradigm for Collaborative LLM Inference over Edge Networks
by: Deng, Xiumei, et al.
Published: (2025)
by: Deng, Xiumei, et al.
Published: (2025)
P3SL: Personalized Privacy-Preserving Split Learning on Heterogeneous Edge Devices
by: Fan, Wei, et al.
Published: (2025)
by: Fan, Wei, et al.
Published: (2025)
SwapNet: Efficient Swapping for DNN Inference on Edge AI Devices Beyond the Memory Budget
by: Wang, Kun, et al.
Published: (2024)
by: Wang, Kun, et al.
Published: (2024)
Multi-Worker Selection based Distributed Swarm Learning for Edge IoT with Non-i.i.d. Data
by: Yao, Zhuoyu, et al.
Published: (2025)
by: Yao, Zhuoyu, et al.
Published: (2025)
Task Graph offloading via Deep Reinforcement Learning in Mobile Edge Computing
by: Liu, Jiagang, et al.
Published: (2023)
by: Liu, Jiagang, et al.
Published: (2023)
FedMHO: Heterogeneous One-Shot Federated Learning Towards Resource-Constrained Edge Devices
by: Yao, Dezhong, et al.
Published: (2025)
by: Yao, Dezhong, et al.
Published: (2025)
Sometimes Painful but Certainly Promising: Feasibility and Trade-offs of Language Model Inference at the Edge
by: Abstreiter, Maximilian, et al.
Published: (2025)
by: Abstreiter, Maximilian, et al.
Published: (2025)
QPART: Adaptive Model Quantization and Dynamic Workload Balancing for Accuracy-aware Edge Inference
by: Li, Xiangchen, et al.
Published: (2025)
by: Li, Xiangchen, et al.
Published: (2025)
Enhancing Kubernetes Automated Scheduling with Deep Learning and Reinforcement Techniques for Large-Scale Cloud Computing Optimization
by: Xu, Zheng, et al.
Published: (2024)
by: Xu, Zheng, et al.
Published: (2024)
Deep Reinforcement Learning for System-on-Chip: Myths and Realities
by: Sung, Tegg Taekyong, et al.
Published: (2022)
by: Sung, Tegg Taekyong, et al.
Published: (2022)
Bitnet.cpp: Efficient Edge Inference for Ternary LLMs
by: Wang, Jinheng, et al.
Published: (2025)
by: Wang, Jinheng, et al.
Published: (2025)
Game-Theoretic Deep Reinforcement Learning to Minimize Carbon Emissions and Energy Costs for AI Inference Workloads in Geo-Distributed Data Centers
by: Hogade, Ninad, et al.
Published: (2024)
by: Hogade, Ninad, et al.
Published: (2024)
Learning the Optimal Path and DNN Partition for Collaborative Edge Inference
by: Huang, Yin, et al.
Published: (2024)
by: Huang, Yin, et al.
Published: (2024)
Two-Timescale Model Caching and Resource Allocation for Edge-Enabled AI-Generated Content Services
by: Liu, Zhang, et al.
Published: (2024)
by: Liu, Zhang, et al.
Published: (2024)
CarbonEdge: Carbon-Aware Deep Learning Inference Framework for Sustainable Edge Computing
by: Zhang, Guilin, et al.
Published: (2026)
by: Zhang, Guilin, et al.
Published: (2026)
Tiny Deep Ensemble: Uncertainty Estimation in Edge AI Accelerators via Ensembling Normalization Layers with Shared Weights
by: Ahmed, Soyed Tuhin, et al.
Published: (2024)
by: Ahmed, Soyed Tuhin, et al.
Published: (2024)
Split Federated Learning Over Heterogeneous Edge Devices: Algorithm and Optimization
by: Sun, Yunrui, et al.
Published: (2024)
by: Sun, Yunrui, et al.
Published: (2024)
Quality Scalable Quantization Methodology for Deep Learning on Edge
by: Khaliq, Salman Abdul, et al.
Published: (2024)
by: Khaliq, Salman Abdul, et al.
Published: (2024)
Acceleration for Deep Reinforcement Learning using Parallel and Distributed Computing: A Survey
by: Liu, Zhihong, et al.
Published: (2024)
by: Liu, Zhihong, et al.
Published: (2024)
Rethinking Inference Placement for Deep Learning across Edge and Cloud Platforms: A Multi-Objective Optimization Perspective and Future Directions
by: Zhang, Zongshun, et al.
Published: (2025)
by: Zhang, Zongshun, et al.
Published: (2025)
Similar Items
-
Infer-EDGE: Dynamic DNN Inference Optimization in 'Just-in-time' Edge-AI Implementations
by: Mounesan, Motahare, et al.
Published: (2025) -
Reinforcement Learning-driven Data-intensive Workflow Scheduling for Volunteer Edge-Cloud
by: Mounesan, Motahare, et al.
Published: (2024) -
Reinforcement Learning-Driven Edge Management for Reliable Multi-view 3D Reconstruction
by: Mounesan, Motahare, et al.
Published: (2025) -
Variational Autoencoder-Based Black-Box Adversarial Attack on Collaborative DNN Inference
by: Yousefi, Shima, et al.
Published: (2025) -
Research on Edge Computing and Cloud Collaborative Resource Scheduling Optimization Based on Deep Reinforcement Learning
by: Wang, Yuqing, et al.
Published: (2025)