ECORE: Energy-Conscious Optimized Routing for Deep Learning Models at the Edge
Fuente:
arXiv
Saved in:
| Main Authors: | Alqahtani, Daghash K., Rodriguez, Maria A., Cheema, Muhammad Aamir, Rezatofighi, Hamid, Toosi, Adel N. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Objective Load Balancing for Heterogeneous Edge-Based Object Detection Systems
by: Alqahtani, Daghash K., et al.
Published: (2026)
by: Alqahtani, Daghash K., et al.
Published: (2026)
A Comprehensive Evaluation of Deep Learning Object Detection Models on Heterogeneous Edge Devices
by: Alqahtani, Daghash K., et al.
Published: (2024)
by: Alqahtani, Daghash K., et al.
Published: (2024)
IntentContinuum: Using LLMs to Support Intent-Based Computing Across the Compute Continuum
by: Akbari, Negin, et al.
Published: (2025)
by: Akbari, Negin, et al.
Published: (2025)
Efficient Routing of Inference Requests across LLM Instances in Cloud-Edge Computing
by: Yu, Shibo, et al.
Published: (2025)
by: Yu, Shibo, et al.
Published: (2025)
REACH: Reinforcement Learning for Adaptive Microservice Rescheduling in the Cloud-Edge Continuum
by: Bai, Xu, et al.
Published: (2025)
by: Bai, Xu, et al.
Published: (2025)
Resilience Evaluation of Kubernetes in Cloud-Edge Environments via Failure Injection
by: Chen, Zihao, et al.
Published: (2025)
by: Chen, Zihao, et al.
Published: (2025)
LLM-Driven Intent-Based Privacy-Aware Orchestration Across the Cloud-Edge Continuum
by: Su, Zijie, et al.
Published: (2026)
by: Su, Zijie, et al.
Published: (2026)
A Multi-Armed Bandit-Based Participant Selection Method for Federated Recommendation Systems
by: Liu, Jintao, et al.
Published: (2025)
by: Liu, Jintao, et al.
Published: (2025)
TempoScale: A Cloud Workloads Prediction Approach Integrating Short-Term and Long-Term Information
by: Wen, Linfeng, et al.
Published: (2024)
by: Wen, Linfeng, et al.
Published: (2024)
RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving
by: Kasnavieh, Hossein Hosseini, et al.
Published: (2026)
by: Kasnavieh, Hossein Hosseini, et al.
Published: (2026)
Multi-Layer Scheduling for MoE-Based LLM Reasoning
by: Sun, Yifan, et al.
Published: (2026)
by: Sun, Yifan, et al.
Published: (2026)
PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving
by: Bai, Xu, et al.
Published: (2026)
by: Bai, Xu, et al.
Published: (2026)
GraphFlash: Enabling Fast and Elastic Graph Processing on Serverless Infrastructure
by: Zhao, Chen, et al.
Published: (2026)
by: Zhao, Chen, et al.
Published: (2026)
DistributedEstimator: Distributed Training of Quantum Neural Networks via Circuit Cutting
by: Singh, Prabhjot, et al.
Published: (2026)
by: Singh, Prabhjot, et al.
Published: (2026)
Covariance-Guided Resource Adaptive Learning for Efficient Edge Inference
by: Nabhaan, Ahmad N. L., et al.
Published: (2026)
by: Nabhaan, Ahmad N. L., et al.
Published: (2026)
Characterizing the Performance of Accelerated Jetson Edge Devices for Training Deep Learning Models
by: K., Prashanthi S., et al.
Published: (2025)
by: K., Prashanthi S., et al.
Published: (2025)
Personalizing Federated Learning for Hierarchical Edge Networks with Non-IID Data
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
POSEIDON : Efficient Function Placement at the Edge using Deep Reinforcement Learning
by: Jain, Prakhar, et al.
Published: (2024)
by: Jain, Prakhar, et al.
Published: (2024)
Scale: Deep Reinforcement Learning for Container Scheduling in Serverless Edge Computing
by: Chen, Chen, et al.
Published: (2026)
by: Chen, Chen, et al.
Published: (2026)
Federated Learning within Global Energy Budget over Heterogeneous Edge Accelerators
by: Banerjee, Roopkatha, et al.
Published: (2025)
by: Banerjee, Roopkatha, et al.
Published: (2025)
Mobile Edge Computing
by: Ahmed, Sohaib, et al.
Published: (2024)
by: Ahmed, Sohaib, et al.
Published: (2024)
HybridFlow: Resource-Adaptive Subtask Routing for Efficient Edge-Cloud LLM Inference
by: Dong, Jiangwen, et al.
Published: (2025)
by: Dong, Jiangwen, et al.
Published: (2025)
Janus: Disaggregating Attention and Experts for Scalable MoE Inference
by: Zhang, Zhexiang, et al.
Published: (2025)
by: Zhang, Zhexiang, et al.
Published: (2025)
LLM-Enhanced Deep Reinforcement Learning for Task Offloading in Collaborative Edge Computing
by: Guo, Hao, et al.
Published: (2026)
by: Guo, Hao, et al.
Published: (2026)
Edge Intelligence-Driven LegalEdge Contracts for EV Charging Stations: A Fedrated Learning with Deep Q-Networks Approach
by: Rahmani, Rahim, et al.
Published: (2026)
by: Rahmani, Rahim, et al.
Published: (2026)
Pagoda: An Energy and Time Roofline Study for DNN Workloads on Edge Accelerators
by: K., Prashanthi S., et al.
Published: (2025)
by: K., Prashanthi S., et al.
Published: (2025)
Stochastic Modeling for Energy-Efficient Edge Infrastructure
by: Rossi, Fabio Diniz
Published: (2025)
by: Rossi, Fabio Diniz
Published: (2025)
Distributed Hierarchical Machine Learning for Joint Resource Allocation and Slice Selection in In-Network Edge Systems
by: Rashid, Sulaiman Muhammad, et al.
Published: (2025)
by: Rashid, Sulaiman Muhammad, et al.
Published: (2025)
CroSatFL: Energy-Efficient Federated Learning with Cross-Aggregation for Satellite Edge Computing
by: Yang, Nan, et al.
Published: (2026)
by: Yang, Nan, et al.
Published: (2026)
Fulcrum: Optimizing Concurrent DNN Training and Inferencing on Edge Accelerators
by: K., Prashanthi S., et al.
Published: (2025)
by: K., Prashanthi S., et al.
Published: (2025)
Energy-aware Distributed Microservice Request Placement at the Edge
by: Toczé, Klervie, et al.
Published: (2024)
by: Toczé, Klervie, et al.
Published: (2024)
Energy Metrics for Edge Microservice Request Placement Strategies
by: Toczé, Klervie, et al.
Published: (2025)
by: Toczé, Klervie, et al.
Published: (2025)
Token Level Routing Inference System for Edge Devices
by: She, Jianshu, et al.
Published: (2025)
by: She, Jianshu, et al.
Published: (2025)
Stable-MoE: Lyapunov-based Token Routing for Distributed Mixture-of-Experts Training over Edge Networks
by: Shi, Long, et al.
Published: (2025)
by: Shi, Long, et al.
Published: (2025)
Decentralized LLM Inference over Edge Networks with Energy Harvesting
by: Khoshsirat, Aria, et al.
Published: (2024)
by: Khoshsirat, Aria, et al.
Published: (2024)
CoEdge-RAG: Optimizing Hierarchical Scheduling for Retrieval-Augmented LLMs in Collaborative Edge Computing
by: Hong, Guihang, et al.
Published: (2025)
by: Hong, Guihang, et al.
Published: (2025)
DeepCompile: A Compiler-Driven Approach to Optimizing Distributed Deep Learning Training
by: Tanaka, Masahiro, et al.
Published: (2025)
by: Tanaka, Masahiro, et al.
Published: (2025)
A Joint Time and Energy-Efficient Federated Learning-based Computation Offloading Method for Mobile Edge Computing
by: Mukherjee, Anwesha, et al.
Published: (2024)
by: Mukherjee, Anwesha, et al.
Published: (2024)
ReinFog: A Deep Reinforcement Learning Empowered Framework for Resource Management in Edge and Cloud Computing Environments
by: Wang, Zhiyu, et al.
Published: (2024)
by: Wang, Zhiyu, et al.
Published: (2024)
Efficient Accelerated Graph Edit Distance Computation on GPU
by: Dabah, Adel, et al.
Published: (2026)
by: Dabah, Adel, et al.
Published: (2026)
Similar Items
-
Multi-Objective Load Balancing for Heterogeneous Edge-Based Object Detection Systems
by: Alqahtani, Daghash K., et al.
Published: (2026) -
A Comprehensive Evaluation of Deep Learning Object Detection Models on Heterogeneous Edge Devices
by: Alqahtani, Daghash K., et al.
Published: (2024) -
IntentContinuum: Using LLMs to Support Intent-Based Computing Across the Compute Continuum
by: Akbari, Negin, et al.
Published: (2025) -
Efficient Routing of Inference Requests across LLM Instances in Cloud-Edge Computing
by: Yu, Shibo, et al.
Published: (2025) -
REACH: Reinforcement Learning for Adaptive Microservice Rescheduling in the Cloud-Edge Continuum
by: Bai, Xu, et al.
Published: (2025)