SlimEdge: Performance and Device Aware Distributed DNN Deployment on Resource-Constrained Edge Hardware
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kumar, Mahadev Sunil, Raha, Arnab, Das, Debayan, G, Gopakumar, Chatterjee, Rounak, Mukherjee, Amitava |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EdgeServing: Deadline-Aware Multi-DNN Serving at the Edge
von: Cao, Jiahe, et al.
Veröffentlicht: (2026)
von: Cao, Jiahe, et al.
Veröffentlicht: (2026)
Evaluating Multi-Instance DNN Inferencing on Multiple Accelerators of an Edge Device
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2025)
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2025)
Improved Decision Module Selection for Hierarchical Inference in Resource-Constrained Edge Devices
von: Behera, Adarsh Prasad, et al.
Veröffentlicht: (2024)
von: Behera, Adarsh Prasad, et al.
Veröffentlicht: (2024)
Performance Characterization of Containerized DNN Training and Inference on Edge Accelerators
von: K., Prashanthi S., et al.
Veröffentlicht: (2023)
von: K., Prashanthi S., et al.
Veröffentlicht: (2023)
Preemption Aware Task Scheduling for Priority and Deadline Constrained DNN Inference Task Offloading in Homogeneous Mobile-Edge Networks
von: Cotter, Jamie, et al.
Veröffentlicht: (2025)
von: Cotter, Jamie, et al.
Veröffentlicht: (2025)
Adaptive Device-Edge Collaboration on DNN Inference in AIoT: A Digital Twin-Assisted Approach
von: Hu, Shisheng, et al.
Veröffentlicht: (2024)
von: Hu, Shisheng, et al.
Veröffentlicht: (2024)
CarbonEdge: Leveraging Mesoscale Spatial Carbon-Intensity Variations for Low Carbon Edge Computing
von: Wu, Li, et al.
Veröffentlicht: (2025)
von: Wu, Li, et al.
Veröffentlicht: (2025)
Fulcrum: Optimizing Concurrent DNN Training and Inferencing on Edge Accelerators
von: K., Prashanthi S., et al.
Veröffentlicht: (2025)
von: K., Prashanthi S., et al.
Veröffentlicht: (2025)
LIME:Accelerating Collaborative Lossless LLM Inference on Memory-Constrained Edge Devices
von: Sun, Mingyu, et al.
Veröffentlicht: (2025)
von: Sun, Mingyu, et al.
Veröffentlicht: (2025)
Pagoda: An Energy and Time Roofline Study for DNN Workloads on Edge Accelerators
von: K., Prashanthi S., et al.
Veröffentlicht: (2025)
von: K., Prashanthi S., et al.
Veröffentlicht: (2025)
Online Optimization of DNN Inference Network Utility in Collaborative Edge Computing
von: Li, Rui, et al.
Veröffentlicht: (2024)
von: Li, Rui, et al.
Veröffentlicht: (2024)
GenAI at the Edge: Comprehensive Survey on Empowering Edge Devices
von: Navardi, Mozhgan, et al.
Veröffentlicht: (2025)
von: Navardi, Mozhgan, et al.
Veröffentlicht: (2025)
FailLite: Failure-Resilient Model Serving for Resource-Constrained Edge Environments
von: Wu, Li, et al.
Veröffentlicht: (2025)
von: Wu, Li, et al.
Veröffentlicht: (2025)
Adaptive Heuristics for Scheduling DNN Inferencing on Edge and Cloud for Personalized UAV Fleets
von: Raj, Suman, et al.
Veröffentlicht: (2024)
von: Raj, Suman, et al.
Veröffentlicht: (2024)
Where to Split? A Pareto-Front Analysis of DNN Partitioning for Edge Inference
von: Masud, Adiba, et al.
Veröffentlicht: (2026)
von: Masud, Adiba, et al.
Veröffentlicht: (2026)
FedFog: Resource-Aware Federated Learning in Edge and Fog Networks
von: Sobati-M, Somayeh
Veröffentlicht: (2025)
von: Sobati-M, Somayeh
Veröffentlicht: (2025)
Accelerating Local LLMs on Resource-Constrained Edge Devices via Distributed Prompt Caching
von: Matsutani, Hiroki, et al.
Veröffentlicht: (2026)
von: Matsutani, Hiroki, et al.
Veröffentlicht: (2026)
Infer-EDGE: Dynamic DNN Inference Optimization in 'Just-in-time' Edge-AI Implementations
von: Mounesan, Motahare, et al.
Veröffentlicht: (2025)
von: Mounesan, Motahare, et al.
Veröffentlicht: (2025)
Will LLMs Scaling Hit the Wall? Breaking Barriers via Distributed Resources on Massive Edge Devices
von: Shen, Tao, et al.
Veröffentlicht: (2025)
von: Shen, Tao, et al.
Veröffentlicht: (2025)
Collaborative Processing for Multi-Tenant Inference on Memory-Constrained Edge TPUs
von: Ng, Nathan, et al.
Veröffentlicht: (2026)
von: Ng, Nathan, et al.
Veröffentlicht: (2026)
Energy-Efficient Joint Offloading and Resource Allocation for Deadline-Constrained Tasks in Multi-Access Edge Computing
von: Gao, Chuanchao, et al.
Veröffentlicht: (2025)
von: Gao, Chuanchao, et al.
Veröffentlicht: (2025)
Failure-Resilient and Carbon-Efficient Deployment of Microservices over the Cloud-Edge Continuum
von: Ponce, Francisco, et al.
Veröffentlicht: (2026)
von: Ponce, Francisco, et al.
Veröffentlicht: (2026)
AdaBridge: Dynamic Data and Computation Reuse for Efficient Multi-task DNN Co-evolution in Edge Systems
von: Wang, Lehao, et al.
Veröffentlicht: (2024)
von: Wang, Lehao, et al.
Veröffentlicht: (2024)
A Survey on Collaborative DNN Inference for Edge Intelligence
von: Ren, Weiqing, et al.
Veröffentlicht: (2022)
von: Ren, Weiqing, et al.
Veröffentlicht: (2022)
ADApt: Edge Device Anomaly Detection and Microservice Replica Prediction
von: Mehran, Narges, et al.
Veröffentlicht: (2025)
von: Mehran, Narges, et al.
Veröffentlicht: (2025)
Dependency-aware Resource Allocation for Serverless Functions at the Edge
von: Baresi, Luciano, et al.
Veröffentlicht: (2023)
von: Baresi, Luciano, et al.
Veröffentlicht: (2023)
Performance Evaluation of Automated Multi-Service Deployment in Edge-Cloud Environments with the CODECO Toolkit
von: Koukis, Georgios, et al.
Veröffentlicht: (2026)
von: Koukis, Georgios, et al.
Veröffentlicht: (2026)
SLICE: SLO-Driven Scheduling for LLM Inference on Edge Computing Devices
von: Chow, Will
Veröffentlicht: (2025)
von: Chow, Will
Veröffentlicht: (2025)
Learning the Optimal Path and DNN Partition for Collaborative Edge Inference
von: Huang, Yin, et al.
Veröffentlicht: (2024)
von: Huang, Yin, et al.
Veröffentlicht: (2024)
A Joint Time and Energy-Efficient Federated Learning-based Computation Offloading Method for Mobile Edge Computing
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
Toward Sustainability-Aware LLM Inference on Edge Clusters
von: Rajashekar, Kolichala, et al.
Veröffentlicht: (2025)
von: Rajashekar, Kolichala, et al.
Veröffentlicht: (2025)
Environment-Aware Dynamic Pruning for Pipelined Edge Inference
von: O'Quinn, Austin, et al.
Veröffentlicht: (2025)
von: O'Quinn, Austin, et al.
Veröffentlicht: (2025)
EnFed: An Energy-aware Federated Learning in Resource Constrained Environments for Human Activity Recognition
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
von: Mukherjee, Anwesha, et al.
Veröffentlicht: (2024)
Distributed Resource Selection for Self-Organising Cloud-Edge Systems
von: Renau, Quentin, et al.
Veröffentlicht: (2025)
von: Renau, Quentin, et al.
Veröffentlicht: (2025)
Resource-efficient Parallel Split Learning in Heterogeneous Edge Computing
von: Zhang, Mingjin, et al.
Veröffentlicht: (2024)
von: Zhang, Mingjin, et al.
Veröffentlicht: (2024)
Characterizing the Performance of Accelerated Jetson Edge Devices for Training Deep Learning Models
von: K., Prashanthi S., et al.
Veröffentlicht: (2025)
von: K., Prashanthi S., et al.
Veröffentlicht: (2025)
OCTOPINF: Workload-Aware Inference Serving for Edge Video Analytics
von: Nguyen, Thanh-Tung, et al.
Veröffentlicht: (2025)
von: Nguyen, Thanh-Tung, et al.
Veröffentlicht: (2025)
Contention-Aware Microservice Deployment in Collaborative Mobile Edge Networks
von: Ge, Xinlei, et al.
Veröffentlicht: (2024)
von: Ge, Xinlei, et al.
Veröffentlicht: (2024)
SkyMemory: A LEO Edge Cache for Transformer Inference Optimization and Scale Out
von: Sandholm, Thomas, et al.
Veröffentlicht: (2025)
von: Sandholm, Thomas, et al.
Veröffentlicht: (2025)
Distributed Edge Analytics in Edge-Fog-Cloud Continuum
von: Srirama, Satish Narayana
Veröffentlicht: (2024)
von: Srirama, Satish Narayana
Veröffentlicht: (2024)
Ähnliche Einträge
-
EdgeServing: Deadline-Aware Multi-DNN Serving at the Edge
von: Cao, Jiahe, et al.
Veröffentlicht: (2026) -
Evaluating Multi-Instance DNN Inferencing on Multiple Accelerators of an Edge Device
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2025) -
Improved Decision Module Selection for Hierarchical Inference in Resource-Constrained Edge Devices
von: Behera, Adarsh Prasad, et al.
Veröffentlicht: (2024) -
Performance Characterization of Containerized DNN Training and Inference on Edge Accelerators
von: K., Prashanthi S., et al.
Veröffentlicht: (2023) -
Preemption Aware Task Scheduling for Priority and Deadline Constrained DNN Inference Task Offloading in Homogeneous Mobile-Edge Networks
von: Cotter, Jamie, et al.
Veröffentlicht: (2025)