Where to Split? A Pareto-Front Analysis of DNN Partitioning for Edge Inference
Fuente:
arXiv
Guardado en:
| Autores principales: | Masud, Adiba, Foley, Nicholas, Rajarajan, Pragathi Durga, Lama, Palden |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Converting Autoencoder Toward Low-latency and Energy-efficient DNN Inference at the Edge
por: Mahmud, Hasanul, et al.
Publicado: (2024)
por: Mahmud, Hasanul, et al.
Publicado: (2024)
Learning the Optimal Path and DNN Partition for Collaborative Edge Inference
por: Huang, Yin, et al.
Publicado: (2024)
por: Huang, Yin, et al.
Publicado: (2024)
Performance Characterization of Containerized DNN Training and Inference on Edge Accelerators
por: K., Prashanthi S., et al.
Publicado: (2023)
por: K., Prashanthi S., et al.
Publicado: (2023)
Fulcrum: Optimizing Concurrent DNN Training and Inferencing on Edge Accelerators
por: K., Prashanthi S., et al.
Publicado: (2025)
por: K., Prashanthi S., et al.
Publicado: (2025)
Online Optimization of DNN Inference Network Utility in Collaborative Edge Computing
por: Li, Rui, et al.
Publicado: (2024)
por: Li, Rui, et al.
Publicado: (2024)
Evaluating Multi-Instance DNN Inferencing on Multiple Accelerators of an Edge Device
por: Tayal, Mumuksh, et al.
Publicado: (2025)
por: Tayal, Mumuksh, et al.
Publicado: (2025)
Adaptive Heuristics for Scheduling DNN Inferencing on Edge and Cloud for Personalized UAV Fleets
por: Raj, Suman, et al.
Publicado: (2024)
por: Raj, Suman, et al.
Publicado: (2024)
Adaptive Device-Edge Collaboration on DNN Inference in AIoT: A Digital Twin-Assisted Approach
por: Hu, Shisheng, et al.
Publicado: (2024)
por: Hu, Shisheng, et al.
Publicado: (2024)
Infer-EDGE: Dynamic DNN Inference Optimization in 'Just-in-time' Edge-AI Implementations
por: Mounesan, Motahare, et al.
Publicado: (2025)
por: Mounesan, Motahare, et al.
Publicado: (2025)
A Survey on Collaborative DNN Inference for Edge Intelligence
por: Ren, Weiqing, et al.
Publicado: (2022)
por: Ren, Weiqing, et al.
Publicado: (2022)
EdgeServing: Deadline-Aware Multi-DNN Serving at the Edge
por: Cao, Jiahe, et al.
Publicado: (2026)
por: Cao, Jiahe, et al.
Publicado: (2026)
HiDP: Hierarchical DNN Partitioning for Distributed Inference on Heterogeneous Edge Platforms
por: Taufique, Zain, et al.
Publicado: (2024)
por: Taufique, Zain, et al.
Publicado: (2024)
EcoFed: Efficient Communication for DNN Partitioning-based Federated Learning
por: Wu, Di, et al.
Publicado: (2023)
por: Wu, Di, et al.
Publicado: (2023)
Collaborative Satellite Computing through Adaptive DNN Task Splitting and Offloading
por: Peng, Shifeng, et al.
Publicado: (2024)
por: Peng, Shifeng, et al.
Publicado: (2024)
Preemption Aware Task Scheduling for Priority and Deadline Constrained DNN Inference Task Offloading in Homogeneous Mobile-Edge Networks
por: Cotter, Jamie, et al.
Publicado: (2025)
por: Cotter, Jamie, et al.
Publicado: (2025)
Harpagon: Minimizing DNN Serving Cost via Efficient Dispatching, Scheduling and Splitting
por: Zhao, Zhixin, et al.
Publicado: (2024)
por: Zhao, Zhixin, et al.
Publicado: (2024)
Pagoda: An Energy and Time Roofline Study for DNN Workloads on Edge Accelerators
por: K., Prashanthi S., et al.
Publicado: (2025)
por: K., Prashanthi S., et al.
Publicado: (2025)
Model Partition and Resource Allocation for Split Learning in Vehicular Edge Networks
por: Yu, Lu, et al.
Publicado: (2024)
por: Yu, Lu, et al.
Publicado: (2024)
S2M3: Split-and-Share Multi-Modal Models for Distributed Multi-Task Inference on the Edge
por: Yoon, JinYi, et al.
Publicado: (2025)
por: Yoon, JinYi, et al.
Publicado: (2025)
Collaborative Inference in DNN-based Satellite Systems with Dynamic Task Streams
por: Guan, Jinglong, et al.
Publicado: (2023)
por: Guan, Jinglong, et al.
Publicado: (2023)
SparOA: Sparse and Operator-aware Hybrid Scheduling for Edge DNN Inference
por: Zhang, Ziyang, et al.
Publicado: (2025)
por: Zhang, Ziyang, et al.
Publicado: (2025)
HarmonyBatch: Batching multi-SLO DNN Inference with Heterogeneous Serverless Functions
por: Chen, Jiabin, et al.
Publicado: (2024)
por: Chen, Jiabin, et al.
Publicado: (2024)
AdaOper: Energy-efficient and Responsive Concurrent DNN Inference on Mobile Devices
por: Lin, Zheng, et al.
Publicado: (2024)
por: Lin, Zheng, et al.
Publicado: (2024)
DARIS: An Oversubscribed Spatio-Temporal Scheduler for Real-Time DNN Inference on GPUs
por: Babaei, Amir Fakhim, et al.
Publicado: (2025)
por: Babaei, Amir Fakhim, et al.
Publicado: (2025)
Cooperative Inference with Interleaved Operator Partitioning for CNNs
por: Liu, Zhibang, et al.
Publicado: (2024)
por: Liu, Zhibang, et al.
Publicado: (2024)
Multi-DNN Inference of Sparse Models on Edge SoCs
por: Luo, Jiawei, et al.
Publicado: (2026)
por: Luo, Jiawei, et al.
Publicado: (2026)
Large Language Model Partitioning for Low-Latency Inference at the Edge
por: Kafetzis, Dimitrios, et al.
Publicado: (2025)
por: Kafetzis, Dimitrios, et al.
Publicado: (2025)
RAPID: Redundancy-Aware and Compatibility-Optimal Edge-Cloud Partitioned Inference for Diverse VLA Models
por: Zheng, Zihao, et al.
Publicado: (2026)
por: Zheng, Zihao, et al.
Publicado: (2026)
Collaborative Inference Acceleration with Non-Penetrative Tensor Partitioning
por: Liu, Zhibang, et al.
Publicado: (2025)
por: Liu, Zhibang, et al.
Publicado: (2025)
ParvaGPU: Efficient Spatial GPU Sharing for Large-Scale DNN Inference in Cloud Environments
por: Lee, Munkyu, et al.
Publicado: (2024)
por: Lee, Munkyu, et al.
Publicado: (2024)
Robust DNN Partitioning and Resource Allocation Under Uncertain Inference Time
por: Nan, Zhaojun, et al.
Publicado: (2025)
por: Nan, Zhaojun, et al.
Publicado: (2025)
AdaBridge: Dynamic Data and Computation Reuse for Efficient Multi-task DNN Co-evolution in Edge Systems
por: Wang, Lehao, et al.
Publicado: (2024)
por: Wang, Lehao, et al.
Publicado: (2024)
Resource-efficient Parallel Split Learning in Heterogeneous Edge Computing
por: Zhang, Mingjin, et al.
Publicado: (2024)
por: Zhang, Mingjin, et al.
Publicado: (2024)
EdgeShard: Efficient LLM Inference via Collaborative Edge Computing
por: Zhang, Mingjin, et al.
Publicado: (2024)
por: Zhang, Mingjin, et al.
Publicado: (2024)
MOPAR: A Model Partitioning Framework for Deep Learning Inference Services on Serverless Platforms
por: Duan, Jiaang, et al.
Publicado: (2024)
por: Duan, Jiaang, et al.
Publicado: (2024)
Communication-Computation Pipeline Parallel Split Learning over Wireless Edge Networks
por: Liu, Chenyu, et al.
Publicado: (2025)
por: Liu, Chenyu, et al.
Publicado: (2025)
Practical Performance Guarantees for Pipelined DNN Inference
por: Archer, Aaron, et al.
Publicado: (2023)
por: Archer, Aaron, et al.
Publicado: (2023)
High-Efficiency Split Computing for Cooperative Edge Systems: A Novel Compressed Sensing Bottleneck
por: Zhong, Hailin, et al.
Publicado: (2025)
por: Zhong, Hailin, et al.
Publicado: (2025)
QEIL v2: Heterogeneous Computing for Edge Intelligence via Roofline-Derived Pareto-Optimal Energy Modeling and Multi-Objective Orchestration
por: Kumar, Satyam, et al.
Publicado: (2026)
por: Kumar, Satyam, et al.
Publicado: (2026)
DSPE: Profit Maximization in Edge-Cloud Storage System using Dynamic Space Partitioning with Erasure Code
por: Roy, Shubhradeep, et al.
Publicado: (2025)
por: Roy, Shubhradeep, et al.
Publicado: (2025)
Ejemplares similares
-
A Converting Autoencoder Toward Low-latency and Energy-efficient DNN Inference at the Edge
por: Mahmud, Hasanul, et al.
Publicado: (2024) -
Learning the Optimal Path and DNN Partition for Collaborative Edge Inference
por: Huang, Yin, et al.
Publicado: (2024) -
Performance Characterization of Containerized DNN Training and Inference on Edge Accelerators
por: K., Prashanthi S., et al.
Publicado: (2023) -
Fulcrum: Optimizing Concurrent DNN Training and Inferencing on Edge Accelerators
por: K., Prashanthi S., et al.
Publicado: (2025) -
Online Optimization of DNN Inference Network Utility in Collaborative Edge Computing
por: Li, Rui, et al.
Publicado: (2024)