Graph Neural Networks Automated Design and Deployment on Device-Edge Co-Inference Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Ao, Yang, Jianlei, Qiao, Tong, Qi, Yingjie, Yang, Zhi, Zhao, Weisheng, Hu, Chunming |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GCoDE: Efficient Device-Edge Co-Inference for GNNs via Architecture-Mapping Co-Search
by: Zhou, Ao, et al.
Published: (2025)
by: Zhou, Ao, et al.
Published: (2025)
HGNAS: Hardware-Aware Graph Neural Architecture Search for Edge Devices
by: Zhou, Ao, et al.
Published: (2024)
by: Zhou, Ao, et al.
Published: (2024)
GNNavigator: Towards Adaptive Training of Graph Neural Networks via Automatic Guideline Exploration
by: Qiao, Tong, et al.
Published: (2024)
by: Qiao, Tong, et al.
Published: (2024)
ACE-GNN: Adaptive GNN Co-Inference with System-Aware Scheduling in Dynamic Edge Environments
by: Zhou, Ao, et al.
Published: (2025)
by: Zhou, Ao, et al.
Published: (2025)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
by: Qiao, Tong, et al.
Published: (2025)
by: Qiao, Tong, et al.
Published: (2025)
Finesse: An Agile Design Framework for Pairing-based Cryptography via Software/Hardware Co-Design
by: Pan, Tianwei, et al.
Published: (2025)
by: Pan, Tianwei, et al.
Published: (2025)
TinyFormer: Efficient Transformer Design and Deployment on Tiny Devices
by: Yang, Jianlei, et al.
Published: (2023)
by: Yang, Jianlei, et al.
Published: (2023)
Combinatorial Optimization with Automated Graph Neural Networks
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
CoVSpec: Efficient Device-Edge Co-Inference for Vision-Language Models via Speculative Decoding
by: Jia, Yuanyuan, et al.
Published: (2026)
by: Jia, Yuanyuan, et al.
Published: (2026)
HPIM: Heterogeneous Processing-In-Memory-based Accelerator for Large Language Models Inference
by: Duan, Cenlin, et al.
Published: (2025)
by: Duan, Cenlin, et al.
Published: (2025)
Task-Oriented Real-time Visual Inference for IoVT Systems: A Co-design Framework of Neural Networks and Edge Deployment
by: Wu, Jiaqi, et al.
Published: (2024)
by: Wu, Jiaqi, et al.
Published: (2024)
On-Demand Multi-Task Sparsity for Efficient Large-Model Deployment on Edge Devices
by: Huang, Lianming, et al.
Published: (2025)
by: Huang, Lianming, et al.
Published: (2025)
LOGIN: A Large Language Model Consulted Graph Neural Network Training Framework
by: Qiao, Yiran, et al.
Published: (2024)
by: Qiao, Yiran, et al.
Published: (2024)
Profiling-Driven Adaptive Distributed Transformer Inference on Embedded Edge Deployment
by: Qazi, Muhammad Azlan, et al.
Published: (2026)
by: Qazi, Muhammad Azlan, et al.
Published: (2026)
DistrEE: Distributed Early Exit of Deep Neural Network Inference on Edge Devices
by: Peng, Xian, et al.
Published: (2025)
by: Peng, Xian, et al.
Published: (2025)
Resource-Efficient Generative AI Model Deployment in Mobile Edge Networks
by: Liang, Yuxin, et al.
Published: (2024)
by: Liang, Yuxin, et al.
Published: (2024)
Co-Designing Binarized Transformer and Hardware Accelerator for Efficient End-to-End Edge Deployment
by: Ji, Yuhao, et al.
Published: (2024)
by: Ji, Yuhao, et al.
Published: (2024)
CIMFlow: An Integrated Framework for Systematic Design and Evaluation of Digital CIM Architectures
by: Qi, Yingjie, et al.
Published: (2025)
by: Qi, Yingjie, et al.
Published: (2025)
RefinementEngine: Automating Intent-to-Device Filtering Policy Deployment under Network Constraints
by: Colaiacomo, Davide, et al.
Published: (2026)
by: Colaiacomo, Davide, et al.
Published: (2026)
CIMinus: Empowering Sparse DNN Workloads Modeling and Exploration on SRAM-based CIM Architectures
by: Qi, Yingjie, et al.
Published: (2025)
by: Qi, Yingjie, et al.
Published: (2025)
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices
by: Qin, Ruiyang, et al.
Published: (2024)
by: Qin, Ruiyang, et al.
Published: (2024)
LPS-GNN : Deploying Graph Neural Networks on Graphs with 100-Billion Edges
by: Cheng, Xu, et al.
Published: (2025)
by: Cheng, Xu, et al.
Published: (2025)
Efficient SRAM-PIM Co-design by Joint Exploration of Value-Level and Bit-Level Sparsity
by: Duan, Cenlin, et al.
Published: (2025)
by: Duan, Cenlin, et al.
Published: (2025)
Edge Prompt Tuning for Graph Neural Networks
by: Fu, Xingbo, et al.
Published: (2025)
by: Fu, Xingbo, et al.
Published: (2025)
Refined Edge Usage of Graph Neural Networks for Edge Prediction
by: Jin, Jiarui, et al.
Published: (2022)
by: Jin, Jiarui, et al.
Published: (2022)
Efficient Edge AI: Deploying Convolutional Neural Networks on FPGA with the Gemmini Accelerator
by: Peccia, Federico Nicolas, et al.
Published: (2024)
by: Peccia, Federico Nicolas, et al.
Published: (2024)
Decision-focused Graph Neural Networks for Combinatorial Optimization
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Relation-Aware Network with Attention-Based Loss for Few-Shot Knowledge Graph Completion
by: Qiao, Qiao, et al.
Published: (2023)
by: Qiao, Qiao, et al.
Published: (2023)
Towards Efficient SRAM-PIM Architecture Design by Exploiting Unstructured Bit-Level Sparsity
by: Duan, Cenlin, et al.
Published: (2024)
by: Duan, Cenlin, et al.
Published: (2024)
Automated Design of Agentic Systems
by: Hu, Shengran, et al.
Published: (2024)
by: Hu, Shengran, et al.
Published: (2024)
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems
by: Zhou, Wenqing, et al.
Published: (2025)
by: Zhou, Wenqing, et al.
Published: (2025)
Deploying Graph Neural Networks in Wireless Networks: A Link Stability Viewpoint
by: Li, Jun, et al.
Published: (2024)
by: Li, Jun, et al.
Published: (2024)
Empowering In-Browser Deep Learning Inference on Edge Devices with Just-in-Time Kernel Optimizations
by: Jia, Fucheng, et al.
Published: (2023)
by: Jia, Fucheng, et al.
Published: (2023)
Edge Graph Intelligence: Reciprocally Empowering Edge Networks with Graph Intelligence
by: Zeng, Liekang, et al.
Published: (2024)
by: Zeng, Liekang, et al.
Published: (2024)
KnowPath: Knowledge-enhanced Reasoning via LLM-generated Inference Paths over Knowledge Graphs
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
Collaborative Compression for Large-Scale MoE Deployment on Edge
by: Chen, Yixiao, et al.
Published: (2025)
by: Chen, Yixiao, et al.
Published: (2025)
Prompt-based Unifying Inference Attack on Graph Neural Networks
by: Wei, Yuecen, et al.
Published: (2024)
by: Wei, Yuecen, et al.
Published: (2024)
EdgeMoE: Empowering Sparse Large Language Models on Mobile Devices
by: Yi, Rongjie, et al.
Published: (2023)
by: Yi, Rongjie, et al.
Published: (2023)
Evolution and Efficiency in Neural Architecture Search: Bridging the Gap Between Expert Design and Automated Optimization
by: Meng, Fanfei, et al.
Published: (2024)
by: Meng, Fanfei, et al.
Published: (2024)
Dynamic Quality-Latency Aware Routing for LLM Inference in Wireless Edge-Device Networks
by: Bao, Rui, et al.
Published: (2025)
by: Bao, Rui, et al.
Published: (2025)
Similar Items
-
GCoDE: Efficient Device-Edge Co-Inference for GNNs via Architecture-Mapping Co-Search
by: Zhou, Ao, et al.
Published: (2025) -
HGNAS: Hardware-Aware Graph Neural Architecture Search for Edge Devices
by: Zhou, Ao, et al.
Published: (2024) -
GNNavigator: Towards Adaptive Training of Graph Neural Networks via Automatic Guideline Exploration
by: Qiao, Tong, et al.
Published: (2024) -
ACE-GNN: Adaptive GNN Co-Inference with System-Aware Scheduling in Dynamic Edge Environments
by: Zhou, Ao, et al.
Published: (2025) -
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
by: Qiao, Tong, et al.
Published: (2025)