Graph Neural Networks Automated Design and Deployment on Device-Edge Co-Inference Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Ao, Yang, Jianlei, Qiao, Tong, Qi, Yingjie, Yang, Zhi, Zhao, Weisheng, Hu, Chunming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GCoDE: Efficient Device-Edge Co-Inference for GNNs via Architecture-Mapping Co-Search
por: Zhou, Ao, et al.
Publicado: (2025)
por: Zhou, Ao, et al.
Publicado: (2025)
HGNAS: Hardware-Aware Graph Neural Architecture Search for Edge Devices
por: Zhou, Ao, et al.
Publicado: (2024)
por: Zhou, Ao, et al.
Publicado: (2024)
GNNavigator: Towards Adaptive Training of Graph Neural Networks via Automatic Guideline Exploration
por: Qiao, Tong, et al.
Publicado: (2024)
por: Qiao, Tong, et al.
Publicado: (2024)
ACE-GNN: Adaptive GNN Co-Inference with System-Aware Scheduling in Dynamic Edge Environments
por: Zhou, Ao, et al.
Publicado: (2025)
por: Zhou, Ao, et al.
Publicado: (2025)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
por: Qiao, Tong, et al.
Publicado: (2025)
por: Qiao, Tong, et al.
Publicado: (2025)
Finesse: An Agile Design Framework for Pairing-based Cryptography via Software/Hardware Co-Design
por: Pan, Tianwei, et al.
Publicado: (2025)
por: Pan, Tianwei, et al.
Publicado: (2025)
TinyFormer: Efficient Transformer Design and Deployment on Tiny Devices
por: Yang, Jianlei, et al.
Publicado: (2023)
por: Yang, Jianlei, et al.
Publicado: (2023)
Combinatorial Optimization with Automated Graph Neural Networks
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
CoVSpec: Efficient Device-Edge Co-Inference for Vision-Language Models via Speculative Decoding
por: Jia, Yuanyuan, et al.
Publicado: (2026)
por: Jia, Yuanyuan, et al.
Publicado: (2026)
HPIM: Heterogeneous Processing-In-Memory-based Accelerator for Large Language Models Inference
por: Duan, Cenlin, et al.
Publicado: (2025)
por: Duan, Cenlin, et al.
Publicado: (2025)
Task-Oriented Real-time Visual Inference for IoVT Systems: A Co-design Framework of Neural Networks and Edge Deployment
por: Wu, Jiaqi, et al.
Publicado: (2024)
por: Wu, Jiaqi, et al.
Publicado: (2024)
On-Demand Multi-Task Sparsity for Efficient Large-Model Deployment on Edge Devices
por: Huang, Lianming, et al.
Publicado: (2025)
por: Huang, Lianming, et al.
Publicado: (2025)
LOGIN: A Large Language Model Consulted Graph Neural Network Training Framework
por: Qiao, Yiran, et al.
Publicado: (2024)
por: Qiao, Yiran, et al.
Publicado: (2024)
Profiling-Driven Adaptive Distributed Transformer Inference on Embedded Edge Deployment
por: Qazi, Muhammad Azlan, et al.
Publicado: (2026)
por: Qazi, Muhammad Azlan, et al.
Publicado: (2026)
DistrEE: Distributed Early Exit of Deep Neural Network Inference on Edge Devices
por: Peng, Xian, et al.
Publicado: (2025)
por: Peng, Xian, et al.
Publicado: (2025)
Resource-Efficient Generative AI Model Deployment in Mobile Edge Networks
por: Liang, Yuxin, et al.
Publicado: (2024)
por: Liang, Yuxin, et al.
Publicado: (2024)
Co-Designing Binarized Transformer and Hardware Accelerator for Efficient End-to-End Edge Deployment
por: Ji, Yuhao, et al.
Publicado: (2024)
por: Ji, Yuhao, et al.
Publicado: (2024)
CIMFlow: An Integrated Framework for Systematic Design and Evaluation of Digital CIM Architectures
por: Qi, Yingjie, et al.
Publicado: (2025)
por: Qi, Yingjie, et al.
Publicado: (2025)
RefinementEngine: Automating Intent-to-Device Filtering Policy Deployment under Network Constraints
por: Colaiacomo, Davide, et al.
Publicado: (2026)
por: Colaiacomo, Davide, et al.
Publicado: (2026)
CIMinus: Empowering Sparse DNN Workloads Modeling and Exploration on SRAM-based CIM Architectures
por: Qi, Yingjie, et al.
Publicado: (2025)
por: Qi, Yingjie, et al.
Publicado: (2025)
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices
por: Qin, Ruiyang, et al.
Publicado: (2024)
por: Qin, Ruiyang, et al.
Publicado: (2024)
LPS-GNN : Deploying Graph Neural Networks on Graphs with 100-Billion Edges
por: Cheng, Xu, et al.
Publicado: (2025)
por: Cheng, Xu, et al.
Publicado: (2025)
Efficient SRAM-PIM Co-design by Joint Exploration of Value-Level and Bit-Level Sparsity
por: Duan, Cenlin, et al.
Publicado: (2025)
por: Duan, Cenlin, et al.
Publicado: (2025)
Edge Prompt Tuning for Graph Neural Networks
por: Fu, Xingbo, et al.
Publicado: (2025)
por: Fu, Xingbo, et al.
Publicado: (2025)
Refined Edge Usage of Graph Neural Networks for Edge Prediction
por: Jin, Jiarui, et al.
Publicado: (2022)
por: Jin, Jiarui, et al.
Publicado: (2022)
Efficient Edge AI: Deploying Convolutional Neural Networks on FPGA with the Gemmini Accelerator
por: Peccia, Federico Nicolas, et al.
Publicado: (2024)
por: Peccia, Federico Nicolas, et al.
Publicado: (2024)
Decision-focused Graph Neural Networks for Combinatorial Optimization
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
Relation-Aware Network with Attention-Based Loss for Few-Shot Knowledge Graph Completion
por: Qiao, Qiao, et al.
Publicado: (2023)
por: Qiao, Qiao, et al.
Publicado: (2023)
Towards Efficient SRAM-PIM Architecture Design by Exploiting Unstructured Bit-Level Sparsity
por: Duan, Cenlin, et al.
Publicado: (2024)
por: Duan, Cenlin, et al.
Publicado: (2024)
Automated Design of Agentic Systems
por: Hu, Shengran, et al.
Publicado: (2024)
por: Hu, Shengran, et al.
Publicado: (2024)
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems
por: Zhou, Wenqing, et al.
Publicado: (2025)
por: Zhou, Wenqing, et al.
Publicado: (2025)
Deploying Graph Neural Networks in Wireless Networks: A Link Stability Viewpoint
por: Li, Jun, et al.
Publicado: (2024)
por: Li, Jun, et al.
Publicado: (2024)
Empowering In-Browser Deep Learning Inference on Edge Devices with Just-in-Time Kernel Optimizations
por: Jia, Fucheng, et al.
Publicado: (2023)
por: Jia, Fucheng, et al.
Publicado: (2023)
Edge Graph Intelligence: Reciprocally Empowering Edge Networks with Graph Intelligence
por: Zeng, Liekang, et al.
Publicado: (2024)
por: Zeng, Liekang, et al.
Publicado: (2024)
KnowPath: Knowledge-enhanced Reasoning via LLM-generated Inference Paths over Knowledge Graphs
por: Zhao, Qi, et al.
Publicado: (2025)
por: Zhao, Qi, et al.
Publicado: (2025)
Collaborative Compression for Large-Scale MoE Deployment on Edge
por: Chen, Yixiao, et al.
Publicado: (2025)
por: Chen, Yixiao, et al.
Publicado: (2025)
Prompt-based Unifying Inference Attack on Graph Neural Networks
por: Wei, Yuecen, et al.
Publicado: (2024)
por: Wei, Yuecen, et al.
Publicado: (2024)
EdgeMoE: Empowering Sparse Large Language Models on Mobile Devices
por: Yi, Rongjie, et al.
Publicado: (2023)
por: Yi, Rongjie, et al.
Publicado: (2023)
Evolution and Efficiency in Neural Architecture Search: Bridging the Gap Between Expert Design and Automated Optimization
por: Meng, Fanfei, et al.
Publicado: (2024)
por: Meng, Fanfei, et al.
Publicado: (2024)
Dynamic Quality-Latency Aware Routing for LLM Inference in Wireless Edge-Device Networks
por: Bao, Rui, et al.
Publicado: (2025)
por: Bao, Rui, et al.
Publicado: (2025)
Ejemplares similares
-
GCoDE: Efficient Device-Edge Co-Inference for GNNs via Architecture-Mapping Co-Search
por: Zhou, Ao, et al.
Publicado: (2025) -
HGNAS: Hardware-Aware Graph Neural Architecture Search for Edge Devices
por: Zhou, Ao, et al.
Publicado: (2024) -
GNNavigator: Towards Adaptive Training of Graph Neural Networks via Automatic Guideline Exploration
por: Qiao, Tong, et al.
Publicado: (2024) -
ACE-GNN: Adaptive GNN Co-Inference with System-Aware Scheduling in Dynamic Edge Environments
por: Zhou, Ao, et al.
Publicado: (2025) -
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
por: Qiao, Tong, et al.
Publicado: (2025)