Rapid Deployment of DNNs for Edge Computing via Structured Pruning at Initialization
Fuente:
arXiv
Saved in:
| Main Authors: | Eccles, Bailey J., Wong, Leon, Varghese, Blesson |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mosaic: Composite Projection Pruning for Resource-efficient LLMs
by: Eccles, Bailey J., et al.
Published: (2025)
by: Eccles, Bailey J., et al.
Published: (2025)
DNNShifter: An Efficient DNN Pruning System for Edge Computing
by: Eccles, Bailey J., et al.
Published: (2023)
by: Eccles, Bailey J., et al.
Published: (2023)
Data-Free Pruning of Self-Attention Layers in LLMs
by: Saikumar, Dhananjay, et al.
Published: (2025)
by: Saikumar, Dhananjay, et al.
Published: (2025)
Signal Collapse in One-Shot Pruning: When Sparse Models Fail to Distinguish Neural Representations
by: Saikumar, Dhananjay, et al.
Published: (2025)
by: Saikumar, Dhananjay, et al.
Published: (2025)
DRIVE: Dual Gradient-Based Rapid Iterative Pruning
by: Saikumar, Dhananjay, et al.
Published: (2024)
by: Saikumar, Dhananjay, et al.
Published: (2024)
Lightweight Edge Learning via Dataset Pruning
by: Ale, Laha, et al.
Published: (2026)
by: Ale, Laha, et al.
Published: (2026)
FedOptima: Optimizing Resource Utilization in Federated Learning
by: Zhang, Zihan, et al.
Published: (2025)
by: Zhang, Zihan, et al.
Published: (2025)
Ampere: Communication-Efficient and High-Accuracy Split Federated Learning
by: Zhang, Zihan, et al.
Published: (2025)
by: Zhang, Zihan, et al.
Published: (2025)
FedPaI: Achieving Extreme Sparsity in Federated Learning via Pruning at Initialization
by: Wang, Haonan, et al.
Published: (2025)
by: Wang, Haonan, et al.
Published: (2025)
Cognitive Edge Computing: A Comprehensive Survey on Optimizing Large Models and AI Agents for Pervasive Deployment
by: Wang, Xubin, et al.
Published: (2025)
by: Wang, Xubin, et al.
Published: (2025)
Carbon Intensity-Aware Adaptive Inference of DNNs
by: Jung, Jiwan
Published: (2024)
by: Jung, Jiwan
Published: (2024)
CEAR: Certified Ensemble Adversarial Robustness in DNNs
by: Sadig, Daniel, et al.
Published: (2026)
by: Sadig, Daniel, et al.
Published: (2026)
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
by: Zhang, Tuo, et al.
Published: (2025)
by: Zhang, Tuo, et al.
Published: (2025)
DTMM: Deploying TinyML Models on Extremely Weak IoT Devices with Pruning
by: Han, Lixiang, et al.
Published: (2024)
by: Han, Lixiang, et al.
Published: (2024)
Collaborative Compression for Large-Scale MoE Deployment on Edge
by: Chen, Yixiao, et al.
Published: (2025)
by: Chen, Yixiao, et al.
Published: (2025)
Balanced Edge Pruning for Graph Anomaly Detection with Noisy Labels
by: Wang, Zhu, et al.
Published: (2024)
by: Wang, Zhu, et al.
Published: (2024)
How DNNs break the Curse of Dimensionality: Compositionality and Symmetry Learning
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Resource-Efficient Generative AI Model Deployment in Mobile Edge Networks
by: Liang, Yuxin, et al.
Published: (2024)
by: Liang, Yuxin, et al.
Published: (2024)
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices
by: Qin, Ruiyang, et al.
Published: (2024)
by: Qin, Ruiyang, et al.
Published: (2024)
Spectral Theory for Edge Pruning in Asynchronous Recurrent Graph Neural Networks
by: Bessone, Nicolas
Published: (2025)
by: Bessone, Nicolas
Published: (2025)
IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning
by: Jung, Jaeheun, et al.
Published: (2025)
by: Jung, Jaeheun, et al.
Published: (2025)
Bit-Identical Medical Deep Learning via Structured Orthogonal Initialization
by: Shkolnikov, Yakov Pyotr
Published: (2026)
by: Shkolnikov, Yakov Pyotr
Published: (2026)
On Background Bias of Post-Hoc Concept Embeddings in Computer Vision DNNs
by: Schwalbe, Gesina, et al.
Published: (2025)
by: Schwalbe, Gesina, et al.
Published: (2025)
ECQ$^{\text{x}}$: Explainability-Driven Quantization for Low-Bit and Sparse DNNs
by: Becking, Daniel, et al.
Published: (2021)
by: Becking, Daniel, et al.
Published: (2021)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
Efficient Triple Modular Redundancy for Reliability Enhancement of DNNs Using Explainable AI
by: Soroush, Kimia, et al.
Published: (2025)
by: Soroush, Kimia, et al.
Published: (2025)
Relationship between Uncertainty in DNNs and Adversarial Attacks
by: Ogonna, Mabel, et al.
Published: (2024)
by: Ogonna, Mabel, et al.
Published: (2024)
Random weights of DNNs and emergence of fixed points
by: Berlyand, L., et al.
Published: (2025)
by: Berlyand, L., et al.
Published: (2025)
Achieving Pareto Optimality using Efficient Parameter Reduction for DNNs in Resource-Constrained Edge Environment
by: Mih, Atah Nuh, et al.
Published: (2024)
by: Mih, Atah Nuh, et al.
Published: (2024)
Co-Designing Binarized Transformer and Hardware Accelerator for Efficient End-to-End Edge Deployment
by: Ji, Yuhao, et al.
Published: (2024)
by: Ji, Yuhao, et al.
Published: (2024)
Graph Neural Networks Automated Design and Deployment on Device-Edge Co-Inference Systems
by: Zhou, Ao, et al.
Published: (2024)
by: Zhou, Ao, et al.
Published: (2024)
Structured vs. Unstructured Pruning: An Exponential Gap
by: Ferre', Davide, et al.
Published: (2026)
by: Ferre', Davide, et al.
Published: (2026)
Less is More: Unlocking Specialization of Time Series Foundation Models via Structured Pruning
by: Zhao, Lifan, et al.
Published: (2025)
by: Zhao, Lifan, et al.
Published: (2025)
StructPrune: Structured Global Pruning asymptotics with $\mathcal{O}(\sqrt{N})$ GPU Memory
by: Song, Xinyuan, et al.
Published: (2025)
by: Song, Xinyuan, et al.
Published: (2025)
DapperFL: Domain Adaptive Federated Learning with Model Fusion Pruning for Edge Devices
by: Jia, Yongzhe, et al.
Published: (2024)
by: Jia, Yongzhe, et al.
Published: (2024)
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
by: Roy, Arjun, et al.
Published: (2026)
by: Roy, Arjun, et al.
Published: (2026)
Edge-free but Structure-aware: Prototype-Guided Knowledge Distillation from GNNs to MLPs
by: Wu, Taiqiang, et al.
Published: (2023)
by: Wu, Taiqiang, et al.
Published: (2023)
HAWX: A Hardware-Aware FrameWork for Fast and Scalable ApproXimation of DNNs
by: Nazari, Samira, et al.
Published: (2026)
by: Nazari, Samira, et al.
Published: (2026)
NeuroFlux: Memory-Efficient CNN Training Using Adaptive Local Learning
by: Saikumar, Dhananjay, et al.
Published: (2024)
by: Saikumar, Dhananjay, et al.
Published: (2024)
SPAP: Structured Pruning via Alternating Optimization and Penalty Methods
by: Hu, Hanyu, et al.
Published: (2025)
by: Hu, Hanyu, et al.
Published: (2025)
Similar Items
-
Mosaic: Composite Projection Pruning for Resource-efficient LLMs
by: Eccles, Bailey J., et al.
Published: (2025) -
DNNShifter: An Efficient DNN Pruning System for Edge Computing
by: Eccles, Bailey J., et al.
Published: (2023) -
Data-Free Pruning of Self-Attention Layers in LLMs
by: Saikumar, Dhananjay, et al.
Published: (2025) -
Signal Collapse in One-Shot Pruning: When Sparse Models Fail to Distinguish Neural Representations
by: Saikumar, Dhananjay, et al.
Published: (2025) -
DRIVE: Dual Gradient-Based Rapid Iterative Pruning
by: Saikumar, Dhananjay, et al.
Published: (2024)