Rapid Deployment of DNNs for Edge Computing via Structured Pruning at Initialization
Fuente:
arXiv
Salvato in:
| Autori principali: | Eccles, Bailey J., Wong, Leon, Varghese, Blesson |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mosaic: Composite Projection Pruning for Resource-efficient LLMs
di: Eccles, Bailey J., et al.
Pubblicazione: (2025)
di: Eccles, Bailey J., et al.
Pubblicazione: (2025)
DNNShifter: An Efficient DNN Pruning System for Edge Computing
di: Eccles, Bailey J., et al.
Pubblicazione: (2023)
di: Eccles, Bailey J., et al.
Pubblicazione: (2023)
Data-Free Pruning of Self-Attention Layers in LLMs
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2025)
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2025)
Signal Collapse in One-Shot Pruning: When Sparse Models Fail to Distinguish Neural Representations
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2025)
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2025)
DRIVE: Dual Gradient-Based Rapid Iterative Pruning
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2024)
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2024)
Lightweight Edge Learning via Dataset Pruning
di: Ale, Laha, et al.
Pubblicazione: (2026)
di: Ale, Laha, et al.
Pubblicazione: (2026)
FedOptima: Optimizing Resource Utilization in Federated Learning
di: Zhang, Zihan, et al.
Pubblicazione: (2025)
di: Zhang, Zihan, et al.
Pubblicazione: (2025)
Ampere: Communication-Efficient and High-Accuracy Split Federated Learning
di: Zhang, Zihan, et al.
Pubblicazione: (2025)
di: Zhang, Zihan, et al.
Pubblicazione: (2025)
FedPaI: Achieving Extreme Sparsity in Federated Learning via Pruning at Initialization
di: Wang, Haonan, et al.
Pubblicazione: (2025)
di: Wang, Haonan, et al.
Pubblicazione: (2025)
Cognitive Edge Computing: A Comprehensive Survey on Optimizing Large Models and AI Agents for Pervasive Deployment
di: Wang, Xubin, et al.
Pubblicazione: (2025)
di: Wang, Xubin, et al.
Pubblicazione: (2025)
Carbon Intensity-Aware Adaptive Inference of DNNs
di: Jung, Jiwan
Pubblicazione: (2024)
di: Jung, Jiwan
Pubblicazione: (2024)
CEAR: Certified Ensemble Adversarial Robustness in DNNs
di: Sadig, Daniel, et al.
Pubblicazione: (2026)
di: Sadig, Daniel, et al.
Pubblicazione: (2026)
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
di: Zhang, Tuo, et al.
Pubblicazione: (2025)
di: Zhang, Tuo, et al.
Pubblicazione: (2025)
DTMM: Deploying TinyML Models on Extremely Weak IoT Devices with Pruning
di: Han, Lixiang, et al.
Pubblicazione: (2024)
di: Han, Lixiang, et al.
Pubblicazione: (2024)
Collaborative Compression for Large-Scale MoE Deployment on Edge
di: Chen, Yixiao, et al.
Pubblicazione: (2025)
di: Chen, Yixiao, et al.
Pubblicazione: (2025)
Balanced Edge Pruning for Graph Anomaly Detection with Noisy Labels
di: Wang, Zhu, et al.
Pubblicazione: (2024)
di: Wang, Zhu, et al.
Pubblicazione: (2024)
How DNNs break the Curse of Dimensionality: Compositionality and Symmetry Learning
di: Jacot, Arthur, et al.
Pubblicazione: (2024)
di: Jacot, Arthur, et al.
Pubblicazione: (2024)
Resource-Efficient Generative AI Model Deployment in Mobile Edge Networks
di: Liang, Yuxin, et al.
Pubblicazione: (2024)
di: Liang, Yuxin, et al.
Pubblicazione: (2024)
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices
di: Qin, Ruiyang, et al.
Pubblicazione: (2024)
di: Qin, Ruiyang, et al.
Pubblicazione: (2024)
Spectral Theory for Edge Pruning in Asynchronous Recurrent Graph Neural Networks
di: Bessone, Nicolas
Pubblicazione: (2025)
di: Bessone, Nicolas
Pubblicazione: (2025)
IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning
di: Jung, Jaeheun, et al.
Pubblicazione: (2025)
di: Jung, Jaeheun, et al.
Pubblicazione: (2025)
Bit-Identical Medical Deep Learning via Structured Orthogonal Initialization
di: Shkolnikov, Yakov Pyotr
Pubblicazione: (2026)
di: Shkolnikov, Yakov Pyotr
Pubblicazione: (2026)
On Background Bias of Post-Hoc Concept Embeddings in Computer Vision DNNs
di: Schwalbe, Gesina, et al.
Pubblicazione: (2025)
di: Schwalbe, Gesina, et al.
Pubblicazione: (2025)
ECQ$^{\text{x}}$: Explainability-Driven Quantization for Low-Bit and Sparse DNNs
di: Becking, Daniel, et al.
Pubblicazione: (2021)
di: Becking, Daniel, et al.
Pubblicazione: (2021)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
Efficient Triple Modular Redundancy for Reliability Enhancement of DNNs Using Explainable AI
di: Soroush, Kimia, et al.
Pubblicazione: (2025)
di: Soroush, Kimia, et al.
Pubblicazione: (2025)
Relationship between Uncertainty in DNNs and Adversarial Attacks
di: Ogonna, Mabel, et al.
Pubblicazione: (2024)
di: Ogonna, Mabel, et al.
Pubblicazione: (2024)
Random weights of DNNs and emergence of fixed points
di: Berlyand, L., et al.
Pubblicazione: (2025)
di: Berlyand, L., et al.
Pubblicazione: (2025)
Achieving Pareto Optimality using Efficient Parameter Reduction for DNNs in Resource-Constrained Edge Environment
di: Mih, Atah Nuh, et al.
Pubblicazione: (2024)
di: Mih, Atah Nuh, et al.
Pubblicazione: (2024)
Co-Designing Binarized Transformer and Hardware Accelerator for Efficient End-to-End Edge Deployment
di: Ji, Yuhao, et al.
Pubblicazione: (2024)
di: Ji, Yuhao, et al.
Pubblicazione: (2024)
Graph Neural Networks Automated Design and Deployment on Device-Edge Co-Inference Systems
di: Zhou, Ao, et al.
Pubblicazione: (2024)
di: Zhou, Ao, et al.
Pubblicazione: (2024)
Structured vs. Unstructured Pruning: An Exponential Gap
di: Ferre', Davide, et al.
Pubblicazione: (2026)
di: Ferre', Davide, et al.
Pubblicazione: (2026)
Less is More: Unlocking Specialization of Time Series Foundation Models via Structured Pruning
di: Zhao, Lifan, et al.
Pubblicazione: (2025)
di: Zhao, Lifan, et al.
Pubblicazione: (2025)
StructPrune: Structured Global Pruning asymptotics with $\mathcal{O}(\sqrt{N})$ GPU Memory
di: Song, Xinyuan, et al.
Pubblicazione: (2025)
di: Song, Xinyuan, et al.
Pubblicazione: (2025)
DapperFL: Domain Adaptive Federated Learning with Model Fusion Pruning for Edge Devices
di: Jia, Yongzhe, et al.
Pubblicazione: (2024)
di: Jia, Yongzhe, et al.
Pubblicazione: (2024)
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
di: Roy, Arjun, et al.
Pubblicazione: (2026)
di: Roy, Arjun, et al.
Pubblicazione: (2026)
Edge-free but Structure-aware: Prototype-Guided Knowledge Distillation from GNNs to MLPs
di: Wu, Taiqiang, et al.
Pubblicazione: (2023)
di: Wu, Taiqiang, et al.
Pubblicazione: (2023)
HAWX: A Hardware-Aware FrameWork for Fast and Scalable ApproXimation of DNNs
di: Nazari, Samira, et al.
Pubblicazione: (2026)
di: Nazari, Samira, et al.
Pubblicazione: (2026)
NeuroFlux: Memory-Efficient CNN Training Using Adaptive Local Learning
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2024)
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2024)
SPAP: Structured Pruning via Alternating Optimization and Penalty Methods
di: Hu, Hanyu, et al.
Pubblicazione: (2025)
di: Hu, Hanyu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Mosaic: Composite Projection Pruning for Resource-efficient LLMs
di: Eccles, Bailey J., et al.
Pubblicazione: (2025) -
DNNShifter: An Efficient DNN Pruning System for Edge Computing
di: Eccles, Bailey J., et al.
Pubblicazione: (2023) -
Data-Free Pruning of Self-Attention Layers in LLMs
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2025) -
Signal Collapse in One-Shot Pruning: When Sparse Models Fail to Distinguish Neural Representations
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2025) -
DRIVE: Dual Gradient-Based Rapid Iterative Pruning
di: Saikumar, Dhananjay, et al.
Pubblicazione: (2024)