Saved in:
| Main Authors: | Jiang, Mengnan, Wang, Jingcun, Eldebiky, Amro, Yin, Xunzhao, Zhuo, Cheng, Lin, Ing-Chao, Zhang, Grace Li |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2312.05875 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BasisN: Reprogramming-Free RRAM-Based In-Memory-Computing by Basis Combination for Deep Neural Networks
by: Eldebiky, Amro, et al.
Published: (2024)
by: Eldebiky, Amro, et al.
Published: (2024)
KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs
by: Chen, Chuangtao, et al.
Published: (2026)
by: Chen, Chuangtao, et al.
Published: (2026)
LiveMind: Low-latency Large Language Models with Simultaneous Inference
by: Chen, Chuangtao, et al.
Published: (2024)
by: Chen, Chuangtao, et al.
Published: (2024)
Classification-Based Automatic HDL Code Generation Using LLMs
by: Sun, Wenhao, et al.
Published: (2024)
by: Sun, Wenhao, et al.
Published: (2024)
An Efficient General-Purpose Optical Accelerator for Neural Networks
by: Fei, Sijie, et al.
Published: (2024)
by: Fei, Sijie, et al.
Published: (2024)
Early-Exit with Class Exclusion for Efficient Inference of Neural Networks
by: Wang, Jingcun, et al.
Published: (2023)
by: Wang, Jingcun, et al.
Published: (2023)
CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation
by: Lin, Xiaolin, et al.
Published: (2025)
by: Lin, Xiaolin, et al.
Published: (2025)
FactorHD: A Hyperdimensional Computing Model for Multi-Object Multi-Class Representation and Factorization
by: Zhou, Yifei, et al.
Published: (2025)
by: Zhou, Yifei, et al.
Published: (2025)
Basis Sharing: Cross-Layer Parameter Sharing for Large Language Model Compression
by: Wang, Jingcun, et al.
Published: (2024)
by: Wang, Jingcun, et al.
Published: (2024)
Attribution Explanations for Deep Neural Networks: A Theoretical Perspective
by: Deng, Huiqi, et al.
Published: (2025)
by: Deng, Huiqi, et al.
Published: (2025)
Orthogonal Soft Pruning for Efficient Class Unlearning
by: Gong, Qinghui, et al.
Published: (2025)
by: Gong, Qinghui, et al.
Published: (2025)
Graph Neural Networks with Feature and Structure Aware Random Walk
by: Zhuo, Wei, et al.
Published: (2021)
by: Zhuo, Wei, et al.
Published: (2021)
C-SWAP: Explainability-Aware Structured Pruning for Efficient Neural Networks Compression
by: Bauvin, Baptiste, et al.
Published: (2025)
by: Bauvin, Baptiste, et al.
Published: (2025)
Enhanced Structured Lasso Pruning with Class-wise Information
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Resource-Aware Neural Network Pruning Using Graph-based Reinforcement Learning
by: Balemans, Dieter, et al.
Published: (2025)
by: Balemans, Dieter, et al.
Published: (2025)
PruneSymNet: A Symbolic Neural Network and Pruning Algorithm for Symbolic Regression
by: Wu, Min, et al.
Published: (2024)
by: Wu, Min, et al.
Published: (2024)
LANS: A Layout-Aware Neural Solver for Plane Geometry Problem
by: Li, Zhong-Zhi, et al.
Published: (2023)
by: Li, Zhong-Zhi, et al.
Published: (2023)
ATP: Adaptive Threshold Pruning for Efficient Data Encoding in Quantum Neural Networks
by: Afane, Mohamed, et al.
Published: (2025)
by: Afane, Mohamed, et al.
Published: (2025)
Efficient Training of Spiking Neural Networks by Spike-aware Data Pruning
by: Ma, Chenxiang, et al.
Published: (2025)
by: Ma, Chenxiang, et al.
Published: (2025)
SepPrune: Structured Pruning for Efficient Deep Speech Separation
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Towards Efficient Deep Spiking Neural Networks Construction with Spiking Activity based Pruning
by: Li, Yaxin, et al.
Published: (2024)
by: Li, Yaxin, et al.
Published: (2024)
An Efficient Sparse Fine-Tuning with Low Quantization Error via Neural Network Pruning
by: Li, Cen-Jhih, et al.
Published: (2025)
by: Li, Cen-Jhih, et al.
Published: (2025)
PrunePEFT: Iterative Hybrid Pruning for Parameter-Efficient Fine-tuning of LLMs
by: Yu, Tongzhou, et al.
Published: (2025)
by: Yu, Tongzhou, et al.
Published: (2025)
EncodingNet: A Novel Encoding-based MAC Design for Efficient Neural Network Acceleration
by: Liu, Bo, et al.
Published: (2024)
by: Liu, Bo, et al.
Published: (2024)
PDTrim: Targeted Pruning for Prefill-Decode Disaggregation in Inference
by: Zhang, Hao, et al.
Published: (2025)
by: Zhang, Hao, et al.
Published: (2025)
Prune-Quantize-Distill: An Ordered Pipeline for Efficient Neural Network Compression
by: Zhou, Longsheng, et al.
Published: (2026)
by: Zhou, Longsheng, et al.
Published: (2026)
SAGMAN: Stability Analysis of Graph Neural Networks on the Manifolds
by: Cheng, Wuxinlin, et al.
Published: (2024)
by: Cheng, Wuxinlin, et al.
Published: (2024)
Temporal Aware Pruning for Efficient Diffusion-based Video Generation
by: Li, Sheng, et al.
Published: (2026)
by: Li, Sheng, et al.
Published: (2026)
Physics-informed Attention-enhanced Fourier Neural Operator for Solar Magnetic Field Extrapolations
by: Cao, Jinghao, et al.
Published: (2025)
by: Cao, Jinghao, et al.
Published: (2025)
Pruning as a Game: Equilibrium-Driven Sparsification of Neural Networks
by: Shah, Zubair, et al.
Published: (2025)
by: Shah, Zubair, et al.
Published: (2025)
Efficient Single-Step Framework for Incremental Class Learning in Neural Networks
by: Dopico-Castro, Alejandro, et al.
Published: (2025)
by: Dopico-Castro, Alejandro, et al.
Published: (2025)
Small Contributions, Small Networks: Efficient Neural Network Pruning Based on Relative Importance
by: Hussien, Mostafa, et al.
Published: (2024)
by: Hussien, Mostafa, et al.
Published: (2024)
Unveiling Project-Specific Bias in Neural Code Models
by: Li, Zhiming, et al.
Published: (2022)
by: Li, Zhiming, et al.
Published: (2022)
QAPruner: Quantization-Aware Vision Token Pruning for Multimodal Large Language Models
by: Wang, Xinhao, et al.
Published: (2026)
by: Wang, Xinhao, et al.
Published: (2026)
HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space
by: Li, Ke, et al.
Published: (2025)
by: Li, Ke, et al.
Published: (2025)
Commute Graph Neural Networks
by: Zhuo, Wei, et al.
Published: (2024)
by: Zhuo, Wei, et al.
Published: (2024)
SWAP: Sparse Entropic Wasserstein Regression for Robust Network Pruning
by: You, Lei, et al.
Published: (2023)
by: You, Lei, et al.
Published: (2023)
FoPru: Focal Pruning for Efficient Large Vision-Language Models
by: Jiang, Lei, et al.
Published: (2024)
by: Jiang, Lei, et al.
Published: (2024)
CLASP: Class-Adaptive Layer Fusion and Dual-Stage Pruning for Multimodal Large Language Models
by: Dang, Yunkai, et al.
Published: (2026)
by: Dang, Yunkai, et al.
Published: (2026)
Application-Specific Component-Aware Structured Pruning of Deep Neural Networks in Control via Soft Coefficient Optimization
by: Sundaram, Ganesh, et al.
Published: (2025)
by: Sundaram, Ganesh, et al.
Published: (2025)
Similar Items
-
BasisN: Reprogramming-Free RRAM-Based In-Memory-Computing by Basis Combination for Deep Neural Networks
by: Eldebiky, Amro, et al.
Published: (2024) -
KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs
by: Chen, Chuangtao, et al.
Published: (2026) -
LiveMind: Low-latency Large Language Models with Simultaneous Inference
by: Chen, Chuangtao, et al.
Published: (2024) -
Classification-Based Automatic HDL Code Generation Using LLMs
by: Sun, Wenhao, et al.
Published: (2024) -
An Efficient General-Purpose Optical Accelerator for Neural Networks
by: Fei, Sijie, et al.
Published: (2024)