Quantization-Aware Regularizers for Deep Neural Networks Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Malchiodi, Dario, Ferraretto, Mattia, Frasca, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Support Vector Based Anomaly Detection in Federated Learning
by: Frasson, Massimo, et al.
Published: (2024)
by: Frasson, Massimo, et al.
Published: (2024)
Resource-Limited Automated Ki67 Index Estimation in Breast Cancer
by: Gliozzo, J., et al.
Published: (2023)
by: Gliozzo, J., et al.
Published: (2023)
Exploring Quantization and Mapping Synergy in Hardware-Aware Deep Neural Network Accelerators
by: Klhufek, Jan, et al.
Published: (2024)
by: Klhufek, Jan, et al.
Published: (2024)
Saliency-Aware Regularized Graph Neural Network
by: Pei, Wenjie, et al.
Published: (2024)
by: Pei, Wenjie, et al.
Published: (2024)
Neural Image Compression with Quantization Rectifier
by: Luo, Wei, et al.
Published: (2024)
by: Luo, Wei, et al.
Published: (2024)
Invariant Representations with Stochastically Quantized Neural Networks
by: Cerrato, Mattia, et al.
Published: (2022)
by: Cerrato, Mattia, et al.
Published: (2022)
Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks
by: Zou, Jinping, et al.
Published: (2024)
by: Zou, Jinping, et al.
Published: (2024)
Adaptive Tabu Dropout for Regularization of Deep Neural Network
by: Hasan, Md. Tarek, et al.
Published: (2024)
by: Hasan, Md. Tarek, et al.
Published: (2024)
Lattice-based Deep Neural Networks: Regularity and Tailored Regularization
by: Keller, Alexander, et al.
Published: (2026)
by: Keller, Alexander, et al.
Published: (2026)
Regularization-based Framework for Quantization-, Fault- and Variability-Aware Training
by: Biswas, Anmol, et al.
Published: (2025)
by: Biswas, Anmol, et al.
Published: (2025)
ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression
by: Yu, Wneya, et al.
Published: (2026)
by: Yu, Wneya, et al.
Published: (2026)
Activation Compression of Graph Neural Networks using Block-wise Quantization with Improved Variance Minimization
by: Eliassen, Sebastian, et al.
Published: (2023)
by: Eliassen, Sebastian, et al.
Published: (2023)
Learning to Compress: Local Rank and Information Compression in Deep Neural Networks
by: Patel, Niket, et al.
Published: (2024)
by: Patel, Niket, et al.
Published: (2024)
Prune-Quantize-Distill: An Ordered Pipeline for Efficient Neural Network Compression
by: Zhou, Longsheng, et al.
Published: (2026)
by: Zhou, Longsheng, et al.
Published: (2026)
Frequency Regularization: Unveiling the Spectral Inductive Bias of Deep Neural Networks
by: Lu, Jiahao
Published: (2025)
by: Lu, Jiahao
Published: (2025)
Saliency-Aware Regularized Quantization Calibration for Large Language Models
by: Zhao, Yanlong, et al.
Published: (2026)
by: Zhao, Yanlong, et al.
Published: (2026)
Compressing Deep Neural Networks Using Explainable AI
by: Soroush, Kimia, et al.
Published: (2025)
by: Soroush, Kimia, et al.
Published: (2025)
Inshrinkerator: Compressing Deep Learning Training Checkpoints via Dynamic Quantization
by: Agrawal, Amey, et al.
Published: (2023)
by: Agrawal, Amey, et al.
Published: (2023)
TinyM$^2$Net-V3: Memory-Aware Compressed Multimodal Deep Neural Networks for Sustainable Edge Deployment
by: Rashid, Hasib-Al, et al.
Published: (2024)
by: Rashid, Hasib-Al, et al.
Published: (2024)
DQA: An Efficient Method for Deep Quantization of Deep Neural Network Activations
by: Hu, Wenhao, et al.
Published: (2024)
by: Hu, Wenhao, et al.
Published: (2024)
Deep Neural Networks are Adaptive to Function Regularity and Data Distribution in Approximation and Estimation
by: Liu, Hao, et al.
Published: (2024)
by: Liu, Hao, et al.
Published: (2024)
Joint Pruning and Channel-wise Mixed-Precision Quantization for Efficient Deep Neural Networks
by: Motetti, Beatrice Alessandra, et al.
Published: (2024)
by: Motetti, Beatrice Alessandra, et al.
Published: (2024)
Low-bit Model Quantization for Deep Neural Networks: A Survey
by: Liu, Kai, et al.
Published: (2025)
by: Liu, Kai, et al.
Published: (2025)
On-Chip Hardware-Aware Quantization for Mixed Precision Neural Networks
by: Huang, Wei, et al.
Published: (2023)
by: Huang, Wei, et al.
Published: (2023)
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers
by: Súkeník, Peter, et al.
Published: (2025)
by: Súkeník, Peter, et al.
Published: (2025)
"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach
by: Gu, Lingyu, et al.
Published: (2024)
by: Gu, Lingyu, et al.
Published: (2024)
Q-SENN: Quantized Self-Explaining Neural Networks
by: Norrenbrock, Thomas, et al.
Published: (2023)
by: Norrenbrock, Thomas, et al.
Published: (2023)
SQUAT: Stateful Quantization-Aware Training in Recurrent Spiking Neural Networks
by: Venkatesh, Sreyes, et al.
Published: (2024)
by: Venkatesh, Sreyes, et al.
Published: (2024)
An Efficient Compression of Deep Neural Network Checkpoints Based on Prediction and Context Modeling
by: Kim, Yuriy, et al.
Published: (2025)
by: Kim, Yuriy, et al.
Published: (2025)
Theoretical Guarantees for Low-Rank Compression of Deep Neural Networks
by: Zhang, Shihao, et al.
Published: (2025)
by: Zhang, Shihao, et al.
Published: (2025)
DAQ: Delta-Aware Quantization for Post-Training LLM Weight Compression
by: Yu, Xiaoming, et al.
Published: (2026)
by: Yu, Xiaoming, et al.
Published: (2026)
Adaptive Gradient Regularization: A Faster and Generalizable Optimization Technique for Deep Neural Networks
by: Jiang, Huixiu, et al.
Published: (2024)
by: Jiang, Huixiu, et al.
Published: (2024)
Pruning Deep Neural Networks via a Combination of the Marchenko-Pastur Distribution and Regularization
by: Berlyand, Leonid, et al.
Published: (2025)
by: Berlyand, Leonid, et al.
Published: (2025)
Frame Quantization of Neural Networks
by: Czaja, Wojciech, et al.
Published: (2024)
by: Czaja, Wojciech, et al.
Published: (2024)
Regularized Gradient Clipping Provably Trains Wide and Deep Neural Networks
by: Tucat, Matteo, et al.
Published: (2024)
by: Tucat, Matteo, et al.
Published: (2024)
Inverse Evolution Layers: Physics-informed Regularizers for Deep Neural Networks
by: Liu, Chaoyu, et al.
Published: (2023)
by: Liu, Chaoyu, et al.
Published: (2023)
Verifying Quantized Graph Neural Networks is PSPACE-complete
by: Sälzer, Marco, et al.
Published: (2025)
by: Sälzer, Marco, et al.
Published: (2025)
On-Device Training of Fully Quantized Deep Neural Networks on Cortex-M Microcontrollers
by: Deutel, Mark, et al.
Published: (2024)
by: Deutel, Mark, et al.
Published: (2024)
Quantization of Spiking Neural Networks Beyond Accuracy
by: Smith, Evan Gibson, et al.
Published: (2026)
by: Smith, Evan Gibson, et al.
Published: (2026)
Pruning and Quantization Impact on Graph Neural Networks
by: Khedri, Khatoon, et al.
Published: (2025)
by: Khedri, Khatoon, et al.
Published: (2025)
Similar Items
-
Support Vector Based Anomaly Detection in Federated Learning
by: Frasson, Massimo, et al.
Published: (2024) -
Resource-Limited Automated Ki67 Index Estimation in Breast Cancer
by: Gliozzo, J., et al.
Published: (2023) -
Exploring Quantization and Mapping Synergy in Hardware-Aware Deep Neural Network Accelerators
by: Klhufek, Jan, et al.
Published: (2024) -
Saliency-Aware Regularized Graph Neural Network
by: Pei, Wenjie, et al.
Published: (2024) -
Neural Image Compression with Quantization Rectifier
by: Luo, Wei, et al.
Published: (2024)