Quantization-Aware Regularizers for Deep Neural Networks Compression
Fuente:
arXiv
Guardado en:
| Autores principales: | Malchiodi, Dario, Ferraretto, Mattia, Frasca, Marco |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Support Vector Based Anomaly Detection in Federated Learning
por: Frasson, Massimo, et al.
Publicado: (2024)
por: Frasson, Massimo, et al.
Publicado: (2024)
Resource-Limited Automated Ki67 Index Estimation in Breast Cancer
por: Gliozzo, J., et al.
Publicado: (2023)
por: Gliozzo, J., et al.
Publicado: (2023)
Exploring Quantization and Mapping Synergy in Hardware-Aware Deep Neural Network Accelerators
por: Klhufek, Jan, et al.
Publicado: (2024)
por: Klhufek, Jan, et al.
Publicado: (2024)
Saliency-Aware Regularized Graph Neural Network
por: Pei, Wenjie, et al.
Publicado: (2024)
por: Pei, Wenjie, et al.
Publicado: (2024)
Neural Image Compression with Quantization Rectifier
por: Luo, Wei, et al.
Publicado: (2024)
por: Luo, Wei, et al.
Publicado: (2024)
Invariant Representations with Stochastically Quantized Neural Networks
por: Cerrato, Mattia, et al.
Publicado: (2022)
por: Cerrato, Mattia, et al.
Publicado: (2022)
Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks
por: Zou, Jinping, et al.
Publicado: (2024)
por: Zou, Jinping, et al.
Publicado: (2024)
Adaptive Tabu Dropout for Regularization of Deep Neural Network
por: Hasan, Md. Tarek, et al.
Publicado: (2024)
por: Hasan, Md. Tarek, et al.
Publicado: (2024)
Lattice-based Deep Neural Networks: Regularity and Tailored Regularization
por: Keller, Alexander, et al.
Publicado: (2026)
por: Keller, Alexander, et al.
Publicado: (2026)
Regularization-based Framework for Quantization-, Fault- and Variability-Aware Training
por: Biswas, Anmol, et al.
Publicado: (2025)
por: Biswas, Anmol, et al.
Publicado: (2025)
ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression
por: Yu, Wneya, et al.
Publicado: (2026)
por: Yu, Wneya, et al.
Publicado: (2026)
Activation Compression of Graph Neural Networks using Block-wise Quantization with Improved Variance Minimization
por: Eliassen, Sebastian, et al.
Publicado: (2023)
por: Eliassen, Sebastian, et al.
Publicado: (2023)
Learning to Compress: Local Rank and Information Compression in Deep Neural Networks
por: Patel, Niket, et al.
Publicado: (2024)
por: Patel, Niket, et al.
Publicado: (2024)
Prune-Quantize-Distill: An Ordered Pipeline for Efficient Neural Network Compression
por: Zhou, Longsheng, et al.
Publicado: (2026)
por: Zhou, Longsheng, et al.
Publicado: (2026)
Frequency Regularization: Unveiling the Spectral Inductive Bias of Deep Neural Networks
por: Lu, Jiahao
Publicado: (2025)
por: Lu, Jiahao
Publicado: (2025)
Saliency-Aware Regularized Quantization Calibration for Large Language Models
por: Zhao, Yanlong, et al.
Publicado: (2026)
por: Zhao, Yanlong, et al.
Publicado: (2026)
Compressing Deep Neural Networks Using Explainable AI
por: Soroush, Kimia, et al.
Publicado: (2025)
por: Soroush, Kimia, et al.
Publicado: (2025)
Inshrinkerator: Compressing Deep Learning Training Checkpoints via Dynamic Quantization
por: Agrawal, Amey, et al.
Publicado: (2023)
por: Agrawal, Amey, et al.
Publicado: (2023)
TinyM$^2$Net-V3: Memory-Aware Compressed Multimodal Deep Neural Networks for Sustainable Edge Deployment
por: Rashid, Hasib-Al, et al.
Publicado: (2024)
por: Rashid, Hasib-Al, et al.
Publicado: (2024)
DQA: An Efficient Method for Deep Quantization of Deep Neural Network Activations
por: Hu, Wenhao, et al.
Publicado: (2024)
por: Hu, Wenhao, et al.
Publicado: (2024)
Deep Neural Networks are Adaptive to Function Regularity and Data Distribution in Approximation and Estimation
por: Liu, Hao, et al.
Publicado: (2024)
por: Liu, Hao, et al.
Publicado: (2024)
Joint Pruning and Channel-wise Mixed-Precision Quantization for Efficient Deep Neural Networks
por: Motetti, Beatrice Alessandra, et al.
Publicado: (2024)
por: Motetti, Beatrice Alessandra, et al.
Publicado: (2024)
Low-bit Model Quantization for Deep Neural Networks: A Survey
por: Liu, Kai, et al.
Publicado: (2025)
por: Liu, Kai, et al.
Publicado: (2025)
On-Chip Hardware-Aware Quantization for Mixed Precision Neural Networks
por: Huang, Wei, et al.
Publicado: (2023)
por: Huang, Wei, et al.
Publicado: (2023)
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers
por: Súkeník, Peter, et al.
Publicado: (2025)
por: Súkeník, Peter, et al.
Publicado: (2025)
"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach
por: Gu, Lingyu, et al.
Publicado: (2024)
por: Gu, Lingyu, et al.
Publicado: (2024)
Q-SENN: Quantized Self-Explaining Neural Networks
por: Norrenbrock, Thomas, et al.
Publicado: (2023)
por: Norrenbrock, Thomas, et al.
Publicado: (2023)
SQUAT: Stateful Quantization-Aware Training in Recurrent Spiking Neural Networks
por: Venkatesh, Sreyes, et al.
Publicado: (2024)
por: Venkatesh, Sreyes, et al.
Publicado: (2024)
An Efficient Compression of Deep Neural Network Checkpoints Based on Prediction and Context Modeling
por: Kim, Yuriy, et al.
Publicado: (2025)
por: Kim, Yuriy, et al.
Publicado: (2025)
Theoretical Guarantees for Low-Rank Compression of Deep Neural Networks
por: Zhang, Shihao, et al.
Publicado: (2025)
por: Zhang, Shihao, et al.
Publicado: (2025)
DAQ: Delta-Aware Quantization for Post-Training LLM Weight Compression
por: Yu, Xiaoming, et al.
Publicado: (2026)
por: Yu, Xiaoming, et al.
Publicado: (2026)
Adaptive Gradient Regularization: A Faster and Generalizable Optimization Technique for Deep Neural Networks
por: Jiang, Huixiu, et al.
Publicado: (2024)
por: Jiang, Huixiu, et al.
Publicado: (2024)
Pruning Deep Neural Networks via a Combination of the Marchenko-Pastur Distribution and Regularization
por: Berlyand, Leonid, et al.
Publicado: (2025)
por: Berlyand, Leonid, et al.
Publicado: (2025)
Frame Quantization of Neural Networks
por: Czaja, Wojciech, et al.
Publicado: (2024)
por: Czaja, Wojciech, et al.
Publicado: (2024)
Regularized Gradient Clipping Provably Trains Wide and Deep Neural Networks
por: Tucat, Matteo, et al.
Publicado: (2024)
por: Tucat, Matteo, et al.
Publicado: (2024)
Inverse Evolution Layers: Physics-informed Regularizers for Deep Neural Networks
por: Liu, Chaoyu, et al.
Publicado: (2023)
por: Liu, Chaoyu, et al.
Publicado: (2023)
Verifying Quantized Graph Neural Networks is PSPACE-complete
por: Sälzer, Marco, et al.
Publicado: (2025)
por: Sälzer, Marco, et al.
Publicado: (2025)
On-Device Training of Fully Quantized Deep Neural Networks on Cortex-M Microcontrollers
por: Deutel, Mark, et al.
Publicado: (2024)
por: Deutel, Mark, et al.
Publicado: (2024)
Quantization of Spiking Neural Networks Beyond Accuracy
por: Smith, Evan Gibson, et al.
Publicado: (2026)
por: Smith, Evan Gibson, et al.
Publicado: (2026)
Pruning and Quantization Impact on Graph Neural Networks
por: Khedri, Khatoon, et al.
Publicado: (2025)
por: Khedri, Khatoon, et al.
Publicado: (2025)
Ejemplares similares
-
Support Vector Based Anomaly Detection in Federated Learning
por: Frasson, Massimo, et al.
Publicado: (2024) -
Resource-Limited Automated Ki67 Index Estimation in Breast Cancer
por: Gliozzo, J., et al.
Publicado: (2023) -
Exploring Quantization and Mapping Synergy in Hardware-Aware Deep Neural Network Accelerators
por: Klhufek, Jan, et al.
Publicado: (2024) -
Saliency-Aware Regularized Graph Neural Network
por: Pei, Wenjie, et al.
Publicado: (2024) -
Neural Image Compression with Quantization Rectifier
por: Luo, Wei, et al.
Publicado: (2024)