Error Diffusion: Post Training Quantization with Block-Scaled Number Formats for Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Khodamoradi, Alireza, Denolf, Kristof, Dellinger, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization
by: Bouquet, Yann, et al.
Published: (2026)
by: Bouquet, Yann, et al.
Published: (2026)
AdaHOP: Fast and Accurate Low-Precision Training via Outlier-Pattern-Aware Rotation
by: Kim, Seonggon, et al.
Published: (2026)
by: Kim, Seonggon, et al.
Published: (2026)
Pushing the Limits of Block Rotations in Post-Training Quantization
by: Sanjeet, Sai, et al.
Published: (2026)
by: Sanjeet, Sai, et al.
Published: (2026)
Quantization Error Propagation: Revisiting Layer-Wise Post-Training Quantization
by: Arai, Yamato, et al.
Published: (2025)
by: Arai, Yamato, et al.
Published: (2025)
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
by: Shing, Makoto, et al.
Published: (2025)
by: Shing, Makoto, et al.
Published: (2025)
TesseraQ: Ultra Low-Bit LLM Post-Training Quantization with Block Reconstruction
by: Li, Yuhang, et al.
Published: (2024)
by: Li, Yuhang, et al.
Published: (2024)
Post Training Quantization of Large Language Models with Microscaling Formats
by: Sharify, Sayeh, et al.
Published: (2024)
by: Sharify, Sayeh, et al.
Published: (2024)
Rethinking Post-Training Quantization: Introducing a Statistical Pre-Calibration Approach
by: Ghaffari, Alireza, et al.
Published: (2025)
by: Ghaffari, Alireza, et al.
Published: (2025)
Interactions Across Blocks in Post-Training Quantization of Large Language Models
by: Shabanovi, Khasmamad, et al.
Published: (2024)
by: Shabanovi, Khasmamad, et al.
Published: (2024)
A Diagnostic Evaluation of Neural Networks Trained with the Error Diffusion Learning Algorithm
by: Fujita, Kazuhisa
Published: (2025)
by: Fujita, Kazuhisa
Published: (2025)
Scaling Laws for Post Training Quantized Large Language Models
by: Xu, Zifei, et al.
Published: (2024)
by: Xu, Zifei, et al.
Published: (2024)
Training Dynamics Impact Post-Training Quantization Robustness
by: Catalan-Tatjer, Albert, et al.
Published: (2025)
by: Catalan-Tatjer, Albert, et al.
Published: (2025)
Towards Next-Level Post-Training Quantization of Hyper-Scale Transformers
by: Kim, Junhan, et al.
Published: (2024)
by: Kim, Junhan, et al.
Published: (2024)
The Quantization Model of Neural Scaling
by: Michaud, Eric J., et al.
Published: (2023)
by: Michaud, Eric J., et al.
Published: (2023)
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
by: Zhang, Jingxuan, et al.
Published: (2026)
by: Zhang, Jingxuan, et al.
Published: (2026)
Gradient-Aligned Calibration for Post-Training Quantization of Diffusion Models
by: Hoang, Dung Anh, et al.
Published: (2026)
by: Hoang, Dung Anh, et al.
Published: (2026)
Gradient-Free Training of Quantized Neural Networks
by: Cohen, Noa, et al.
Published: (2024)
by: Cohen, Noa, et al.
Published: (2024)
SKIM: Any-bit Quantization Pushing The Limits of Post-Training Quantization
by: Bai, Runsheng, et al.
Published: (2024)
by: Bai, Runsheng, et al.
Published: (2024)
DyBit: Dynamic Bit-Precision Numbers for Efficient Quantized Neural Network Inference
by: Zhou, Jiajun, et al.
Published: (2023)
by: Zhou, Jiajun, et al.
Published: (2023)
Activation Compression of Graph Neural Networks using Block-wise Quantization with Improved Variance Minimization
by: Eliassen, Sebastian, et al.
Published: (2023)
by: Eliassen, Sebastian, et al.
Published: (2023)
Z-Error Loss for Training Neural Networks
by: Godin, Guillaume
Published: (2025)
by: Godin, Guillaume
Published: (2025)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
by: Lee, Dongyeun, et al.
Published: (2025)
by: Lee, Dongyeun, et al.
Published: (2025)
Improving Quantization-aware Training of Low-Precision Network via Block Replacement on Full-Precision Counterpart
by: Yu, Chengting, et al.
Published: (2024)
by: Yu, Chengting, et al.
Published: (2024)
Data Generation for Hardware-Friendly Post-Training Quantization
by: Dikstein, Lior, et al.
Published: (2024)
by: Dikstein, Lior, et al.
Published: (2024)
Training Neural Networks at Any Scale
by: Pethick, Thomas, et al.
Published: (2025)
by: Pethick, Thomas, et al.
Published: (2025)
RepQuant: Towards Accurate Post-Training Quantization of Large Transformer Models via Scale Reparameterization
by: Li, Zhikai, et al.
Published: (2024)
by: Li, Zhikai, et al.
Published: (2024)
EDA-DM: Enhanced Distribution Alignment for Post-Training Quantization of Diffusion Models
by: Liu, Xuewen, et al.
Published: (2024)
by: Liu, Xuewen, et al.
Published: (2024)
BlockDialect: Block-wise Fine-grained Mixed Format Quantization for Energy-Efficient LLM Inference
by: Jang, Wonsuk, et al.
Published: (2025)
by: Jang, Wonsuk, et al.
Published: (2025)
Understanding the Difficulty of Low-Precision Post-Training Quantization for LLMs
by: Xu, Zifei, et al.
Published: (2024)
by: Xu, Zifei, et al.
Published: (2024)
MatGPTQ: Accurate and Efficient Post-Training Matryoshka Quantization
by: Kleinegger, Maximilian, et al.
Published: (2026)
by: Kleinegger, Maximilian, et al.
Published: (2026)
Assessing the Potential for Catastrophic Failure in Dynamic Post-Training Quantization
by: Frank, Logan, et al.
Published: (2025)
by: Frank, Logan, et al.
Published: (2025)
PreNeT: Leveraging Computational Features to Predict Deep Neural Network Training Time
by: Pourali, Alireza, et al.
Published: (2024)
by: Pourali, Alireza, et al.
Published: (2024)
MCEL: Margin-Based Cross-Entropy Loss for Error-Tolerant Quantized Neural Networks
by: Yayla, Mikail, et al.
Published: (2026)
by: Yayla, Mikail, et al.
Published: (2026)
Efficient Post-training Quantization with FP8 Formats
by: Shen, Haihao, et al.
Published: (2023)
by: Shen, Haihao, et al.
Published: (2023)
Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models
by: Zhou, Chenxi, et al.
Published: (2025)
by: Zhou, Chenxi, et al.
Published: (2025)
Beacon: Post-Training Quantization with Integrated Grid Selection
by: Zhang, Shihao, et al.
Published: (2025)
by: Zhang, Shihao, et al.
Published: (2025)
Scaling Law for Quantization-Aware Training
by: Chen, Mengzhao, et al.
Published: (2025)
by: Chen, Mengzhao, et al.
Published: (2025)
ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models
by: Lucas, Ryan, et al.
Published: (2026)
by: Lucas, Ryan, et al.
Published: (2026)
Quantized and Interpretable Learning Scheme for Deep Neural Networks in Classification Task
by: Maleki, Alireza, et al.
Published: (2024)
by: Maleki, Alireza, et al.
Published: (2024)
The Computational Complexity of Almost Stable Clustering with Penalties
by: Khodamoradi, Kamyar, et al.
Published: (2025)
by: Khodamoradi, Kamyar, et al.
Published: (2025)
Similar Items
-
LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization
by: Bouquet, Yann, et al.
Published: (2026) -
AdaHOP: Fast and Accurate Low-Precision Training via Outlier-Pattern-Aware Rotation
by: Kim, Seonggon, et al.
Published: (2026) -
Pushing the Limits of Block Rotations in Post-Training Quantization
by: Sanjeet, Sai, et al.
Published: (2026) -
Quantization Error Propagation: Revisiting Layer-Wise Post-Training Quantization
by: Arai, Yamato, et al.
Published: (2025) -
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
by: Shing, Makoto, et al.
Published: (2025)