Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Jiang, Jiacheng, Meng, Yuan, Tang, Chen, Yu, Han, Li, Qun, Wang, Zhi, Zhu, Wenwu
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911142352257024
author Jiang, Jiacheng
Meng, Yuan
Tang, Chen
Yu, Han
Li, Qun
Wang, Zhi
Zhu, Wenwu
author_facet Jiang, Jiacheng
Meng, Yuan
Tang, Chen
Yu, Han
Li, Qun
Wang, Zhi
Zhu, Wenwu
contents Current quantization-aware training (QAT) methods primarily focus on enhancing the performance of quantized models on in-distribution (I.D) data, while overlooking the potential performance degradation on out-of-distribution (OOD) data. In this paper, we first substantiate this problem through rigorous experiment, showing that QAT can lead to a significant OOD generalization performance degradation. Further, we find the contradiction between the perspective that flatness of loss landscape gives rise to superior OOD generalization and the phenomenon that QAT lead to a sharp loss landscape, can cause the above problem. Therefore, we propose a flatness-oriented QAT method, FQAT, to achieve generalizable QAT. Specifically, i) FQAT introduces a layer-wise freezing mechanism to mitigate the gradient conflict issue between dual optimization objectives (i.e., vanilla QAT and flatness). ii) FQAT proposes an disorder-guided adaptive freezing algorithm to dynamically determines which layers to freeze at each training step, effectively addressing the challenges caused by interference between layers. A gradient disorder metric is designed to help the algorithm identify unstable layers during training. Extensive experiments on influential OOD benchmark demonstrate the superiority of our method over state-of-the-art baselines under both I.D and OOD image classification tasks.
format Preprint
id arxiv_https___arxiv_org_abs_2509_00859
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
Jiang, Jiacheng
Meng, Yuan
Tang, Chen
Yu, Han
Li, Qun
Wang, Zhi
Zhu, Wenwu
Computer Vision and Pattern Recognition
Current quantization-aware training (QAT) methods primarily focus on enhancing the performance of quantized models on in-distribution (I.D) data, while overlooking the potential performance degradation on out-of-distribution (OOD) data. In this paper, we first substantiate this problem through rigorous experiment, showing that QAT can lead to a significant OOD generalization performance degradation. Further, we find the contradiction between the perspective that flatness of loss landscape gives rise to superior OOD generalization and the phenomenon that QAT lead to a sharp loss landscape, can cause the above problem. Therefore, we propose a flatness-oriented QAT method, FQAT, to achieve generalizable QAT. Specifically, i) FQAT introduces a layer-wise freezing mechanism to mitigate the gradient conflict issue between dual optimization objectives (i.e., vanilla QAT and flatness). ii) FQAT proposes an disorder-guided adaptive freezing algorithm to dynamically determines which layers to freeze at each training step, effectively addressing the challenges caused by interference between layers. A gradient disorder metric is designed to help the algorithm identify unstable layers during training. Extensive experiments on influential OOD benchmark demonstrate the superiority of our method over state-of-the-art baselines under both I.D and OOD image classification tasks.
title Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2509.00859