Exploiting the Partly Scratch-off Lottery Ticket for Quantization-Aware Training

2022-11-12 06:11:36

Yunshan Zhong, Mingbao Lin, Yuxin Zhang, Gongrui Nan, Fei Chao, Rongrong Ji

arXiv_CV

arXiv_CV QA Pose Quantization

Abstract
Abstract (translated)
URL
PDF

Abstract

Quantization-aware training (QAT) receives extensive popularity as it well retains the performance of quantized networks. In QAT, the contemporary experience is that all quantized weights are updated for an entire training process. In this paper, this experience is challenged based on an interesting phenomenon we observed. Specifically, a large portion of quantized weights reaches the optimal quantization level after a few training epochs, which we refer to as the partly scratch-off lottery ticket. This straightforward-yet-valuable observation naturally inspires us to zero out gradient calculations of these weights in the remaining training period to avoid meaningless updating. To effectively find the ticket, we develop a heuristic method, dubbed as lottery ticket scratcher (LTS), which freezes a weight once the distance between the full-precision one and its quantization level is smaller than a controllable threshold. Surprisingly, the proposed LTS typically eliminates 30\%-60\% weight updating and 15\%-30\% FLOPs of the backward pass, while still resulting on par with or even better performance than the compared baseline. For example, compared with the baseline, LTS improves 2-bit ResNet-18 by 1.41\%, eliminating 56\% weight updating and 28\% FLOPs of the backward pass.

Abstract (translated)

URL

https://arxiv.org/abs/2211.08544

PDF

https://arxiv.org/pdf/2211.08544.pdf