Training Multi-Bit Quantized and Binarized Networks with a Learnable Symmetric Quantizer
Quantizing weights and activations of deep neural networks is essential for deploying them in resource-constrained devices, or cloud platforms for at-scale services. While binarization is a special case of quantization, this extreme case often leads to several training difficulties, and necessitates...
| Published in: | IEEE Access |
|---|---|
| Main Authors: | , , |
| Format: | Article |
| Language: | English |
| Published: |
IEEE
2021-01-01
|
| Subjects: | |
| Online Access: | https://ieeexplore.ieee.org/document/9383003/ |
