Fetching the paper…
Reading the bibliography…
Mixed-precision quantization has been widely applied on deep neural networks (DNNs) as it leads to significantly better efficiency-accuracy tradeoffs compared to uniform quantization.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng et al · 2009
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Yoshua Bengio et al · 2013
Earlier work this paper cites.
1.1 computing’s energy problem (and what we can do about it)
Mark Horowitz · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients
Shuchang Zhou et al · 2016
Earlier work this paper cites.
Xnor-net: Imagenet classification using binary convolutional neural networks
Mohammad Rastegari et al · 2016
Earlier work this paper cites.
Fengfu Li et al · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He et al · 2016
Earlier work this paper cites.
Learning sparse neural networks through l _ 0 l\_0 regularization
Christos Louizos et al · 2017
Cited alongside, same era.
Bayesian compression for deep learning
Christos Louizos et al · 2017
Cited alongside, same era.
Mobilenetv2: Inverted residuals and linear bottlenecks
Mark Sandler et al · 2018
Cited alongside, same era.
Squeeze-and-excitation networks
Jie Hu et al · 2018
Cited alongside, same era.
Pact: Parameterized clipping activation for quantized neural networks
Jungwook Choi et al · 2018
Cited alongside, same era.
Lq-nets: Learned quantization for highly accurate and compact deep neural networks
Dongqing Zhang et al · 2018
Hawq-v2: Hessian aware trace-weighted quantization of neural networks
Zhen Dong et al · 2020
Later among the works it cites.
Winning the lottery with continuous sparsification
Pedro Savarese et al · 2020
Later among the works it cites.
Growing efficient deep networks by structured continuous sparsification
Xin Yuan et al · 2020
Later among the works it cites.
Finding non-uniform quantization schemes using multi-task gaussian processes
Marcelo Gennari do Nascimento et al · 2020
Later among the works it cites.
Zeroq: A novel zero shot quantization framework
Yaohui Cai et al · 2020
Later among the works it cites.
Quanos: adversarial noise sensitivity driven hybrid quantization of neural networks
Priyadarshini Panda · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Model compression via distillation and quantization
A. Polino et al · 2018
Cited alongside, same era.
Hawq: Hessian aware quantization of neural networks with mixed-precision
Zhen Dong et al · 2019
Cited alongside, same era.
Haq: Hardware-aware automated quantization with mixed precision
Kuan Wang et al · 2019
Cited alongside, same era.
Later among the works it cites.
Bsq: Exploring bit-level sparsity for mixed-precision neural network quantization
Huanrui Yang et al · 2021
Later among the works it cites.
Hawq-v3: Dyadic neural network quantization
Zhewei Yao et al · 2021
Later among the works it cites.
Zero-shot adversarial quantization
Yuang Liu et al · 2021
Later among the works it cites.