Fetching the paper…
Reading the bibliography…
Model compression is vital to the deployment of deep learning on edge devices.
Autoq: Automated kernel-wise neural network quantization, 2019
Qian Lou, Feng Guo, Lantao Liu, Minje Kim, and Lei Jiang · 1902
Earlier work this paper cites.
Rao’s distance measure
Colin Atkinson and Ann F. S. Mitchell · 1981
Earlier work this paper cites.
Improving the convergence of back-propagation learning with second-order methods
Suzanna Becker and Yann Lecun · 1989
Earlier work this paper cites.
A stochastic estimator of the trace of the influence matrix for laplacian smoothing splines
M.F. Hutchinson · 1989
Earlier work this paper cites.
Pruning versus clipping in neural networks
Steven A. Janowsky · 1989
Earlier work this paper cites.
Natural gradient works efficiently in learning
Shun-ichi Amari · 1998
Earlier work this paper cites.
Quantization
R.M. Gray and D.L. Neuhoff · 1998
Earlier work this paper cites.
Disentangling adaptive gradient methods from learning rates, 2020
Naman Agarwal, Rohan Anil, Elad Hazan, Tomer Koren, and Cyril Zhang · 2002
Earlier work this paper cites.
Channel-wise hessian aware trace-weighted quantization of neural networks, 2020
Xu Qian, Victor Li, and Crews Darren · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
New perspectives on the natural gradient method
James Martens · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Information geometry and its applications
Shun’ichi Amari · 2016
Earlier work this paper cites.
Towards the limit of network quantization, 2016
Yoojin Choi, Mostafa El-Khamy, and Jungwon Lee · 2016
Cited alongside, same era.
The cityscapes dataset for semantic urban scene understanding
Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele · 2016
Cited alongside, same era.
Binarized neural networks
Itay Hubara, Matthieu Courbariaux, Daniel Soudry, Ran El-Yaniv, and Yoshua Bengio · 2016
Cited alongside, same era.
Reducing the model order of deep neural networks using information theory
Ming Tu, Visar Berisha, Yu Cao, and Jae-sun Seo · 2016
Cited alongside, same era.
Ai benchmark: Running deep neural networks on android smartphones
Andrey Ignatov, Radu Timofte, William Chou, Ke Wang, Max Wu, Tim Hartley, and Luc Van Gool · 2018
Cited alongside, same era.
Haq: Hardware-aware automated quantization with mixed precision
Kuan Wang, Zhijian Liu, Yujun Lin, Ji Lin, and Song Han · 2019
Later among the works it cites.
Hawq-v2: Hessian aware trace-weighted quantization of neural networks
Zhen Dong, Zhewei Yao, Daiyaan Arfeen, Amir Gholami, Michael W Mahoney, and Kurt Keutzer · 2020
Later among the works it cites.
Comparing fisher information regularization with distillation for dnn quantization
Prad Kadambi · 2020
Later among the works it cites.
Hessian based analysis of sgd for deep nets: Dynamics and generalization
Xinyan Li, Qilong Gu, Yingxue Zhou, Tiancong Chen, and Arindam Banerjee · 2020
Later among the works it cites.
Bn-nas: Neural architecture search with batch normalization
Boyu Chen, Peixia Li, Baopu Li, Chen Lin, Chuming Li, Ming Sun, Junjie Yan, and Wanli Ouyang · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Benoit Jacob, Skirmantas Kligys, Bo Chen, Menglong Zhu, Matthew Tang, Andrew Howard, Hartwig Adam, and Dmitry Kalenichenko · 2018
Cited alongside, same era.
An elementary introduction to information geometry
Frank Nielsen · 2018
Cited alongside, same era.
Mixed precision quantization of convnets via differentiable neural architecture search, 2018
Bichen Wu, Yanghan Wang, Peizhao Zhang, Yuandong Tian, Peter Vajda, and Kurt Keutzer · 2018
Cited alongside, same era.
Hawq: Hessian aware quantization of neural networks with mixed-precision
Zhen Dong, Zhewei Yao, Amir Gholami, Michael W Mahoney, and Kurt Keutzer · 2019
Cited alongside, same era.
Universal statistics of fisher information in deep neural networks: Mean field approach
Ryo Karakida, Shotaro Akaho, and Shun-ichi Amari · 2019
Cited alongside, same era.
Limitations of the empirical fisher approximation for natural gradient descent
Frederik Kunstner, Philipp Hennig, and Lukas Balles · 2019
Cited alongside, same era.
Quantifying the carbon emissions of machine learning
Alexandre Lacoste, Alexandra Luccioni, Victor Schmidt, and Thomas Dandres · 2019
Cited alongside, same era.
Claudionor N Coelho, Aki Kuusela, Shan Li, Hao Zhuang, Jennifer Ngadiuba, Thea Klaeboe Aarrestad, Vladimir Loncar, Maurizio Pierini, Adrian Alan Pol, and Sioni Summers · 2021
Later among the works it cites.
A survey of quantization methods for efficient neural network inference, 2021
Amir Gholami, Sehoon Kim, Zhen Dong, Zhewei Yao, Michael W. Mahoney, and Kurt Keutzer · 2021
Later among the works it cites.
Bmpq: Bit-gradient sensitivity driven mixed-precision quantization of dnns from scratch, 2021
Souvik Kundu, Shikai Wang, Qirui Sun, Peter A. Beerel, and Massoud Pedram · 2021
Later among the works it cites.
Brecq: Pushing the limit of post-training quantization by block reconstruction
Yuhang Li, Ruihao Gong, Xu Tan, Yang Yang, Peng Hu, Qi Zhang, Fengwei Yu, Wei Wang, and Shi Gu · 2021
Later among the works it cites.
Layer importance estimation with imprinting for neural network quantization
Hongyang Liu, Sara Elkerdawy, Nilanjan Ray, and Mostafa Elhoushi · 2021
Later among the works it cites.
Hawq-v3: Dyadic neural network quantization
Zhewei Yao, Zhen Dong, Zhangcheng Zheng, Amir Gholami, Jiali Yu, Eric Tan, Leyuan Wang, Qijing Huang, Yida Wang, Michael Mahoney, et al · 2021
Later among the works it cites.
Hessian-aware pruning and optimal neural implant
Shixing Yu, Zhewei Yao, Amir Gholami, Zhen Dong, Sehoon Kim, Michael W Mahoney, and Kurt Keutzer · 2021
Later among the works it cites.
Learning rate grafting: Transferability of optimizer tuning, 2022
Naman Agarwal, Rohan Anil, Elad Hazan, Tomer Koren, and Cyril Zhang · 2022
Closest in time.
Mixed-precision neural network quantization via learned layer-wise importance, 2022
Chen Tang, Kai Ouyang, Zhi Wang, Yifei Zhu, Yaowei Wang, Wen Ji, and Wenwu Zhu · 2022
Closest in time.