Fetching the paper…
Reading the bibliography…
Second-order information has proven to be very effective in determining the redundancy of neural network weights and activations.
Haim Avron · 2011
Earlier work this paper cites.
Exploiting linear structure within convolutional networks for efficient evaluation, 2014
Emily Denton, Wojciech Zaremba, Joan Bruna, Yann LeCun, and Rob Fergus · 2014
Earlier work this paper cites.
Binaryconnect: Training deep neural networks with binary weights during propagations, 2015
Matthieu Courbariaux, Yoshua Bengio, and Jean-Pierre David · 2015
Earlier work this paper cites.
Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding, 2015
Song Han, Huizi Mao, and William J. Dally · 2015
Earlier work this paper cites.
Deep residual learning for image recognition, 2015
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning, 2015
Timothy P. Lillicrap, Jonathan J. Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Earlier work this paper cites.
Xnor-net: Imagenet classification using binary convolutional neural networks, 2016
Mohammad Rastegari, Vicente Ordonez, Joseph Redmon, and Ali Farhadi · 2016
Earlier work this paper cites.
Trained ternary quantization, 2016
Chenzhuo Zhu, Song Han, Huizi Mao, and William J. Dally · 2016
Earlier work this paper cites.
Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients, 2016
Shuchang Zhou, Yuxin Wu, Zekun Ni, Xinyu Zhou, He Wen, and Yuheng Zou · 2016
Earlier work this paper cites.
Designing neural network architectures using reinforcement learning, 2016
Bowen Baker, Otkrist Gupta, Nikhil Naik, and Ramesh Raskar · 2016
Cited alongside, same era.
Towards accurate binary convolutional neural network, 2017
Xiaofan Lin, Cong Zhao, and Wei Pan · 2017
Cited alongside, same era.
Incremental network quantization: Towards lossless cnns with low-precision weights, 2017
Aojun Zhou, Anbang Yao, Yiwen Guo, Lin Xu, and Yurong Chen · 2017
Cited alongside, same era.
Quantization and training of neural networks for efficient integer-arithmetic-only inference, 2017
Benoit Jacob, Skirmantas Kligys, Bo Chen, Menglong Zhu, Matthew Tang, Andrew Howard, Hartwig Adam, and Dmitry Kalenichenko · 2017
Cited alongside, same era.
Learning transferable architectures for scalable image recognition, 2017
Barret Zoph, Vijay Vasudevan, Jonathon Shlens, and Quoc V. Le · 2017
Cited alongside, same era.
Quantizing deep convolutional networks for efficient inference: A whitepaper, 2018
Raghuraman Krishnamoorthi · 2018
Later among the works it cites.
Post-training 4-bit quantization of convolution networks for rapid-deployment, 2018
Ron Banner, Yury Nahshan, Elad Hoffer, and Daniel Soudry · 2018
Later among the works it cites.
Bridging the accuracy gap for 2-bit quantized neural networks (qnn), 2018
Jungwook Choi, Pierce I-Jen Chuang, Zhuo Wang, Swagath Venkataramani, Vijayalakshmi Srinivasan, and Kailash Gopalakrishnan · 2018
Later among the works it cites.
Hawq-v2: Hessian aware trace-weighted quantization of neural networks, 2019
Zhen Dong, Zhewei Yao, Yaohui Cai, Daiyaan Arfeen, Amir Gholami, Michael W. Mahoney, and Kurt Keutzer · 2019
Later among the works it cites.
Q-bert: Hessian based ultra low precision quantization of bert, 2019
Sheng Shen, Zhen Dong, Jiayu Ye, Linjian Ma, Zhewei Yao, Amir Gholami, Michael W. Mahoney, and Kurt Keutzer · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mark Sandler, Andrew Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen · 2018
Cited alongside, same era.
Amc: Automl for model compression and acceleration on mobile devices, 2018
Yihui He, Ji Lin, Zhijian Liu, Hanrui Wang, Li-Jia Li, and Song Han · 2018
Cited alongside, same era.
Haq: Hardware-aware automated quantization with mixed precision, 2018
Kuan Wang, Zhijian Liu, Yujun Lin, Ji Lin, and Song Han · 2018
Cited alongside, same era.
Pact: Parameterized clipping activation for quantized neural networks, 2018
Jungwook Choi, Zhuo Wang, Swagath Venkataramani, Pierce I-Jen Chuang, Vijayalakshmi Srinivasan, and Kailash Gopalakrishnan · 2018
Cited alongside, same era.
Hawq: Hessian aware quantization of neural networks with mixed-precision, 2019
Zhen Dong, Zhewei Yao, Amir Gholami, Michael Mahoney, and Kurt Keutzer · 2019
Later among the works it cites.
Low-bit quantization of neural networks for efficient inference, 2019
Yoni Choukroun, Eli Kravchik, Fan Yang, and Pavel Kisilev · 2019
Later among the works it cites.
Autoq: Automated kernel-wise neural network quantization, 2019
Qian Lou, Feng Guo, Lantao Liu, Minje Kim, and Lei Jiang · 2019
Later among the works it cites.