Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients, 2016
Shuchang Zhou, Yuxin Wu, Zekun Ni, Xinyu Zhou, He Wen, and Yuheng Zou · 2016
Later among the works it cites.
Trained ternary quantization, 2016
Chenzhuo Zhu, Song Han, Huizi Mao, and William J. Dally · 2016
Later among the works it cites.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollar, and Ross Girshick · 2017
Later among the works it cites.
Channel pruning for accelerating very deep neural networks
Yihui He, Xiangyu Zhang, and Jian Sun · 2017
Later among the works it cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Original
Andrew G Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam · 2017
Later among the works it cites.
Towards accurate binary convolutional neural network
Xiaofan Lin, Cong Zhao, and Wei Pan · 2017
Later among the works it cites.
Learning efficient convolutional networks through network slimming
Zhuang Liu, Jianguo Li, B Shen, Gao Huang, Shoumeng Yan, and Changshui Zhang · 2017
Later among the works it cites.
Bnn+: Improved binary network training
Original
Sajad Darabi, Mouloud Belbahri, Matthieu Courbariaux, and Vahid Partovi Nia · 2018
Later among the works it cites.
Soft filter pruning for accelerating deep convolutional neural networks
Original
Yang He, Guoliang Kang, Xuanyi Dong, Yanwei Fu, and Yi Yang · 2018
Later among the works it cites.
Bi-real net: Enhancing the performance of 1-bit cnns with improved representational capability and advanced training algorithm
Zechun Liu, Baoyuan Wu, Wenhan Luo, Xin Yang, Wei Liu, and Kwang-Ting Cheng · 2018
Later among the works it cites.
Lq-nets: Learned quantization for highly accurate and compact deep neural networks
Dongqing Zhang, Jiaolong Yang, Dongqiangzi Ye, and Gang Hua · 2018
Later among the works it cites.
Shufflenet: An extremely efficient convolutional neural network for mobile devices
Xiangyu Zhang, Xinyu Zhou, Mengxiao Lin, and Jian Sun · 2018
Later among the works it cites.
Filter pruning via geometric median for deep convolutional neural networks acceleration
Yang He, Ping Liu, Ziwei Wang, Zhilan Hu, and Yi Yang · 2019
Later among the works it cites.
Learning instance-wise sparsity for accelerating deep models
Chuanjian Liu, Yunhe Wang, Kai Han, Chunjing Xu, and Chang Xu · 2019
Later among the works it cites.
Searching for accurate binary neural architectures
Mingzhu Shen, Kai Han, Chunjing Xu, and Yunhe Wang · 2019
Later among the works it cites.
Binary ensemble neural network: More bits per network or more networks per bit?
Shilin Zhu, Xin Dong, and Hao Su · 2019
Later among the works it cites.