Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude
Tijmen Tieleman and Geoffrey Hinton · 2012
Cited alongside, same era.
Exploiting sparseness in deep neural networks for large vocabulary speech recognition
Dong Yu, Frank Seide, Gang Li, and Li Deng · 2012
Cited alongside, same era.
Fast training of convolutional networks through ffts
Original
Michael Mathieu, Mikael Henaff, and Yann LeCun · 2013
Cited alongside, same era.
Memory bounded deep convolutional networks
Original
Maxwell D Collins and Pushmeet Kohli · 2014
Cited alongside, same era.
Fixed-point feedforward deep neural network design using weights+ 1, 0, and- 1
Kyuyeon Hwang and Wonyong Sung · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Original
Karen Simonyan and Andrew Zisserman · 2014
Cited alongside, same era.
Structured pruning of deep convolutional neural networks
Original
Sajid Anwar, Kyuyeon Hwang, and Wonyong Sung
Cited in the paper.
A deep neural network compression pipeline: Pruning, quantization, huffman encoding
Original
Song Han, Huizi Mao, and William J Dally
Cited in the paper.
Learning both weights and connections for efficient neural network
Song Han, Jeff Pool, John Tran, and William Dally
Cited in the paper.
Resiliency of deep neural networks under quantization
Wonyong Sung, Sungho Shin, and Kyuyeon Hwang
Cited in the paper.