Fetching the paper…
Reading the bibliography…
The exponential growth in numbers of parameters of neural networks over the past years has been accompanied by an increase in performance across several fields.
Stabilizing the lottery ticket hypothesis
Jonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, and Michael Carbin · 1903
Earlier work this paper cites.
Balanced Binary Neural Networks with Gated Residual
Mingzhu Shen, Xianglong Liu, Ruihao Gong, and Kai Han · 1909
Earlier work this paper cites.
Rigging the Lottery: Making All Tickets Winners
Utku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro, and Erich Elsen · 1911
Earlier work this paper cites.
Pruning versus clipping in neural networks
Steven A. Janowsky · 1989
Earlier work this paper cites.
Optimal brain damage
Yann LeCun, John S Denker, and Sara A Solla · 1990
Earlier work this paper cites.
Progressive Skeletonization: Trimming more fat from a network at initialization
Pau de Jorge, Amartya Sanyal, Harkirat S. Behl, Philip H. S. Torr, Gregory Rogez, and Puneet K. Dokania · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images, 2009
Alex Krizhevsky · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Mnist handwritten digit database
Yann LeCun, Corinna Cortes, and C J Burges · 2010
Earlier work this paper cites.
Trained ternary quantization
Chenzhuo Zhu, Song Han, Huizi Mao, and William J. Dally · 2010
Earlier work this paper cites.
Neural networks for machine learning. coursera, video lectures: Lecture 9c, 2012
Geoffrey Hinton · 2012
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Yoshua Bengio, Nicholas Léonard, and Aaron Courville · 2013
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting, 2014
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, and Ruslan Salakhutdinov · 2014
Cited alongside, same era.
Fast and accurate deep network learning by exponential linear units (elus)
Djork-Arné Clevert, Thomas Unterthiner, and Sepp Hochreiter · 2015
Cited alongside, same era.
BinaryConnect: Training Deep Neural Networks with binary weights during propagations
Matthieu Courbariaux, Yoshua Bengio, and Jean-Pierre David · 2015
Cited alongside, same era.
Learning both weights and connections for efficient neural networks
Song Han, Jeff Pool, John Tran, and William J. Dally · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
The lottery ticket hypothesis: Finding sparse, trainable neural networks
Jonathan Frankle and Michael Carbin · 2018
Later among the works it cites.
Importance estimation for neural network pruning
Pavlo Molchanov, Arun Mallya, Stephen Tyree, Iuri Frosio, and Jan Kautz · 2019
Later among the works it cites.
Deconstructing lottery tickets: Zeros, signs, and the supermask
Hattie Zhou, Janice Lan, Rosanne Liu, and Jason Yosinski · 2019
Later among the works it cites.
Textual evidence for the perfunctoriness of independent medical reviews
Adrian Brasoveanu, Megan Moodie, and Rakshit Agrawal · 2020
Later among the works it cites.
An embarrassingly simple approach to training ternary weight networks
Xiang Deng and Zhongfei Zhang · 2020
Later among the works it cites.
At-scale sparse deep neural network inference with efficient gpu implementation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Variational dropout and the local reparameterization trick
Diederik P. Kingma, Tim Salimans, and Max Welling · 2015
Cited alongside, same era.
Ternary neural networks for resource-efficient ai applications
Hande Alemdar, Vincent Leroy, Adrien Prost-Boucle, and Frédéric Pétrot · 2016
Cited alongside, same era.
Binarized neural networks: Training deep neural networks with weights and activations constrained to +1 or -1
Matthieu Courbariaux, Itay Hubara, Daniel Soudry, Ran El-Yaniv, and Yoshua Bengio · 2016
Cited alongside, same era.
Fifty years of graph matching, network alignment and network comparison
Frank Emmert-Streib, Matthias Dehmer, and Yongtang Shi · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Deep Residual Networks with Exponential Linear Unit
Anish Shah, Eashan Kadam, Hena Shah, Sameer Shinde, and Sandip Shingade · 2016
Cited alongside, same era.
Mert Hidayetoglu, Carl Pearson, Vikram Sharma Mailthody, Eiman Ebrahimi, Jinjun Xiong, Rakesh Nagi, and Wen Mei Hwu · 2020
Later among the works it cites.
What’s Hidden in a Randomly Weighted Neural Network?
Vivek Ramanujan, Mitchell Wortsman, Aniruddha Kembhavi, Ali Farhadi, and Mohammad Rastegari · 2020
Later among the works it cites.
Energy and policy considerations for deep learning in nlp
Emma Strubell, Ananya Ganesh, and Andrew McCallum · 2020
Later among the works it cites.
Pruning Randomly Initialized Neural Networks with Iterative Randomization
Daiki Chijiwa, Shin’ya Yamaguchi, Yasutoshi Ida, Kenji Umakoshi, and Tomohiro Inoue · 2021
Later among the works it cites.
James Diffenderfer and Bhavya Kailkhura · 2021
Later among the works it cites.
Sparse Training via Boosting Pruning Plasticity with Neuroregeneration
Shiwei Liu, Tianlong Chen, Xiaohan Chen, Zahra Atashgahi, Lu Yin, Huanyu Kou, Li Shen, Mykola Pechenizkiy, Zhangyang Wang, and Decebal Constantin Mocanu · 2021
Later among the works it cites.
Scalable training of artificial neural networks with adaptive sparse connectivity inspired by network science
Decebal Constantin Mocanu, Elena Mocanu, Peter Stone, Phuong H. Nguyen, Madeleine Gibescu, and Antonio Liotta · 2041
Closest in time.