Snip: Single-shot network pruning based on connection sensitivity
Original
Namhoon Lee, Thalaiyasingam Ajanthan, and Philip HS Torr · 2018
Later among the works it cites.
Rethinking the value of network pruning
Original
Zhuang Liu, Mingjie Sun, Tinghui Zhou, Gao Huang, and Trevor Darrell · 2018
Later among the works it cites.
Scalable training of artificial neural networks with adaptive sparse connectivity inspired by network science
Decebal Constantin Mocanu, Elena Mocanu, Peter Stone, Phuong H Nguyen, Madeleine Gibescu, and Antonio Liotta · 2018
Later among the works it cites.
Evolutionary pruning of transfer learned deep convolutional neural network for breast cancer diagnosis in digital breast tomosynthesis
Ravi K Samala, Heang-Ping Chan, Lubomir M Hadjiiski, Mark A Helvie, Caleb Richter, and Kenny Cha · 2018
Later among the works it cites.
Adaptive methods for nonconvex optimization
Manzil Zaheer, Sashank Reddi, Devendra Sachan, Satyen Kale, and Sanjiv Kumar · 2018
Later among the works it cites.
Bridging the gap between deep learning and sparse matrix format selection
Yue Zhao, Jiajia Li, Chunhua Liao, and Xipeng Shen · 2018
Later among the works it cites.
How can we be so dense? the benefits of using highly sparse representations
Original
Subutai Ahmad and Luiz Scheinkman · 2019
Later among the works it cites.
Sparse networks from scratch: Faster training without losing performance
Original
Tim Dettmers and Luke Zettlemoyer · 2019
Later among the works it cites.
Rigging the Lottery: Making All Tickets Winners
Original
Utku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro, and Erich Elsen · 2019
Later among the works it cites.
The difficulty of training sparse neural networks
Original
Utku Evci, Fabian Pedregosa, Aidan Gomez, and Erich Elsen · 2019
Later among the works it cites.
The lottery ticket hypothesis: Finding sparse, trainable neural networks
Jonathan Frankle and Michael Carbin · 2019
Later among the works it cites.
The State of Sparsity in Deep Neural Networks
Original
Trevor Gale, Erich Elsen, and Sara Hooker · 2019
Later among the works it cites.
What Do Compressed Deep Neural Networks Forget?
Original
Sara Hooker, Aaron Courville, Gregory Clark, Yann Dauphin, and Andrea Frome · 2019
Later among the works it cites.
Fantastic generalization measures and where to find them
Original
Yiding Jiang, Behnam Neyshabur, Hossein Mobahi, Dilip Krishnan, and Samy Bengio · 2019
Later among the works it cites.
Characterizing well-behaved vs. pathological deep neural networks
Antoine Labatie · 2019
Later among the works it cites.
A signal propagation perspective for pruning neural networks at initialization
Original
Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould, and Philip HS Torr · 2019
Later among the works it cites.
Optimizing sparse tensor times matrix on gpus
Yuchen Ma, Jiajia Li, Xiaolong Wu, Chenggang Yan, Jimeng Sun, and Richard Vuduc · 2019
Later among the works it cites.
Parameter efficient training of deep convolutional neural networks by dynamic sparse reparameterization
Original
Hesham Mostafa and Xin Wang · 2019
Later among the works it cites.
Energy and policy considerations for deep learning in nlp, 2019
Emma Strubell, Ananya Ganesh, and Andrew McCallum · 2019
Later among the works it cites.
TinyML: Machine Learning with TensorFlow Lite on Arduino and Ultra-Low-Power Microcontrollers
P. Warden and D. Situnayake · 2019
Later among the works it cites.
A mean field theory of batch normalization
Original
Greg Yang, Jeffrey Pennington, Vinay Rao, Jascha Sohl-Dickstein, and Samuel S Schoenholz · 2019
Later among the works it cites.
Dive into deep learning
Aston Zhang, Zachary C Lipton, Mu Li, and Alexander J Smola · 2019
Later among the works it cites.
Snap: A 1.67—21.55 tops/w sparse neural acceleration processor for unstructured sparse deep neural network inference in 16nm cmos
Jie-Fang Zhang, Ching-En Lee, Chester Liu, Yakun Sophia Shao, Stephen W Keckler, and Zhengya Zhang · 2019
Later among the works it cites.
Activation function impact on sparse neural networks
Adam Dubowski · 2020
Later among the works it cites.
Gradient flow in sparse neural networks and how lottery tickets win
Original
Utku Evci, Yani A Ioannou, Cem Keskin, and Yann Dauphin · 2020
Later among the works it cites.
Pruning neural networks at initialization: Why are we missing the mark?, 2020
Jonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, and Michael Carbin · 2020
Later among the works it cites.
Finding trainable sparse networks through neural tangent transfer
Original
Tianlin Liu and Friedemann Zenke · 2020
Later among the works it cites.
Pruning neural networks without any data by iteratively conserving synaptic flow
Original
Hidenori Tanaka, Daniel Kunin, Daniel LK Yamins, and Surya Ganguli · 2020
Later among the works it cites.
The Computational Limits of Deep Learning
Original
Neil C. Thompson, Kristjan Greenewald, Keeheon Lee, and Gabriel F. Manso · 2020
Later among the works it cites.
Picking winning tickets before training by preserving gradient flow
Original
Chaoqi Wang, Guodong Zhang, and Roger Grosse · 2020
Later among the works it cites.
Sparch: Efficient architecture for sparse matrix multiplication
Zhekai Zhang, Hanrui Wang, Song Han, and William J Dally · 2020
Later among the works it cites.
A gradient flow framework for analyzing network pruning
Ekdeep Singh Lubana and Robert Dick · 2021
Closest in time.