Optimal brain damage
Yann LeCun, John S Denker, and Sara A Solla · 1990
Earlier work this paper cites.
A neural network for factoid question answering over paragraphs
Mohit Iyyer, Jordan L. Boyd-Graber, Leonardo Max Batista Claudino, Richard Socher, and Hal Daumé · 2014
Earlier work this paper cites.
A thorough examination of the cnn/daily mail reading comprehension task
Original
Danqi Chen, Jason Bolton, and Christopher D Manning · 2016
Earlier work this paper cites.
Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding
Song Han, Huizi Mao, and William J Dally · 2016
Earlier work this paper cites.
Pointer sentinel mixture models
Original
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher · 2016
Earlier work this paper cites.
Abstractive text summarization using sequence-to-sequence rnns and beyond
Original
Ramesh Nallapati, Bowen Zhou, Caglar Gulcehre, Bing Xiang, et al · 2016
Earlier work this paper cites.
Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension
Original
Mandar Joshi, Eunsol Choi, Daniel S Weld, and Luke Zettlemoyer · 2017
Earlier work this paper cites.
To prune, or not to prune: exploring the efficacy of pruning for model compression
Original
Michael Zhu and Suyog Gupta · 2017
Earlier work this paper cites.
Sparse networks from scratch: Faster training without losing performance
Original
Tim Dettmers and Luke Zettlemoyer · 2019
Earlier work this paper cites.
The state of sparsity in deep neural networks
Original
Trevor Gale, Erich Elsen, and Sara Hooker · 2019
Earlier work this paper cites.
Freebaseqa: A new factoid qa data set matching trivia-style question-answer pairs with freebase
Kelvin Jiang, Dekun Wu, and Hui Jiang · 2019
Earlier work this paper cites.
Parameter efficient training of deep convolutional neural networks by dynamic sparse reparameterization
Hesham Mostafa and Xin Wang · 2019
Earlier work this paper cites.
Iteratively training look-up tables for network quantization
Fabien Cardinaux, Stefan Uhlich, Kazuki Yoshiyama, Javier Alonso García, Lukas Mauch, Stephen Tiedemann, Thomas Kemp, and Akira Nakamura · 2020
Earlier work this paper cites.
Rigging the lottery: Making all tickets winners
Utku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro, and Erich Elsen · 2020
Earlier work this paper cites.
Measuring massive multitask language understanding
Original
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt · 2020
Earlier work this paper cites.
Dynamic model pruning with feedback
Tao Lin, Sebastian U. Stich, Luis Barba, Daniil Dmitriev, and Martin Jaggi · 2020
Earlier work this paper cites.
Permute, quantize, and fine-tune: Efficient compression of neural networks
Julieta Martinez, Jashan Shewakramani, Ting Liu, Ioan Andrei Bârsan, Wenyuan Zeng, and Raquel Urtasun · 2020
Earlier work this paper cites.
Nvidia a100 tensor core gpu architecture
Nvidia · 2020
Earlier work this paper cites.
Movement pruning: Adaptive sparsity by fine-tuning
Victor Sanh, Thomas Wolf, and Alexander Rush · 2020
Earlier work this paper cites.
Woodfisher: Efficient second-order approximation for neural network compression
Sidak Pal Singh and Dan Alistarh · 2020
Earlier work this paper cites.
M-fac: Efficient matrix-free approximations of second-order information
Elias Frantar, Eldar Kurtic, and Dan Alistarh · 2021
Earlier work this paper cites.