Fetching the paper…
Reading the bibliography…
The past few years have witnessed a growth in size and computational requirements for training and inference with neural networks.
An efficient heuristic procedure for partitioning graphs
Kernighan, Brian W and Lin, Shen · 1970
Earlier work this paper cites.
Optimization by simulated annealing
Kirkpatrick, Scott, Vecchi, Mario P, et al · 1983
Earlier work this paper cites.
A linear-time heuristic for improving network partitions
Fiduccia, Charles M and Mattheyses, Robert M · 1988
Earlier work this paper cites.
Optimization by simulated annealing: an experimental evaluation; part i, graph partitioning
Johnson, David S, Aragon, Cecilia R, McGeoch, Lyle A, and Schevon, Catherine · 1989
Earlier work this paper cites.
New spectral methods for ratio cut partitioning and clustering
Hagen, Lars and Kahng, Andrew B · 1992
Earlier work this paper cites.
Simple statistical gradient following algorithms for connectionnist reinforcement learning
Williams, Ronald · 1992
Earlier work this paper cites.
A multilevel algorithm for partitioning graphs
Hendrickson, B. and Leland, R · 1993
Earlier work this paper cites.
A fast multilevel implementation of recursive spectral bisection for partitioning unstructured problems
Barnard, S. T. and Simon, H. D · 1994
Earlier work this paper cites.
Experimental analysis of the dual recursive bipartitioning algorithm for static mapping
Pellegrini, F. and Roman, J · 1996
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jurgen · 1997
Earlier work this paper cites.
Improvement of the efficiency of genetic algorithms for scalable parallel graph partitioning in a multi-level framework
Chevalier, C. and Pellegrini, F · 2006
Earlier work this paper cites.
A parallelisable multi-level banded diffusion scheme for computing balanced partitions with smooth boundaries
Pellegrini, F · 2007
Earlier work this paper cites.
Distillating knowledge about scotch
Pellegrini, F · 2009
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Hinton, Geoffrey, Deng, Li, Yu, Dong, Dahl, George E., Mohamed, Abdel-rahman, Jaitly, Navdeep, Senior, Andrew, Vanhoucke, Vincent, Nguyen, Patrick, Sainath, Tara N., et al · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Cited alongside, same era.
Lecture 6.5—RmsProp: Divide the gradient by a running average of its recent magnitude
Tieleman, T. and Hinton, G · 2012
Cited alongside, same era.
On the difficulty of training recurrent neural networks
Pascanu, Razvan, Mikolov, Tomas, and Bengio, Yoshua · 2013
Cited alongside, same era.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, Kyunghyun, Van Merriënboer, Bart, Gulcehre, Caglar, Bahdanau, Dzmitry, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2014
Cited alongside, same era.
Towards end-to-end speech recognition with recurrent neural networks
Graves, Alex and Jaitly, Navdeep · 2014
Cited alongside, same era.
Pointer networks
Vinyals, Oriol, Fortunato, Meire, and Jaitly, Navdeep · 2015
Later among the works it cites.
Tensorflow: A system for large-scale machine learning
Abadi, Martín, Barham, Paul, Chen, Jianmin, Chen, Zhifeng, Davis, Andy, Dean, Jeffrey, Devin, Matthieu, Ghemawat, Sanjay, Irving, Geoffrey, Isard, Michael, Kudlur, Manjunath, Levenberg, Josh, Monga, Rajat, Moore, Sherry, Murray, Derek G., Steiner, Benoit, Tucker, Paul, Vasudevan, Vijay, Warden, Pete, Wicke, Martin, Yu, Yuan, and Zheng, Xiaoqiang · 2016
Later among the works it cites.
Neural combinatorial optimization with reinforcement learning
Bello, Irwan, Pham, Hieu, Le, Quoc V., Norouzi, Mohammad, and Bengio, Samy · 2016
Later among the works it cites.
Dermatologist-level classification of skin cancer
Esteva, Andre, Kuprel, Brett, Novoa, Rob, Ko, Justin, Swetter, Susan, Blau, Helen M., and Thrun, Sebastian · 2016
Later among the works it cites.
Deep residual learning for image recognition
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hannun, Awni, Case, Carl, Casper, Jared, Catanzaro, Bryan, Diamos, Greg, Elsen, Erich, Prenger, Ryan, Satheesh, Sanjeev, Sengupta, Shubho, Coates, Adam, et al · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc V · 2014
Cited alongside, same era.
Recurrent neural network regularization
Zaremba, Wojciech, Sutskever, Ilya, and Vinyals, Oriol · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2015
Cited alongside, same era.
Chan, William, Jaitly, Navdeep, Le, Quoc V, and Vinyals, Oriol · 2015
Cited alongside, same era.
ImageNet Large Scale Visual Recognition Challenge
Russakovsky, Olga, Deng, Jia, Su, Hao, Krause, Jonathan, Satheesh, Sanjeev, Ma, Sean, Huang, Zhiheng, Karpathy, Andrej, Khosla, Aditya, Bernstein, Michael, Berg, Alexander C., and Fei-Fei, Li · 2015
Cited alongside, same era.
Later among the works it cites.
Exploring the limits of language modeling
Jozefowicz, Rafal, Vinyals, Oriol, Schuster, Mike, Shazeer, Noam, and Wu, Yonghui · 2016
Later among the works it cites.
Achieving budget-optimality with adaptive schemes in crowdsourcing
Khetan, Ashish and Oh, Sewoong · 2016
Later among the works it cites.
Resource management with deep reinforcement learning
Mao, Hongzi, Alizadeh, Mohammad, Menache, Ishai, and Kandula, Srikanth · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
Oord, Aaron van den, Dieleman, Sander, Zen, Heiga, Simonyan, Karen, Vinyals, Oriol, Graves, Alex, Kalchbrenner, Nal, Senior, Andrew, and Kavukcuoglu, Koray · 2016
Later among the works it cites.
Rethinking the inception architecture for computer vision
Szegedy, Christian, Vanhoucke, Vincent, Ioffe, Sergey, Shlens, Jon, and Wojna, Zbigniew · 2016
Later among the works it cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Yonghui, Schuster, Mike, Chen, Zhifeng, Le, Quoc V., Norouzi, Mohammad, Macherey, Wolfgang, Krikun, Maxim, Cao, Yuan, Gao, Qin, Macherey, Klaus, Klingner, Jeff, Shah, Apurva, Johnson, Melvin, Liu, Xiaobing, Łukasz Kaiser, Gouws, Stephan, Kato, Yoshikiyo, Kudo, Taku, Kazawa, Hideto, Stevens, Keith, Kurian, George, Patil, Nishant, Wang, Wei, Young, Cliff, Smith, Jason, Riesa, Jason, Rudnick, Alex, Vinyals, Oriol, Corrado, Greg, Hughes, Macduff, and Dean, Jeffrey · 2016
Later among the works it cites.
Deep voice: Real-time neural text-to-speech
Arik, Sercan O, Chrzanowski, Mike, Coates, Adam, Diamos, Gregory, Gibiansky, Andrew, Kang, Yongguo, Li, Xian, Miller, John, Raiman, Jonathan, Sengupta, Shubho, et al · 2017
Closest in time.
Tacotron: A fully end-to-end text-to-speech synthesis model
Wang, Yuxuan, Skerry-Ryan, R. J., Stanton, Daisy, Wu, Yonghui, Weiss, Ron J., Jaitly, Navdeep, Yang, Zongheng, Xiao, Ying, Chen, Zhifeng, Bengio, Samy, Le, Quoc V., Agiomyrgiannakis, Yannis, Clark, Rob, and Saurous, Rif A · 2017
Closest in time.