Fetching the paper…
Reading the bibliography…
Neural network architectures found by sophistic search algorithms achieve strikingly good test performance, surpassing most human-crafted network models by significant margins.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning”, Machine Learning
1992
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory”, In Neural Computations
1997
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images”, Tech Report
2009
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever and G. E. Hinton, “ImageNet classification with deep convolutional neural networks”, In NIPS
2012
Earlier work this paper cites.
K. Simonyan, A. Zisserman, “Very deep convolutional networks for large-scale image recognition”, Arxiv, 1409.1556, 2014
2014
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke and A. Rabinovich, “Going deeper with convolutions”, In CVPR
2015
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift”, Arxiv, 1502.03167, 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. L. Ba, “Adam: A method for stochastic optimization”, In ICLR
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, “Identity Mappings in Deep Residual Networks”, Arxiv, 1603.050274, 2016
2016
Earlier work this paper cites.
G. Huang, Z. Liu, L. van der Maaten and K. Q. Weinberger, “Densely Connected Convolutional Networks”, Arxiv, 1608.06993, 2016
2016
Earlier work this paper cites.
B. Zoph and Q. V. Le, “Neural Architecture Search with Reinforcement Learning”, Arxiv, 1611.01578, 2016
2016
Cited alongside, same era.
G. Larsson, M. Maire and G. Shakhnarovich, “FractalNet: Ultra-Deep Neural Networks without Residuals”, Arxiv, 1605.07648v4, 2016
2016
Cited alongside, same era.
I. Loshchilov and F. Hutter, “SGDR: Stochastic Gradient Descent with Warm Restarts”, Arxiv, 1608.03983, 2016
2016
Cited alongside, same era.
S. Zagoruyko and N. Komodakis, “Wide Residual Networks”, Arxiv, 1605.07146, 2016
2016
Cited alongside, same era.
D. Han, J. Kim and J. Kim, “Deep Pyramidal Residual Networks”, Arxiv, 1610.02915, 2016
2016
Cited alongside, same era.
B. Zoph, V. Vasudevan, J. Shlens and Q. V. Le, “Learning Transferable Architectures for Scalable Image Recognition”, Arxiv, 1707.07012, 2017
X. Zhang, X. Zhou, M. Lin and J. Sun, “ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices”, Arxiv, 1707.01083, 2017
2017
Later among the works it cites.
T. DeVries and G. W. Taylor, “Improved Regularization of Convolutional Neural Networks with Cutout”, Arxiv, 1708.04552, 2017
2017
Later among the works it cites.
H. Pham, M. Y. Guan, B. Zoph, Q. V. Le and J. Dean, “Efficient Neural Architecture Search via Parameter Sharing”, Arxiv, 1802.03268, 2018
2018
Closest in time.
E. Real, A. Aggarwal, Y. Huang and Q. V. Le, “Regularized Evolution for Image Classifier Architecture Search”, Arxiv, 1802.01548, 2018
2018
Closest in time.
L. Wang, Y. Zhao and Y. Jinnai, “AlphaX: eXploring Neural Architectures with Deep Neural Networks and Monte Carlo Tree Search”, Arxiv, 1805.07440, 2018
2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
C. Liu, B. Zoph, M. Neumann, J. Shlens, W. Hua, L. Li, L. Fei-Fei, A. Yuille, J. Huang and K. Murphy, “Progressive Neural Architecture Search”, Arxiv, 1712.00559, 2017
2017
Cited alongside, same era.
T. Elsken, J. Metzen and F. Hutter, “Simple And Efficient Architecture Search for Convolutional Neural Networks”, Arxiv, 1711.04528, 2017
2017
Cited alongside, same era.
L. Xie and A. Yuille, “Genetic CNN”, Arxiv, 1703.01513, 2017
2017
Cited alongside, same era.
A. Brock, T. Lim, J.M. Ritchie and N. Weston, “SMASH: One-Shot Model Architecture Search through HyperNetworks”, Arxiv, 1708.05344, 2017
2017
Cited alongside, same era.
M. Abadi et al, “TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems”, https://www.tensorflow.org/
Cited in the paper.
R. Luo, F. Tian, T. Qin, E. Chen and T. Liu, “Neural Architecture Optimization”, Arxiv, 1808.07233, 2018
2018
Closest in time.
H. Liu, K. Simonyan and Y. Yang, “DARTS: Differentiable Architecture Search”, Arxiv, 1806.09055, 2018
2018
Closest in time.
J. Dong, A. Cheng, D. Juan, W. Wei and M. Sun, “DPP-Net: Device-aware Progressive Search for Pareto-optimal Neural Architectures”, Arxiv, 1806.08198, 2018
2018
Closest in time.
N. Ma, X. Zhang, H. Zheng and J. Sun, “ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design”, Arxiv, 1807.11164, 2018
2018
Closest in time.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov and L. Chen, “MobileNetV2: Inverted Residuals and Linear Bottlenecks”, Arxiv, 1801.04381, 2018
2018
Closest in time.