Fetching the paper…
Reading the bibliography…
Convolutional neural networks have gained a remarkable success in computer vision.
C. J. C. H. Watkins, “Learning from delayed rewards,” Ph.D. dissertation, King’s College, Cambridge, 1989
1989
Earlier work this paper cites.
J. D. Schaffer, D. Whitley, and L. J. Eshelman, “Combinations of genetic algorithms and neural networks: A survey of the state of the art,” in International Workshop on Combinations of Genetic Algorithms and Neural Networks . IEEE, 1992, pp. 1–37
1992
Earlier work this paper cites.
L.-J. Lin, “Reinforcement learning for robots using neural networks,” Carnegie-Mellon Univ Pittsburgh PA School of Computer Science, Tech. Rep., 1993
1993
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural Computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press Cambridge, 1998, vol. 1, no. 1
1998
Earlier work this paper cites.
A. Y. Ng, D. Harada, and S. Russell, “Policy invariance under reward transformations: Theory and application to reward shaping,” in International Conference on Machine Learning , vol. 99, 1999, pp. 278–287
1999
Earlier work this paper cites.
S. Hochreiter, A. S. Younger, and P. R. Conwell, “Learning to learn using gradient descent,” in International Conference on Artificial Neural Networks . Springer, 2001, pp. 87–94
2001
Earlier work this paper cites.
K. O. Stanley and R. Miikkulainen, “Evolving neural networks through augmenting topologies,” Evolutionary Computation , vol. 10, no. 2, pp. 99–127, 2002
2002
Earlier work this paper cites.
R. Vilalta and Y. Drissi, “A perspective view and survey of meta-learning,” Artificial Intelligence Review , vol. 18, no. 2, pp. 77–95, 2002
2002
Earlier work this paper cites.
K. O. Stanley, D. B. D’Ambrosio, and J. Gauci, “A hypercube-based encoding for evolving large-scale neural networks,” Artificial Life , vol. 15, no. 2, pp. 185–212, 2009
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in IEEE Conference on Computer Vision and Pattern Recognition , 2009, pp. 248–255
2009
Earlier work this paper cites.
T. Mikolov, M. Karafiát, L. Burget, J. Černockỳ, and S. Khudanpur, “Recurrent neural network based language model,” in Eleventh Annual Conference of the International Speech Communication Association , 2010
2010
Earlier work this paper cites.
J. S. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl, “Algorithms for hyper-parameter optimization,” in Advances in Neural Information Processing Systems , 2011, pp. 2546–2554
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in Neural Information Processing Systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
J. Dean, G. Corrado, R. Monga, K. Chen, M. Devin, M. Mao, A. Senior, P. Tucker, K. Yang, Q. V. Le et al. , “Large scale distributed deep networks,” in Advances in Neural Information Processing Systems , 2012, pp. 1223–1231
2012
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan, “Network in network,” in International Conference on Learning Representations , 2013
2013
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Advances in Neural Information Processing Systems , 2013, pp. 3111–3119
2013
Earlier work this paper cites.
M. Li, L. Zhou, Z. Yang, A. Li, F. Xia, D. G. Andersen, and A. Smola, “Parameter server for distributed machine learning,” in Big Learning NIPS Workshop , vol. 6, 2013, p. 2
2013
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,” Nature , vol. 521, no. 7553, pp. 436–444, 2015
2015
Earlier work this paper cites.
R. Girshick, “Fast r-cnn,” in IEEE International Conference on Computer Vision , 2015, pp. 1440–1448
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” in Advances in Neural Information Processing Systems , 2015, pp. 91–99
2015
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 3431–3440
2015
Cited alongside, same era.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in 3rd International Conference for Learning Representations , 2015
2015
Cited alongside, same era.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 1–9
2015
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in International Conference on Machine Learning , 2015, pp. 448–456
2015
Cited alongside, same era.
2017
Later among the works it cites.
2017
Later among the works it cites.
F. Chollet, “Xception: Deep learning with depthwise separable convolutions,” in IEEE Conference on Computer Vision and Pattern Recognition , 2017
2017
Later among the works it cites.
M. Suganuma, S. Shirakawa, and T. Nagao, “A genetic programming approach to designing convolutional neural network architectures,” in Genetic and Evolutionary Computation Conference , 2017, pp. 497–504
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Cited alongside, same era.
T. Domhan, J. T. Springenberg, and F. Hutter, “Speeding up automatic hyperparameter optimization of deep neural networks by extrapolation of learning curves.” in International Joint Conference on Artificial Intelligence , 2015, pp. 3460–3468
2015
Cited alongside, same era.
K. He and J. Sun, “Convolutional neural networks at constrained time cost,” in IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 5353–5360
2015
Cited alongside, same era.
J. Yue-Hei Ng, M. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici, “Beyond short snippets: Deep networks for video classification,” in IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 4694–4702
2015
Cited alongside, same era.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in 3rd International Conference on Learning Representations , 2015
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in IEEE International Conference on Computer Vision , 2015, pp. 1026–1034
2015
Cited alongside, same era.
H. Nam and B. Han, “Learning multi-domain convolutional neural networks for visual tracking,” in IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 4293–4302
2016
Cited alongside, same era.
L. Bertinetto, J. Valmadre, J. F. Henriques, A. Vedaldi, and P. H. Torr, “Fully-convolutional siamese networks for object tracking,” in European Conference on Computer Vision . Springer, 2016, pp. 850–865
2016
Cited alongside, same era.
L. Xie and A. Yuille, “Genetic cnn,” in International Conference on Computer Vision , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
A. Klein, S. Falkner, J. T. Springenberg, and F. Hutter, “Learning curve prediction with bayesian neural networks,” International Conference on Learning Representations , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
G. Huang, Z. Liu, K. Q. Weinberger, and L. van der Maaten, “Densely connected convolutional networks,” in IEEE Conference on Computer Vision and Pattern Recognition , 2017
2017
Later among the works it cites.
S. Xie, R. Girshick, P. Dollár, Z. Tu, and K. He, “Aggregated residual transformations for deep neural networks,” in IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 5987–5995
2017
Later among the works it cites.
Y. Chen, J. Li, H. Xiao, X. Jin, S. Yan, and J. Feng, “Dual path networks,” in Advances in Neural Information Processing Systems , 2017, pp. 4467–4475
2017
Later among the works it cites.
C. Szegedy, S. Ioffe, V. Vanhoucke, and A. A. Alemi, “Inception-v4, inception-resnet and the impact of residual connections on learning.” in AAAI Conference on Artificial Intelligence , vol. 4, 2017, p. 12
2017
Later among the works it cites.
X. Zhang, Z. Li, C. C. Loy, and D. Lin, “Polynet: A pursuit of structural diversity in very deep networks,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2017, pp. 3900–3908
2017
Later among the works it cites.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 40, no. 4, pp. 834–848, 2018
2018
Closest in time.
Z. Zhong, J. Yan, W. Wu, J. Shao, and C.-L. Liu, “Blockqnn: Practical block-wise neural network architecture generation,” in IEEE Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.
2018
Closest in time.
H. Pham, M. Y. Guan, B. Zoph, Q. V. Le, and J. Dean, “Faster discovery of neural architectures by searching for paths in a large model,” International Conference on Learning Representations Workshop , 2018
2018
Closest in time.
2018
Closest in time.
B. Zoph, V. Vasudevan, J. Shlens, and Q. V. Le, “Learning transferable architectures for scalable image recognition,” in IEEE Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in IEEE Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.