Fetching the paper…
Reading the bibliography…
Perceptual capabilities of artificial systems have come a long way since the advent of deep learning.
Backpropagation applied to handwritten zip code recognition
LeCun, Y., Boser, B., Denker, J.S., Henderson, D., Howard, R.E., Hubbard, W., Jackel, L.D., 1989 · 1989
Earlier work this paper cites.
Neural mechanisms of selective visual attention
Desimone, R., Duncan, J., 1995 · 1995
Earlier work this paper cites.
Convolutional networks for images, speech, and time series
LeCun, Y., Bengio, Y., et al., 1995 · 1995
Earlier work this paper cites.
Cortical mechanisms of space-based and object-based attentional control
Yantis, S., Serences, J.T., 2003 · 2003
Earlier work this paper cites.
Feature-based attention in visual cortex
Maunsell, J.H., Treue, S., 2006 · 2006
Earlier work this paper cites.
Expectation (and attention) in visual cognition
Summerfield, C., Egner, T., 2009 · 2009
Earlier work this paper cites.
Representation of multiple, independent categories in the primate prefrontal cortex
Cromer, J.A., Roy, J.E., Miller, E.K., 2010 · 2010
Earlier work this paper cites.
Top-down control of visual attention
Noudoost, B., Chang, M.H., Steinmetz, N.A., Moore, T., 2010 · 2010
Earlier work this paper cites.
Visual attention: The past 25 years
Carrasco, M., 2011 · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks, in: Advances in neural information processing systems, pp. 1097–1105
Krizhevsky, A., Sutskever, I., Hinton, G.E., 2012 · 2012
Cited alongside, same era.
Top-down influences on visual processing
Gilbert, C.D., Li, W., 2013 · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., Bengio, Y., 2014 · 2014
Cited alongside, same era.
The cifar-10 dataset
Krizhevsky, A., Nair, V., Hinton, G., 2014 · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K., Zisserman, A., 2014 · 2014
Cited alongside, same era.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and< 0.5 mb model size
Iandola, F.N., Han, S., Moskewicz, M.W., Ashraf, K., Dally, W.J., Keutzer, K., 2016 · 2016
Later among the works it cites.
Encode, review, and decode: Reviewer module for caption generation
Wu, Z., Ye, Y., Yuexin, Y., Cohen, R.S.W.W., 2016 · 2016
Later among the works it cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Howard, A.G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., Adam, H., 2017 · 2017
Later among the works it cites.
Knowing when to look: Adaptive attention via a visual sentinel for image captioning, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), p. 2
Lu, J., Xiong, C., Parikh, D., Socher, R., 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
How transferable are features in deep neural networks?, in: Advances in neural information processing systems, pp. 3320–3328
Yosinski, J., Clune, J., Bengio, Y., Lipson, H., 2014 · 2014
Cited alongside, same era.
Deep learning
LeCun, Y., Bengio, Y., Hinton, G., 2015 · 2015
Cited alongside, same era.
Going deeper with convolutions, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 1–9
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., Rabinovich, A., 2015 · 2015
Cited alongside, same era.
Deep residual learning for image recognition, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 770–778
He, K., Zhang, X., Ren, S., Sun, J., 2016 · 2016
Cited alongside, same era.
Neural mechanisms of selective visual attention
Moore, T., Zirnsak, M., 2017 · 2017
Later among the works it cites.
An overview of multi-task learning in deep neural networks
Ruder, S., 2017 · 2017
Later among the works it cites.
Inception-v4, inception-resnet and the impact of residual connections on learning., in: AAAI, p. 12
Szegedy, C., Ioffe, S., Vanhoucke, V., Alemi, A.A., 2017 · 2017
Later among the works it cites.
Shufflenet: An extremely efficient convolutional neural network for mobile devices. arxiv 2017
Zhang, X., Zhou, X., Lin, M., Sun, J., 2017 · 2017
Later among the works it cites.
Show, attend and tell: Neural image caption generation with visual attention, in: International conference on machine learning, pp. 2048–2057
Xu, K., Ba, J., Kiros, R., Cho, K., Courville, A., Salakhudinov, R., Zemel, R., Bengio, Y., 2015 · 2057
Closest in time.