Fetching the paper…
Reading the bibliography…
We study deep learning approaches to inferring numerical coordinates for points of interest in an input image.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning internal representations by error propagation,” DTIC Document, Tech. Rep., 1985
1985
Earlier work this paper cites.
Y. LeCun, B. E. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. E. Hubbard, and L. D. Jackel, “Handwritten digit recognition with a back-propagation network,” in Advances in neural information processing systems , 1990, pp. 396–404
1990
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
T. Tieleman and G. Hinton, “Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude,” COURSERA: Neural networks for machine learning , vol. 4, no. 2, 2012
2012
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan, “Network in network,” in International Conference on Learning Representations , 2014
2014
Earlier work this paper cites.
J. J. Tompson, A. Jain, Y. LeCun, and C. Bregler, “Joint training of a convolutional network and a graphical model for human pose estimation,” in Advances in neural information processing systems , 2014, pp. 1799–1807
2014
Earlier work this paper cites.
A. Toshev and C. Szegedy, “Deeppose: Human pose estimation via deep neural networks,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2014, pp. 1653–1660
2014
Earlier work this paper cites.
M. Andriluka, L. Pishchulin, P. Gehler, and B. Schiele, “2d human pose estimation: New benchmark and state of the art analysis,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2014, pp. 3686–3693
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 3431–3440
2015
Cited alongside, same era.
J. Tompson, R. Goroshin, A. Jain, Y. LeCun, and C. Bregler, “Efficient object localization using convolutional networks,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 648–656
2015
Cited alongside, same era.
A. Radford, L. Metz, and S. Chintala, “Unsupervised representation learning with deep convolutional generative adversarial networks,” in International Conference on Learning Representations , 2016
2016
Cited alongside, same era.
A. Newell, K. Yang, and J. Deng, “Stacked hourglass networks for human pose estimation,” in European Conference on Computer Vision . Springer, 2016, pp. 483–499
2016
Cited alongside, same era.
S.-E. Wei, V. Ramakrishna, T. Kanade, and Y. Sheikh, “Convolutional pose machines,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 4724–4732
2016
Later among the works it cites.
A. Bulat and G. Tzimiropoulos, “Human pose estimation via convolutional part heatmap regression,” in European Conference on Computer Vision . Springer, 2016, pp. 717–732
2016
Later among the works it cites.
W. Yang, S. Li, W. Ouyang, H. Li, and X. Wang, “Learning feature pyramids for human pose estimation,” in The IEEE International Conference on Computer Vision , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 770–778
2016
Cited alongside, same era.
K. M. Yi, E. Trulls, V. Lepetit, and P. Fua, “Lift: Learned invariant feature transform,” in European Conference on Computer Vision . Springer, 2016, pp. 467–483
2016
Cited alongside, same era.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” Journal of Machine Learning Research , vol. 17, no. 39, pp. 1–40, 2016
2016
Cited alongside, same era.
F. Yu and V. Koltun, “Multi-scale context aggregation by dilated convolutions,” in ICLR , 2016
2016
Cited alongside, same era.
U. Rafi, B. Leibe, J. Gall, and I. Kostrikov, “An efficient convolutional network for human pose estimation.” in BMVC , vol. 1, 2016, p. 2
2016
Cited alongside, same era.
Y. Chen, C. Shen, X.-S. Wei, L. Liu, and J. Yang, “Adversarial posenet: A structure-aware convolutional network for human pose estimation,” in The IEEE International Conference on Computer Vision , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
X. Chu, W. Yang, W. Ouyang, C. Ma, A. L. Yuille, and X. Wang, “Multi-context attention for human pose estimation,” in The IEEE Conference on Computer Vision and Pattern Recognition , 2017
2017
Later among the works it cites.
M. Jaderberg, K. Simonyan, A. Zisserman et al. , “Spatial transformer networks,” in Advances in Neural Information Processing Systems , 2015, pp. 2017–2025
2025
Closest in time.