Fetching the paper…
Reading the bibliography…
Advances in Deep Learning have recently made it possible to recover full 3D meshes of human poses from individual images.
T. Pfister, J. Charles, and A. Zisserman, “Flowing convnets for human pose estimation in videos,” in Proceedings of the IEEE International Conference on Computer Vision , 2015, pp. 1913–1921
1921
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
D. Baraff and A. Witkin, “Large steps in cloth simulation,” in Computer graphics and interactive techniques . ACM, 1998, pp. 43–54
1998
Earlier work this paper cites.
C. J. Taylor, “Reconstruction of articulated objects from point correspondences in a single uncalibrated image,” CVIU , vol. 80, no. 3, pp. 349–363, 2000
2000
Earlier work this paper cites.
P. F. Felzenszwalb and D. P. Huttenlocher, “Pictorial structures for object recognition,” International journal of computer vision , vol. 61, no. 1, pp. 55–79, 2005
2005
Earlier work this paper cites.
M. Hauth, “Numerical techniques for cloth simulation,” system (figure 2 (a) , vol. 15, p. 3, 2005
2005
Earlier work this paper cites.
M. Andriluka, S. Roth, and B. Schiele, “Pictorial structures revisited: People detection and articulated pose estimation,” in CVPR . IEEE, 2009, pp. 1014–1021
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in CVPR . Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
M. Andriluka, S. Roth, and B. Schiele, “Monocular 3d pose estimation and tracking by detection,” in 2010 IEEE Computer Society Conference on Computer Vision and Pattern Recognition . IEEE, 2010, pp. 623–630
2010
Earlier work this paper cites.
L. Sigal, A. O. Balan, and M. J. Black, “Humaneva: Synchronized video and motion capture dataset and baseline algorithm for evaluation of articulated human motion,” IJCV , vol. 87, no. 1-2, p. 4, 2010
2010
Earlier work this paper cites.
B. Sapp, D. Weiss, and B. Taskar, “Parsing human motion with stretchable models,” in CVPR 2011 . IEEE, 2011, pp. 1281–1288
2011
Earlier work this paper cites.
S. Johnson and M. Everingham, “Learning effective human pose estimation from inaccurate annotation,” in CVPR , 2011
2011
Earlier work this paper cites.
2012
Earlier work this paper cites.
V. Ramakrishna, T. Kanade, and Y. Sheikh, “Reconstructing 3d human pose from 2d image landmarks,” in ECCV . Springer, 2012, pp. 573–586
2012
Earlier work this paper cites.
Y. Yang and D. Ramanan, “Articulated human detection with flexible mixtures of parts,” IEEE transactions on pattern analysis and machine intelligence , vol. 35, no. 12, pp. 2878–2890, 2013
2013
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in ECCV . Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
C. Ionescu, D. Papava, V. Olaru, and C. Sminchisescu, “Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments,” PAMI , vol. 36, no. 7, pp. 1325–1339, jul 2014
2014
Earlier work this paper cites.
M. Loper, N. Mahmood, and M. J. Black, “Mosh: Motion and shape capture from sparse markers,” ACM Transactions on Graphics , vol. 33, no. 6, p. 220, 2014
2014
Cited alongside, same era.
M. Cimpoi, S. Maji, I. Kokkinos, S. Mohamed, , and A. Vedaldi, “Describing textures in the wild,” in CVPR , 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
2014
Cited alongside, same era.
D. Mehta, H. Rhodin, D. Casas, P. Fua, O. Sotnychenko, W. Xu, and C. Theobalt, “Monocular 3d human pose estimation in the wild using improved cnn supervision,” in 3D Vision (3DV) . IEEE, 2017, pp. 506–516
2017
Later among the works it cites.
C. Lassner, G. Pons-Moll, and P. V. Gehler, “A generative model of people in clothing,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 853–862
2017
Later among the works it cites.
G. Varol, J. Romero, X. Martin, N. Mahmood, M. J. Black, I. Laptev, and C. Schmid, “Learning from Synthetic Humans,” in CVPR , 2017
2017
Later among the works it cites.
G. Varol, J. Romero, X. Martin, N. Mahmood, M. J. Black, I. Laptev, and C. Schmid, “Learning from synthetic humans,” in CVPR , 2017, pp. 109–117
2017
Later among the works it cites.
M. Wang, X. Chen, W. Liu, C. Qian, L. Lin, and L. Ma, “Drpose3d: Depth ranking in 3d human pose estimation,” IJCAI , 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
M. Andriluka, L. Pishchulin, P. Gehler, and B. Schiele, “2d human pose estimation: New benchmark and state of the art analysis,” in CVPR , 2014, pp. 3686–3693
2014
Cited alongside, same era.
C. Ionescu, D. Papava, V. Olaru, and C. Sminchisescu, “Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments,” PAMI , vol. 36, no. 7, pp. 1325–1339, 2014
2014
Cited alongside, same era.
M. Loper, N. Mahmood, J. Romero, G. Pons-Moll, and M. J. Black, “Smpl: A skinned multi-person linear model,” ACM Transactions on Graphics , vol. 34, no. 6, p. 248, 2015
2015
Cited alongside, same era.
H. Joo, H. Liu, L. Tan, L. Gui, B. Nabbe, I. Matthews, T. Kanade, S. Nobuhara, and Y. Sheikh, “Panoptic studio: A massively multiview system for social motion capture,” in Proceedings of the IEEE International Conference on Computer Vision , 2015, pp. 3334–3342
2015
Cited alongside, same era.
S.-E. Wei, V. Ramakrishna, T. Kanade, and Y. Sheikh, “Convolutional pose machines,” in CVPR , 2016, pp. 4724–4732
2016
Cited alongside, same era.
A. Newell, K. Yang, and J. Deng, “Stacked hourglass networks for human pose estimation,” in European Conference on Computer Vision . Springer, 2016, pp. 483–499
2016
Cited alongside, same era.
C. Kampouris, S. Zafeiriou, A. Ghosh, and S. Malassiotis, “Fine-grained material classification using micro-geometry and reflectance,” in ECCV . Springer, 2016, pp. 778–792
2016
Cited alongside, same era.
2018
Later among the works it cites.
Y. Luo, J. Ren, Z. Wang, W. Sun, J. Pan, J. Liu, J. Pang, and L. Lin, “Lstm pose machines,” in CVPR . IEEE, 2018, pp. 5207–5215
2018
Later among the works it cites.
A. Kanazawa, M. J. Black, D. W. Jacobs, and J. Malik, “End-to-end recovery of human shape and pose,” in CVPR , 2018, pp. 7122–7131
2018
Later among the works it cites.
W. Yang, W. Ouyang, X. Wang, J. Ren, H. Li, and X. Wang, “3d human pose estimation in the wild by adversarial learning,” in CVPR , vol. 1, 2018
2018
Later among the works it cites.
T. Alldieck, M. Magnor, W. Xu, C. Theobalt, and G. Pons-Moll, “Video based reconstruction of 3d people models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 8387–8397
2018
Later among the works it cites.
——, “Detailed human avatars from monocular video,” in 2018 International Conference on 3D Vision (3DV) . IEEE, 2018, pp. 98–109
2018
Later among the works it cites.
T. Yu, Z. Zheng, K. Guo, J. Zhao, Q. Dai, H. Li, G. Pons-Moll, and Y. Liu, “Doublefusion: Real-time capture of human performances with inner body shapes from a single depth sensor,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 7287–7296
2018
Later among the works it cites.
M. Omran, C. Lassner, G. Pons-Moll, P. Gehler, and B. Schiele, “Neural body fitting: Unifying deep learning and model based human pose and shape estimation,” in 2018 International Conference on 3D Vision (3DV) . IEEE, 2018, pp. 484–494
2018
Later among the works it cites.
B. Zhou, A. Lapedriza, A. Khosla, A. Oliva, and A. Torralba, “Places: A 10 million image database for scene recognition,” PAMI , vol. 40, no. 6, pp. 1452–1464, 2018
2018
Later among the works it cites.
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in International conference on machine learning , 2015, pp. 2048–2057
2057
Closest in time.