Fetching the paper…
Reading the bibliography…
Estimating human pose and shape from monocular images is a long-standing problem in computer vision.
——, “End-to-end human pose and mesh reconstruction with transformers,” in CVPR , 2021, pp. 1954–1963
1963
Earlier work this paper cites.
H. Choi, G. Moon, J. Y. Chang, and K. M. Lee, “Beyond static features for temporally consistent 3D human pose and shape from a video,” in CVPR , 2021, pp. 1964–1973
1973
Earlier work this paper cites.
R. Nevatia and T. O. Binford, “Description and recognition of curved objects,” Artificial intelligence , vol. 8, no. 1, pp. 77–98, 1977
1977
Earlier work this paper cites.
D. Marr and H. K. Nishihara, “Representation and recognition of the spatial organization of three-dimensional shapes,” Proceedings of the Royal Society of London. Series B. Biological Sciences , vol. 200, no. 1140, pp. 269–294, 1978
1978
Earlier work this paper cites.
H.-J. Lee and Z. Chen, “Determination of 3D human body postures from a single view,” Computer Vision, Graphics, and Image Processing , vol. 30, no. 2, pp. 148–168, 1985
1985
Earlier work this paper cites.
A. Pentland and B. Horowitz, “Recovery of nonrigid motion and structure,” TPAMI , vol. 13, no. 07, pp. 730–742, 1991
1991
Earlier work this paper cites.
D. Metaxas and D. Terzopoulos, “Shape and nonrigid motion estimation through physics-based synthesis,” TPAMI , vol. 15, no. 6, pp. 580–591, 1993
1993
Earlier work this paper cites.
K. Rohr, “Towards model-based recognition of human movements in image sequences,” CVGIP: Image understanding , vol. 59, no. 1, pp. 94–115, 1994
1994
Earlier work this paper cites.
S. X. Ju, M. J. Black, and Y. Yacoob, “Cardboard people: A parameterized model of articulated image motion,” in FG . IEEE, 1996, pp. 38–44
1996
Earlier work this paper cites.
D. M. Gavrila, Vision-based 3-D tracking of humans in action . University of Maryland, College Park, 1996
1996
Earlier work this paper cites.
T. Vetter and V. Blanz, “Estimating coloured 3D face models from single images: An example based approach,” in ECCV , 1998, pp. 499–513
1998
Earlier work this paper cites.
V. Blanz and T. Vetter, “A morphable model for the synthesis of 3D faces,” in SIGGRAPH , 1999, pp. 187–194
1999
Earlier work this paper cites.
S. Wachter and H.-H. Nagel, “Tracking of persons in monocular image sequences,” CVIU , vol. 74, no. 3, pp. 174–192, 1999
1999
Earlier work this paper cites.
H. Sidenbladh, M. J. Black, and D. J. Fleet, “Stochastic tracking of 3D human figures using 2D image motion,” in ECCV . Springer, 2000, pp. 702–718
2000
Earlier work this paper cites.
L. Kakadiaris and D. Metaxas, “Model-based estimation of 3D human motion,” TPAMI , vol. 22, no. 12, pp. 1453–1459, 2000
2000
Earlier work this paper cites.
R. Plänkers and P. Fua, “Tracking and modeling people in video sequences,” CVIU , vol. 81, no. 3, pp. 285–302, 2001
2001
Earlier work this paper cites.
K. M. Robinette, S. Blackwell, H. Daanen, M. Boehmer, S. Fleming, T. Brill, D. Hoeferlin, and D. Burnsides, “Civilian American and European Surface Anthropometry Resource (CAESAR) final report,” US Air Force Research Laboratory, Tech. Rep. AFRL-HE-WP-TR-2002-0169, 2002
2002
Earlier work this paper cites.
K. Grauman, G. Shakhnarovich, and T. Darrell, “Inferring 3D structure with a statistical image-based shape model,” in ICCV , 2003, pp. 641–648
2003
Earlier work this paper cites.
C. Sminchisescu and B. Triggs, “Estimating articulated human motion with covariance scaled sampling,” International Journal of Robotics Research , vol. 22, no. 6, pp. 371–391, 2003
2003
Earlier work this paper cites.
B. Allen, B. Curless, and Z. Popović, “The space of human body shapes: reconstruction and parameterization from range scans,” TOG , vol. 22, no. 3, pp. 587–594, 2003
2003
Earlier work this paper cites.
A. Mohr and M. Gleicher, “Building efficient, accurate character skins from examples,” TOG , vol. 22, no. 3, pp. 562–568, 2003
2003
Earlier work this paper cites.
A. Agarwal and B. Triggs, “Recovering 3D human pose from monocular images,” TPAMI , vol. 28, no. 1, pp. 44–58, 2005
2005
Earlier work this paper cites.
D. Anguelov, P. Srinivasan, D. Koller, S. Thrun, J. Rodgers, and J. Davis, “SCAPE: shape completion and animation of people,” TOG , vol. 24, pp. 408–416, 2005
2005
Earlier work this paper cites.
M. Teschner, S. Kimmerle, B. Heidelberger, G. Zachmann, L. Raghupathi, A. Fuhrmann, M.-P. Cani, F. Faure, N. Magnenat-Thalmann, W. Strasser et al. , “Collision detection for deformable objects,” in CGF , vol. 24, no. 1. Wiley Online Library, 2005, pp. 61–81
2005
Earlier work this paper cites.
B. Allen, B. Curless, Z. Popović, and A. Hertzmann, “Learning a correlated model of identity and pose-dependent body shape variation for real-time synthesis,” in SCA . ACM, 2006, pp. 147–156
2006
Earlier work this paper cites.
A. O. Balan, L. Sigal, M. J. Black, J. E. Davis, and H. W. Haussecker, “Detailed human shape and pose from images,” in CVPR . IEEE, 2007, pp. 1–8
2007
Earlier work this paper cites.
L. Sigal, A. Balan, and M. Black, “Combined discriminative and generative articulated pose and non-rigid shape estimation,” NeurIPS , vol. 20, pp. 1337–1344, 2007
2007
Earlier work this paper cites.
A. O. Bălan and M. J. Black, “The naked truth: Estimating body shape under clothing,” in ECCV . Springer, 2008, pp. 15–29
2008
Earlier work this paper cites.
N. Hasler, C. Stoll, M. Sunkel, B. Rosenhahn, and H.-P. Seidel, “A statistical model of human pose and body shape,” in CGF , vol. 28. Wiley Online Library, 2009, pp. 337–346
2009
Earlier work this paper cites.
P. Guan, A. Weiss, A. O. Balan, and M. J. Black, “Estimating human shape and pose from a single image,” in ICCV . IEEE, 2009, pp. 1381–1388
2009
Earlier work this paper cites.
P. Paysan, R. Knothe, B. Amberg, S. Romdhani, and T. Vetter, “A 3D face model for pose and illumination invariant face recognition,” in AVSS . Ieee, 2009, pp. 296–301
2009
Earlier work this paper cites.
W.-S. Zheng, S. Gong, and T. Xiang, “Associating groups of people,” in BMVC , vol. 2, no. 6, 2009, pp. 1–11
2009
Earlier work this paper cites.
L. Sigal, A. O. Balan, and M. J. Black, “HumanEva: Synchronized video and motion capture dataset and baseline algorithm for evaluation of articulated human motion,” IJCV , vol. 87, no. 1-2, p. 4, 2010
2010
Earlier work this paper cites.
N. Hasler, T. Thormählen, B. Rosenhahn, and H.-P. Seidel, “Learning skeletons for shape and pose,” in I3D , 2010, pp. 23–30
2010
Earlier work this paper cites.
N. Hasler, H. Ackermann, B. Rosenhahn, T. Thormählen, and H.-P. Seidel, “Multilinear pose and body shape estimation of dressed subjects from image sets,” in CVPR . IEEE, 2010, pp. 1823–1830
2010
Earlier work this paper cites.
S. Zhou, H. Fu, L. Liu, D. Cohen-Or, and X. Han, “Parametric reshaping of human bodies in images,” TOG , vol. 29, no. 4, pp. 1–10, 2010
2010
Earlier work this paper cites.
“Carnegie mellon university - cmu graphics lab - motion capture library,” http://mocap.cs.cmu.edu/ , 2010
2010
Earlier work this paper cites.
S. Johnson and M. Everingham, “Clustered pose and nonlinear appearance models for human pose estimation,” in BMVC , 2010, pp. 12.1–12.11
2010
Earlier work this paper cites.
G. Pons-Moll and B. Rosenhahn, “Model-based pose estimation,” Visual Analysis of Humans , pp. 139–170, 2011
2011
Earlier work this paper cites.
——, “Learning effective human pose estimation from inaccurate annotation,” in CVPR . IEEE, 2011, pp. 1465–1472
2011
Earlier work this paper cites.
O. Freifeld and M. J. Black, “Lie bodies: A manifold representation of 3D human shape,” in ECCV . Springer, 2012, pp. 1–14
2012
Earlier work this paper cites.
D. A. Hirshberg, M. Loper, E. Rachlin, and M. J. Black, “Coregistration: Simultaneous alignment and modeling of articulated 3D shape,” in ECCV . Springer, 2012, pp. 242–255
2012
Earlier work this paper cites.
Y. Chen, Z. Liu, and Z. Zhang, “Tensor-based human body modeling,” in CVPR , 2013, pp. 105–112
2013
Earlier work this paper cites.
C. Cao, Y. Weng, S. Zhou, Y. Tong, and K. Zhou, “FaceWarehouse: a 3D facial expression database for visual computing,” TVCG , vol. 20, no. 3, pp. 413–425, 2013
2013
Earlier work this paper cites.
O. Aldrian and W. A. Smith, “Inverse rendering of faces with a 3D morphable model,” TPAMI , vol. 35, no. 5, pp. 1080–1093, 2013
2013
Earlier work this paper cites.
M. Loper, N. Mahmood, and M. J. Black, “MoSh: Motion and shape capture from sparse markers,” TOG , vol. 33, no. 6, pp. 1–13, 2014
2014
Earlier work this paper cites.
C. Ionescu, D. Papava, V. Olaru, and C. Sminchisescu, “Human3.6M: Large scale datasets and predictive methods for 3D human sensing in natural environments,” TPAMI , vol. 36, no. 7, pp. 1325–1339, 2014
2014
Earlier work this paper cites.
M. M. Loper and M. J. Black, “OpenDR: An approximate differentiable renderer,” in ECCV , 2014, pp. 154–169
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” NeurIPS , vol. 27, 2014
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in ICLR , 2014
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: Common objects in context,” in ECCV . Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
M. Andriluka, L. Pishchulin, P. Gehler, and B. Schiele, “2D human pose estimation: New benchmark and state of the art analysis,” in CVPR , 2014, pp. 3686–3693
2014
Earlier work this paper cites.
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei, “Large-scale video classification with convolutional neural networks,” in CVPR , 2014, pp. 1725–1732
2014
Earlier work this paper cites.
M. Loper, N. Mahmood, J. Romero, G. Pons-Moll, and M. J. Black, “SMPL: A skinned multi-person linear model,” TOG , vol. 34, no. 6, pp. 1–16, 2015
2015
Earlier work this paper cites.
G. Pons-Moll, J. Romero, N. Mahmood, and M. J. Black, “Dyna: A model of dynamic human shape in motion,” TOG , vol. 34, no. 4, pp. 1–14, 2015
2015
Earlier work this paper cites.
S. Zuffi and M. J. Black, “The stitched puppet: A graphical model of 3D human shape and pose,” in CVPR , 2015, pp. 3537–3546
2015
Earlier work this paper cites.
I. Akhter and M. J. Black, “Pose-conditioned joint angle limits for 3D human pose reconstruction,” in CVPR , 2015, pp. 1446–1455
2015
Earlier work this paper cites.
D. Rezende and S. Mohamed, “Variational inference with normalizing flows,” in International conference on machine learning . PMLR, 2015, pp. 1530–1538
2015
Earlier work this paper cites.
H. Joo, H. Liu, L. Tan, L. Gui, B. Nabbe, I. Matthews, T. Kanade, S. Nobuhara, and Y. Sheikh, “Panoptic studio: A massively multiview system for social motion capture,” in ICCV , 2015, pp. 3334–3342
2015
Earlier work this paper cites.
F. Bogo, A. Kanazawa, C. Lassner, P. Gehler, J. Romero, and M. J. Black, “Keep it SMPL: Automatic estimation of 3D human pose and shape from a single image,” in ECCV . Springer, 2016, pp. 561–578
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
J. Thies, M. Zollhöfer, M. Stamminger, C. Theobalt, and M. Nießner, “Face2Face: Real-time face capture and reenactment of RGB videos,” in CVPR , 2016, pp. 2387–2395
2016
Earlier work this paper cites.
H.-S. Fang, S. Xie, Y.-W. Tai, and C. Lu, “RMPE: Regional multi-person pose estimation,” in ICCV , 2017, pp. 2334–2343
2017
Earlier work this paper cites.
J. Martinez, R. Hossain, J. Romero, and J. J. Little, “A simple yet effective baseline for 3D human pose estimation,” in ICCV , 2017, pp. 2659–2668
2017
Earlier work this paper cites.
G. Pavlakos, X. Zhou, K. G. Derpanis, and K. Daniilidis, “Coarse-to-fine volumetric prediction for single-image 3D human pose,” in CVPR , 2017, pp. 1263–1272
2017
Earlier work this paper cites.
Y. Huang, F. Bogo, C. Lassner, A. Kanazawa, P. V. Gehler, J. Romero, I. Akhter, and M. J. Black, “Towards accurate marker-less human shape and pose estimation over time,” in 3DV . IEEE, 2017, pp. 421–430
2017
Earlier work this paper cites.
C. Lassner, J. Romero, M. Kiefel, F. Bogo, M. J. Black, and P. V. Gehler, “Unite the people: Closing the loop between 3D and 2D human representations,” in CVPR , 2017, pp. 6050–6059
2017
Earlier work this paper cites.
H.-Y. F. Tung, H.-W. Tung, E. Yumer, and K. Fragkiadaki, “Self-supervised learning of motion capture,” NeurIPS , pp. 5236–5246, 2017
2017
Earlier work this paper cites.
T. Li, T. Bolkart, M. J. Black, H. Li, and J. Romero, “Learning a model of facial shape and expression from 2D scans,” TOG , vol. 36, no. 6, pp. 194–1, 2017
2017
Earlier work this paper cites.
J. Romero, D. Tzionas, and M. J. Black, “Embodied hands: Modeling and capturing hands and bodies together,” TOG , vol. 36, no. 6, pp. 1–17, 2017
2017
Earlier work this paper cites.
G. Varol, J. Romero, X. Martin, N. Mahmood, M. J. Black, I. Laptev, and C. Schmid, “Learning from synthetic humans,” in CVPR , 2017, pp. 109–117
2017
Earlier work this paper cites.
C. Zimmermann and T. Brox, “Learning to estimate 3D hand pose from single RGB images,” in ICCV , 2017, pp. 4913–4921
2017
Earlier work this paper cites.
A. S. Jackson, A. Bulat, V. Argyriou, and G. Tzimiropoulos, “Large pose 3D face reconstruction from a single image via direct volumetric CNN regression,” in ICCV , 2017, pp. 1031–1039
2017
Earlier work this paper cites.
A. Tewari, M. Zollhöfer, H. Kim, P. Garrido, F. Bernard, P. Perez, and C. Theobalt, “MoFA: model-based deep convolutional face autoencoder for unsupervised monocular reconstruction,” in ICCV , 2017, pp. 3735–3744
2017
Earlier work this paper cites.
T. Simon, H. Joo, I. Matthews, and Y. Sheikh, “Hand keypoint detection in single images using multiview bootstrapping,” in CVPR , 2017, pp. 4645–4653
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask R-CNN,” in ICCV , 2017, pp. 2961–2969
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
L. Dinh, J. Sohl-Dickstein, and S. Bengio, “Density estimation using real NVP,” in ICLR , 2017
2017
Earlier work this paper cites.
M. Trumble, A. Gilbert, C. Malleson, A. Hilton, and J. Collomosse, “Total Capture: 3D human pose estimation fusing video and inertial sensors,” in BMVC , 2017, pp. 1–13
2017
Earlier work this paper cites.
D. Mehta, H. Rhodin, D. Casas, P. Fua, O. Sotnychenko, W. Xu, and C. Theobalt, “Monocular 3D human pose estimation in the wild using improved CNN supervision,” in 3DV . IEEE, 2017, pp. 506–516
2017
Earlier work this paper cites.
G. Lisanti, N. Martinel, A. Del Bimbo, and G. Luca Foresti, “Group re-identification via unsupervised transfer of sparse features encoding,” in ICCV , 2017, pp. 2449–2458
2017
Earlier work this paper cites.
Q. Chen, T. Ge, Y. Xu, Z. Zhang, X. Yang, and K. Gai, “Semantic human matting,” in ACM MM , 2018, pp. 618–626
2018
Earlier work this paper cites.
J. Zhao, J. Li, Y. Cheng, T. Sim, S. Yan, and J. Feng, “Understanding humans in crowded scenes: Deep nested adversarial learning and a new benchmark for multi-human parsing,” in ACM MM , 2018, pp. 792–800
2018
Earlier work this paper cites.
X. Sun, B. Xiao, F. Wei, S. Liang, and Y. Wei, “Integral human pose regression,” in ECCV , 2018, pp. 536–553
2018
Earlier work this paper cites.
A. Zanfir, E. Marinoiu, and C. Sminchisescu, “Monocular 3D pose and shape estimation of multiple people in natural scenes - the importance of multiple scene constraints,” in CVPR , 2018, pp. 2148–2157
2018
Earlier work this paper cites.
A. Kanazawa, M. J. Black, D. W. Jacobs, and J. Malik, “End-to-end recovery of human shape and pose,” in CVPR , 2018, pp. 7122–7131
2018
Earlier work this paper cites.
G. Pavlakos, L. Zhu, X. Zhou, and K. Daniilidis, “Learning to estimate 3D human pose and shape from a single color image,” in CVPR , 2018, pp. 459–468
2018
Earlier work this paper cites.
M. Omran, C. Lassner, G. Pons-Moll, P. Gehler, and B. Schiele, “Neural Body Fitting: Unifying deep learning and model-based human pose and shape estimation,” in 3DV . IEEE, 2018, pp. 484–494
2018
Earlier work this paper cites.
T. Yu, Z. Zheng, K. Guo, J. Zhao, Q. Dai, H. Li, G. Pons-Moll, and Y. Liu, “DoubleFusion: Real-time capture of human performances with inner body shapes from a single depth sensor,” in CVPR , 2018, pp. 7287–7296
2018
Earlier work this paper cites.
N. Hesse, S. Pujades, J. Romero, M. J. Black, C. Bodensteiner, M. Arens, U. G. Hofmann, U. Tacke, M. Hadders-Algra, R. Weinberger et al. , “Learning an infant body model from RGB-D data for accurate full body motion analysis,” in MICCAI . Springer, 2018, pp. 792–800
2018
Earlier work this paper cites.
H. Joo, T. Simon, and Y. Sheikh, “Total Capture: A 3D deformation model for tracking faces, hands, and bodies,” in CVPR , 2018, pp. 8320–8329
2018
Earlier work this paper cites.
T. von Marcard, R. Henschel, M. J. Black, B. Rosenhahn, and G. Pons-Moll, “Recovering accurate 3D human pose in the wild using IMUs and a moving camera,” in ECCV , 2018, pp. 601–617
2018
Earlier work this paper cites.
R. A. Güler, N. Neverova, and I. Kokkinos, “DensePose: Dense human pose estimation in the wild,” in CVPR , 2018, pp. 7297–7306
2018
Earlier work this paper cites.
G. Varol, D. Ceylan, B. Russell, J. Yang, E. Yumer, I. Laptev, and C. Schmid, “BodyNet: Volumetric inference of 3D human body shapes,” in ECCV , 2018, pp. 20–36
2018
Earlier work this paper cites.
A. Zanfir, E. Marinoiu, M. Zanfir, A.-I. Popa, and C. Sminchisescu, “Deep network for the integrated 3D sensing of multiple people in natural images,” in NeurIPS , vol. 31, 2018, pp. 8410–8419
2018
Earlier work this paper cites.
U. Iqbal, P. Molchanov, T. Breuel, J. Gall, and J. Kautz, “Hand pose estimation via latent 2.5D heatmap regression,” in ECCV , 2018, pp. 125–143
2018
Earlier work this paper cites.
F. Mueller, F. Bernard, O. Sotnychenko, D. Mehta, S. Sridhar, D. Casas, and C. Theobalt, “GANerated hands for real-time 3D hand tracking from monocular RGB,” in CVPR , 2018, pp. 49–59
2018
Earlier work this paper cites.
Y. Feng, F. Wu, X. Shao, Y. Wang, and X. Zhou, “Joint 3D face reconstruction and dense alignment with position map regression network,” in ECCV , 2018, pp. 557–574
2018
Earlier work this paper cites.
A. Tewari, M. Zollhöfer, P. Garrido, F. Bernard, H. Kim, P. Pérez, and C. Theobalt, “Self-supervised multi-level face model learning for monocular reconstruction at over 250 Hz,” in CVPR , 2018, pp. 2549–2559
2018
Earlier work this paper cites.
K. Genova, F. Cole, A. Maschinot, A. Sarna, D. Vlasic, and W. T. Freeman, “Unsupervised training for 3D morphable model regression,” in CVPR , 2018, pp. 8377–8386
2018
Earlier work this paper cites.
Q. Cao, L. Shen, W. Xie, O. M. Parkhi, and A. Zisserman, “VGGFace2: A dataset for recognising faces across pose and age,” in FG , 2018, pp. 67–74
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
X. Wang, R. Girshick, A. Gupta, and K. He, “Non-local neural networks,” in CVPR , 2018, pp. 7794–7803
2018
Earlier work this paper cites.
D. Mehta, O. Sotnychenko, F. Mueller, W. Xu, S. Sridhar, G. Pons-Moll, and C. Theobalt, “Single-shot multi-person 3D pose estimation from monocular RGB,” in 3DV . IEEE, 2018, pp. 120–130
2018
Earlier work this paper cites.
M. Andriluka, U. Iqbal, E. Insafutdinov, L. Pishchulin, A. Milan, J. Gall, and B. Schiele, “PoseTrack: A benchmark for human pose estimation and tracking,” in CVPR , 2018, pp. 5167–5176
2018
Earlier work this paper cites.
H. Rhodin, M. Salzmann, and P. Fua, “Unsupervised geometry-aware representation for 3D human pose estimation,” in ECCV , 2018, pp. 750–767
2018
Earlier work this paper cites.
W. Xu, A. Chatterjee, M. Zollhöfer, H. Rhodin, D. Mehta, H.-P. Seidel, and C. Theobalt, “MonoPerfCap: Human performance capture from monocular video,” TOG , vol. 37, no. 2, pp. 1–15, 2018
2018
Earlier work this paper cites.
Z. Cao, G. Hidalgo, T. Simon, S.-E. Wei, and Y. Sheikh, “OpenPose: Realtime multi-person 2D pose estimation using part affinity fields,” TPAMI , vol. 43, no. 1, pp. 172–186, 2019
2019
Earlier work this paper cites.
G. Pavlakos, V. Choutas, N. Ghorbani, T. Bolkart, A. A. Osman, D. Tzionas, and M. J. Black, “Expressive body capture: 3D hands, face, and body from a single image,” in CVPR , 2019, pp. 10 975–10 985
2019
Earlier work this paper cites.
N. Kolotouros, G. Pavlakos, M. J. Black, and K. Daniilidis, “Learning to reconstruct 3D human pose and shape via model-fitting in the loop,” in ICCV , 2019, pp. 2252–2261
2019
Earlier work this paper cites.
N. Mahmood, N. Ghorbani, N. F. Troje, G. Pons-Moll, and M. J. Black, “AMASS: Archive of motion capture as surface shapes,” in ICCV , 2019, pp. 5442–5451
2019
Earlier work this paper cites.
D. Xiang, H. Joo, and Y. Sheikh, “Monocular total capture: Posing face, body, and hands in the wild,” in CVPR , 2019, pp. 10 965–10 974
2019
Earlier work this paper cites.
R. A. Güler and I. Kokkinos, “HoloPose: Holistic 3D human reconstruction in-the-wild,” in CVPR , 2019, pp. 10 884–10 894
2019
Earlier work this paper cites.
H. Zhang, J. Cao, G. Lu, W. Ouyang, and Z. Sun, “DaNet: Decompose-and-aggregate network for 3D human shape and pose estimation,” in ACM MM , 2019, pp. 935–944
2019
Earlier work this paper cites.
A. Kanazawa, J. Y. Zhang, P. Felsen, and J. Malik, “Learning 3D human dynamics from video,” in CVPR , 2019, pp. 5614–5623
2019
Earlier work this paper cites.
Y. Xu, S.-C. Zhu, and T. Tung, “DenseRaC: Joint 3D pose and shape estimation by dense render-and-compare,” in ICCV , 2019, pp. 7760–7770
2019
Cited alongside, same era.
Y. Sun, Y. Ye, W. Liu, W. Gao, Y. Fu, and T. Mei, “Human mesh recovery from monocular images via a skeleton-disentangled representation,” in ICCV , 2019, pp. 5349–5358
2019
Cited alongside, same era.
Y. Zhou, C. Barnes, J. Lu, J. Yang, and H. Li, “On the continuity of rotation representations in neural networks,” in CVPR , 2019, pp. 5745–5753
2019
Cited alongside, same era.
Z. Zheng, T. Yu, Y. Wei, Q. Dai, and Y. Liu, “DeepHuman: 3D human reconstruction from a single image,” in ICCV , 2019, pp. 7738–7748
2019
Cited alongside, same era.
N. Kolotouros, G. Pavlakos, and K. Daniilidis, “Convolutional mesh regression for single-image human shape reconstruction,” in CVPR , 2019, pp. 4501–4510
S. Zhang, Y. Zhang, F. Bogo, M. Pollefeys, and S. Tang, “Learning motion priors for 2D human body capture in 3D scenes,” in ICCV , 2021, pp. 11 343–11 353
2021
Later among the works it cites.
L. Müller, A. A. Osman, S. Tang, C.-H. P. Huang, and M. J. Black, “On self-contact and human pose,” in CVPR , 2021, pp. 9990–9999
2021
Later among the works it cites.
D. Rempe, T. Birdal, A. Hertzmann, J. Yang, S. Sridhar, and L. J. Guibas, “HuMoR: 3D human motion model for robust pose estimation,” in ICCV , 2021
2021
Later among the works it cites.
J. Zhang, D. Yu, J. H. Liew, X. Nie, and J. Feng, “Body meshes as points,” in CVPR , 2021, pp. 546–556
2021
Later among the works it cites.
B. Zhang, Y. Wang, X. Deng, Y. Zhang, P. Tan, C. Ma, and H. Wang, “Interacting two-hand 3D pose and shape reconstruction from single color image,” in ICCV , 2021, pp. 11 354–11 363
2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
2019
Cited alongside, same era.
C. Doersch and A. Zisserman, “Sim2real transfer learning for 3D human pose estimation: motion to the rescue,” NeurIPS , vol. 32, pp. 12 949–12 961, 2019
2019
Cited alongside, same era.
Y. Rong, Z. Liu, C. Li, K. Cao, and C. C. Loy, “Delving deep into hybrid annotations for 3D human recovery in the wild,” in ICCV , 2019, pp. 5340–5348
2019
Cited alongside, same era.
A. Arnab, C. Doersch, and A. Zisserman, “Exploiting temporal context for 3D human pose estimation in the wild,” in CVPR , 2019, pp. 3395–3404
2019
Cited alongside, same era.
M. Hassan, V. Choutas, D. Tzionas, and M. J. Black, “Resolving 3D human pose ambiguities with 3D scene constraints,” in ICCV , 2019, pp. 2282–2292
2019
Cited alongside, same era.
S. Baek, K. I. Kim, and T.-K. Kim, “Pushing the envelope for RGB-based dense 3D hand pose estimation via neural rendering,” in CVPR , 2019, pp. 1067–1076
2019
Cited alongside, same era.
A. Boukhayma, R. d. Bem, and P. H. Torr, “3D hand shape and pose from images in the wild,” in CVPR , 2019, pp. 10 843–10 852
2019
Cited alongside, same era.
Later among the works it cites.
L. Huang, B. Zhang, Z. Guo, Y. Xiao, Z. Cao, and J. Yuan, “Survey on depth and RGB image-based 3D hand shape and pose estimation,” Virtual Reality & Intelligent Hardware , vol. 3, no. 3, pp. 207–234, 2021
2021
Later among the works it cites.
Y. Feng, H. Feng, M. J. Black, and T. Bolkart, “Learning an animatable detailed 3D face model from in-the-wild images,” TOG , vol. 40, no. 4, pp. 88:1–88:13, 2021
2021
Later among the works it cites.
Y. Rong, T. Shiratori, and H. Joo, “FrankMocap: A monocular 3D whole-body pose estimation system via regression and integration,” in ICCVW , 2021
2021
Later among the works it cites.
S. Guan, J. Xu, Y. Wang, B. Ni, and X. Yang, “Bilevel online adaptation for out-of-domain human mesh reconstruction,” in CVPR , 2021, pp. 10 472–10 481
2021
Later among the works it cites.
Z. Weng and S. Yeung, “Holistic 3D human and scene mesh estimation from single view images,” in CVPR , 2021, pp. 334–343
2021
Later among the works it cites.
H. Cho, Y. Cho, J. Yu, and J. Kim, “Camera distortion-aware 3D human pose estimation in video with optimization-based meta-learning,” in ICCV , 2021, pp. 11 169–11 178
2021
Later among the works it cites.
K. Xie, T. Wang, U. Iqbal, Y. Guo, S. Fidler, and F. Shkurti, “Physics-based human motion estimation and synthesis from videos,” in ICCV , 2021, pp. 11 532–11 541
2021
Later among the works it cites.
M. Fieraru, M. Zanfir, E. Oneata, A.-I. Popa, V. Olaru, and C. Sminchisescu, “Learning complex 3D human self-contact,” in AAAI , 2021
2021
Later among the works it cites.
Y. He, A. Pang, X. Chen, H. Liang, M. Wu, Y. Ma, and L. Xu, “ChallenCap: Monocular 3D capture of challenging human performances using multi-modal references,” in CVPR , 2021, pp. 11 400–11 411
2021
Later among the works it cites.
J. Li, R. Villegas, D. Ceylan, J. Yang, Z. Kuang, H. Li, and Y. Zhao, “Task-generic hierarchical human motion prior using VAEs,” in 3DV . IEEE, 2021, pp. 771–781
2021
Later among the works it cites.
J. Xu, M. Wang, J. Gong, W. Liu, C. Qian, Y. Xie, and L. Ma, “Exploring versatile prior for human motion via motion frequency guidance,” in 3DV . IEEE, 2021, pp. 606–616
2021
Later among the works it cites.
2021
Later among the works it cites.
P. Patel, C.-H. P. Huang, J. Tesch, D. T. Hoffmann, S. Tripathi, and M. J. Black, “AGORA: Avatars in geography optimized for regression analysis,” in CVPR , 2021, pp. 13 468–13 478
2021
Later among the works it cites.
T. Yu, Z. Zheng, K. Guo, P. Liu, Q. Dai, and Y. Liu, “Function4D: Real-time human volumetric capture from very sparse consumer RGBD sensors,” in CVPR , 2021, pp. 5746–5756
2021
Later among the works it cites.
2021
Later among the works it cites.
Q. Fang, Q. Shuai, J. Dong, H. Bao, and X. Zhou, “Reconstructing 3D human pose by watching humans in the mirror,” in CVPR , 2021, pp. 12 814–12 823
2021
Later among the works it cites.
zju3dv, “EasyMoCap - make human motion capture easier.” GitHub, 2021. [Online]. Available: https://github.com/zju3dv/EasyMocap
2021
Later among the works it cites.
M. Liu, D. Yang, Y. Zhang, Z. Cui, J. M. Rehg, and S. Tang, “4D human body capture from egocentric video via 3D scene grounding,” in 3DV . IEEE, 2021, pp. 930–939
2021
Later among the works it cites.
Q. Ma, S. Saito, J. Yang, S. Tang, and M. J. Black, “SCALE: Modeling clothed humans with a surface codec of articulated local elements,” in CVPR , 2021, pp. 16 082–16 093
2021
Later among the works it cites.
G. Moon, H. Choi, and K. M. Lee, “Accurate 3D hand pose estimation for whole-body 3D human mesh estimation,” in CVPRW , 2022, pp. 2308–2317
2022
Closest in time.
Q. Feng, Y. Liu, Y.-K. Lai, J. Yang, and K. Li, “FOF: Learning fourier occupancy field for monocular real-time human reconstruction,” in NeurIPS , 2022
2022
Closest in time.
Y. Xiu, J. Yang, D. Tzionas, and M. J. Black, “ICON: implicit clothed humans obtained from normals,” in CVPR , 2022, pp. 13 286–13 296
2022
Closest in time.
T. Hu, T. Yu, Z. Zheng, H. Zhang, Y. Liu, and M. Zwicker, “HVTR: Hybrid volumetric-textural rendering for human avatars,” in 3DV . IEEE, 2022, pp. 197–208
2022
Closest in time.
Z. Zheng, H. Huang, T. Yu, H. Zhang, Y. Guo, and Y. Liu, “Structured local radiance fields for human avatar modeling,” in CVPR , 2022, pp. 15 893–15 903
2022
Closest in time.
W. Liu, Q. Bao, Y. Sun, and T. Mei, “Recent advances of monocular 2D and 3D human pose estimation: A deep learning perspective,” ACM Computing Surveys , vol. 55, no. 4, pp. 1–41, 2022
2022
Closest in time.
M. Mihajlovic, S. Saito, A. Bansal, M. Zollhoefer, and S. Tang, “COAP: Compositional articulated occupancy of people,” in CVPR , 2022, pp. 13 201–13 210
2022
Closest in time.
A. A. Osman, T. Bolkart, D. Tzionas, and M. J. Black, “SUPR: A sparse unified part-based human representation,” in ECCV . Springer, 2022, pp. 568–585
2022
Closest in time.
Z. Li, J. Liu, Z. Zhang, S. Xu, and Y. Yan, “CLIFF: Carrying location information in full frames into human pose and shape estimation,” in ECCV . Springer, 2022, pp. 590–606
2022
Closest in time.
Z. Li, B. Xu, H. Huang, C. Lu, and Y. Guo, “Deep two-stream video inference for human body pose and shape estimation,” in WACV , 2022, pp. 430–439
2022
Closest in time.
X. Gong, M. Zheng, B. Planche, S. Karanam, T. Chen, D. Doermann, and Z. Wu, “Self-supervised human mesh recovery with cross-representation alignment,” in ECCV . Springer, 2022, pp. 212–230
2022
Closest in time.
H. Choi, G. Moon, J. Park, and K. M. Lee, “Learning to estimate robust 3D human mesh from in-the-wild crowded scenes,” in CVPR , 2022
2022
Closest in time.
G. Pavlakos, J. Malik, and A. Kanazawa, “Human mesh recovery from multiple shots,” in CVPR , 2022, pp. 1485–1495
2022
Closest in time.
J. Cho, K. Youwang, and T.-H. Oh, “Cross-attention of disentangled modalities for 3D human mesh recovery with transformers,” in ECCV . Springer, 2022, pp. 342–359
2022
Closest in time.
V. Choutas, L. Müller, C.-H. P. Huang, S. Tang, D. Tzionas, and M. J. Black, “Accurate 3D body shape regression using metric and semantic attributes,” in CVPR , 2022, pp. 2718–2728
2022
Closest in time.
Z. Wang, J. Yang, and C. Fowlkes, “The best of both worlds: combining model-based and nonparametric approaches for 3D human body estimation,” in CVPRW , 2022, pp. 2318–2327
2022
Closest in time.
J. Cha, M. Saqlain, G. Kim, M. Shin, and S. Baek, “Multi-person 3D pose and shape estimation via inverse kinematics and refinement,” in ECCV . Springer, 2022, pp. 660–677
2022
Closest in time.
R. Khirodkar, S. Tripathi, and K. Kitani, “Occluded human mesh recovery,” in CVPR , 2022, pp. 1715–1725
2022
Closest in time.
W.-L. Wei, J.-C. Lin, T.-L. Liu, and H.-Y. M. Liao, “Capturing humans in motion: temporal-attentive 3D human pose and shape estimation from monocular video,” in CVPR , 2022, pp. 13 211–13 220
2022
Closest in time.
Y. Yuan, U. Iqbal, P. Molchanov, K. Kitani, and J. Kautz, “GLAMR: Global occlusion-aware human mesh recovery with dynamic cameras,” in CVPR , 2022, pp. 11 038–11 049
2022
Closest in time.
J. Park, Y. Oh, G. Moon, H. Choi, and K. M. Lee, “HandOccNet: Occlusion-robust 3D hand mesh estimation network,” in CVPR , 2022, pp. 1496–1505
2022
Closest in time.
M. Li, L. An, H. Zhang, L. Wu, F. Chen, T. Yu, and Y. Liu, “Interacting attention graph for single image two-hand reconstruction,” in CVPR , 2022
2022
Closest in time.
L. Wang, Z. Chen, T. Yu, C. Ma, L. Li, and Y. Liu, “FaceVerse: a fine-grained and detail-controllable 3D face morphable model from a hybrid dataset,” in CVPR , 2022
2022
Closest in time.
W. Zielonka, T. Bolkart, and J. Thies, “Towards metrical reconstruction of human faces,” in ECCV . Springer, 2022, pp. 250–269
2022
Closest in time.
Y. Sun, T. Huang, Q. Bao, W. Liu, G. Wenpeng, and Y. Fu, “Learning monocular mesh recovery of multiple body parts via synthesis,” in ICASSP , 2022
2022
Closest in time.
J. Li, S. Bian, C. Xu, G. Liu, G. Yu, and C. Lu, “D &D: Learning human dynamics from dynamic camera,” in ECCV . Springer, 2022, pp. 479–496
2022
Closest in time.
X. Xie, B. L. Bhatnagar, and G. Pons-Moll, “CHORE: Contact, human and object reconstruction from a single RGB image,” in ECCV . Springer, 2022, pp. 125–145
2022
Closest in time.
H. Yi, C.-H. P. Huang, D. Tzionas, M. Kocabas, M. Hassan, S. Tang, J. Thies, and M. J. Black, “Human-aware object placement for visual environment reconstruction,” in CVPR , 2022, pp. 3959–3970
2022
Closest in time.
Z. Luo, S. Iwase, Y. Yuan, and K. M. Kitani, “Embodied scene-aware human pose estimation,” in NeurIPS , 2022
2022
Closest in time.
E. Gärtner, M. Andriluka, E. Coumans, and C. Sminchisescu, “Differentiable dynamics for articulated 3D human motion reconstruction,” in CVPR , 2022, pp. 13 190–13 200
2022
Closest in time.
E. Gärtner, M. Andriluka, H. Xu, and C. Sminchisescu, “Trajectory optimization for physics-based reconstruction of 3D human pose from monocular video,” in CVPR , 2022, pp. 13 106–13 115
2022
Closest in time.
B. Huang, L. Pan, Y. Yang, J. Ju, and Y. Wang, “Neural MoCon: Neural motion control for physically plausible human motion capture,” in CVPR , 2022, pp. 6417–6426
2022
Closest in time.
Y. Rong, Z. Liu, and C. C. Loy, “Chasing the tail in monocular 3D human reconstruction with prototype memory,” TIP , vol. 31, pp. 2907–2919, 2022
2022
Closest in time.
A. Davydov, A. Remizova, V. Constantin, S. Honari, M. Salzmann, and P. Fua, “Adversarial parametric pose prior,” in CVPR , 2022, pp. 10 997–11 005
2022
Closest in time.
G. Moon, H. Choi, and K. M. Lee, “NeuralAnnot: Neural annotator for 3D human mesh training sets,” in CVPRW , 2022, pp. 2299–2307
2022
Closest in time.
B. L. Bhatnagar, X. Xie, I. A. Petrov, C. Sminchisescu, C. Theobalt, and G. Pons-Moll, “BEHAVE: Dataset and method for tracking human object interactions,” in CVPR , 2022, pp. 15 935–15 946
2022
Closest in time.
C.-H. P. Huang, H. Yi, M. Höschle, M. Safroshkin, T. Alexiadis, S. Polikovsky, D. Scharstein, and M. J. Black, “Capturing and inferring dense full-body human-scene contact,” in CVPR , 2022, pp. 13 274–13 285
2022
Closest in time.
A. Zeng, L. Yang, X. Ju, J. Li, J. Wang, and Q. Xu, “SmoothNet: A plug-and-play network for refining human poses in videos,” in ECCV . Springer, 2022, pp. 625–642
2022
Closest in time.
S. Lin, H. Zhang, Z. Zheng, R. Shao, and Y. Liu, “Learning implicit templates for point-based clothed human modeling,” in ECCV . Springer, 2022, pp. 210–228
2022
Closest in time.
T. Alldieck, M. Zanfir, and C. Sminchisescu, “Photorealistic monocular 3D reconstruction of humans wearing clothing,” in CVPR , 2022, pp. 1506–1515
2022
Closest in time.
R. Shao, H. Zhang, H. Zhang, M. Chen, Y. Cao, T. Yu, and Y. Liu, “DoubleField: Bridging the neural surface and radiance fields for high-fidelity human reconstruction and rendering,” in CVPR , 2022
2022
Closest in time.
Y. Feng, J. Yang, M. Pollefeys, M. J. Black, and T. Bolkart, “SCARF: Capturing and animation of body and clothing from monocular video,” in SIGGRAPH Asia Conference Papers , 2022, p. 9
2022
Closest in time.
G. Moon, H. Nam, T. Shiratori, and K. M. Lee, “3D clothed human reconstruction in the wild,” in ECCV . Springer, 2022, pp. 184–200
2022
Closest in time.
Y. Xiu, J. Yang, X. Cao, D. Tzionas, and M. J. Black, “ECON: Explicit clothed humans optimized via normal integration,” in CVPR , 2023, pp. 512–523
2023
Closest in time.
Z. Zheng, X. Zhao, H. Zhang, B. Liu, and Y. Liu, “AvatarReX: Real-time expressive full-body avatars,” ACM TOG , vol. 42, no. 4, 2023
2023
Closest in time.
X. Sun, Q. Feng, X. Li, J. Zhang, Y.-K. Lai, J. Yang, and K. Li, “Learning semantic-aware disentangled representation for flexible 3D human body editing,” in CVPR , 2023
2023
Closest in time.
J. Li, S. Bian, Q. Liu, J. Tang, F. Wang, and C. Lu, “NIKI: Neural inverse kinematics with invertible neural networks for 3D human pose and shape estimation,” in CVPR , 2023, pp. 12 933–12 942
2023
Closest in time.
K. Shetty, A. Birkhold, S. Jaganathan, N. Strobel, M. Kowarschik, A. Maier, and B. Egger, “PLIKS: A pseudo-linear inverse kinematic solver for 3D human body estimation,” in CVPR , 2023, pp. 574–584
2023
Closest in time.
D. Wang and S. Zhang, “3D human mesh recovery with sequentially global rotation estimation,” in ICCV , 2023, pp. 14 953–14 962
2023
Closest in time.
Q. Fang, K. Chen, Y. Fan, Q. Shuai, J. Li, and W. Zhang, “Learning analytical posterior probability for human mesh recovery,” in CVPR , 2023, pp. 8781–8791
2023
Closest in time.
A. Sengupta, I. Budvytis, and R. Cipolla, “HuManiFlow: Ancestor-conditioned normalising flows on SO (3) manifolds for human pose and shape distribution estimation,” in CVPR , 2023, pp. 4779–4789
2023
Closest in time.
H. Cho, Y. Cho, J. Ahn, and J. Kim, “Implicit 3D human mesh recovery using consistency with pose and shape from unseen-view,” in CVPR , 2023, pp. 21 148–21 158
2023
Closest in time.
S. Goel, G. Pavlakos, J. Rajasegaran, A. Kanazawa, and J. Malik, “Humans in 4D: Reconstructing and tracking humans with transformers,” in ICCV , 2023
2023
Closest in time.
C. Zheng, X. Liu, G.-J. Qi, and C. Chen, “POTTER: Pooling attention transformer for efficient human mesh recovery,” in CVPR , 2023, pp. 1611–1620
2023
Closest in time.
Z. Dou, Q. Wu, C. Lin, Z. Cao, Q. Wu, W. Wan, T. Komura, and W. Wang, “TORE: Token reduction for efficient human mesh recovery with transformer,” in ICCV , 2023, pp. 15 143–15 155
2023
Closest in time.
J. Kim, M.-G. Gwon, H. Park, H. Kwon, G.-M. Um, and W. Kim, “Sampling is matter: Point-guided 3D human mesh reconstruction,” in CVPR , 2023, pp. 12 880–12 889
2023
Closest in time.
Y. Yoshiyasu, “Deformable mesh transformer for 3D human mesh recovery,” in CVPR , 2023, pp. 17 006–17 015
2023
Closest in time.
J. Li, Z. Yang, X. Wang, J. Ma, C. Zhou, and Y. Yang, “JOTR: 3D joint contrastive learning with transformers for occluded human mesh recovery,” in ICCV , 2023, pp. 9110–9121
2023
Closest in time.
X. Ma, J. Su, C. Wang, W. Zhu, and Y. Wang, “3D human mesh estimation from virtual markers,” in CVPR , 2023, pp. 534–543
2023
Closest in time.
H. Zhang, Y. Tian, Y. Zhang, M. Li, L. An, Z. Sun, and Y. Liu, “PyMAF-X: Towards well-aligned full-body model regression from monocular images,” TPAMI , 2023
2023
Closest in time.
Z. Qiu, Q. Yang, J. Wang, H. Feng, J. Han, E. Ding, C. Xu, D. Fu, and J. Wang, “PSVT: End-to-end multi-person 3D pose and shape estimation with progressive video transformers,” in CVPR , 2023, pp. 21 254–21 263
2023
Closest in time.
C. Wang, F. Zhu, and S. Wen, “MeMaHand: Exploiting mesh-mano interaction for single image two-hand reconstruction,” in CVPR , 2023, pp. 564–573
2023
Closest in time.
J. Lee, M. Sung, H. Choi, and T.-K. Kim, “Im2Hands: Learning attentive implicit representation of interacting two-hand shapes,” in CVPR , 2023, pp. 21 169–21 178
2023
Closest in time.
Z. Yu, S. Huang, C. Fang, T. P. Breckon, and J. Wang, “ACR: Attention collaboration-based regressor for arbitrary two-hand reconstruction,” in CVPR , 2023, pp. 12 955–12 964
2023
Closest in time.
G. Moon, “Bringing inputs to shared domains for 3D interacting hands recovery in the wild,” in CVPR , 2023, pp. 17 028–17 037
2023
Closest in time.
H. Yi, H. Liang, Y. Liu, Q. Cao, Y. Wen, T. Bolkart, D. Tao, and M. J. Black, “Generating holistic 3D human motion from speech,” in CVPR , 2023, pp. 469–480
2023
Closest in time.
N. Zioulis and J. F. O’Brien, “KBody: Towards general, robust, and aligned monocular whole-body estimation,” in CVPRW , 2023, pp. 6214–6224
2023
Closest in time.
2023
Closest in time.
J. Lin, A. Zeng, H. Wang, L. Zhang, and Y. Li, “One-stage 3D whole-body mesh recovery with component aware transformer,” in CVPR , 2023, pp. 21 159–21 168
2023
Closest in time.
Z. Cai, W. Yin, A. Zeng, C. Wei, Q. Sun, Y. Wang, H. E. Pang, H. Mei, M. Zhang, L. Zhang, C. C. Loy, L. Yang, and Z. Liu, “SMPLer-X: Scaling up expressive human pose and shape estimation,” in NeurIPS Datasets and Benchmarks , 2023
2023
Closest in time.
H. E. Pang, Z. Cai, L. Yang, T. Qingyi, W. Zhonghua, T. Zhang, and Z. Liu, “Towards robust and expressive whole-body human pose and shape estimation,” NeurIPS , 2023
2023
Closest in time.
M.-P. Forte, P. Kulits, C.-H. P. Huang, V. Choutas, D. Tzionas, K. J. Kuchenbecker, and M. J. Black, “Reconstructing signing avatars from video using linguistic priors,” in CVPR , 2023, pp. 12 791–12 801
2023
Closest in time.
H. Wen, J. Huang, H. Cui, H. Lin, Y.-K. Lai, L. Fang, and K. Li, “Crowd3D: Towards hundreds of people reconstruction from a single image,” in CVPR , 2023, pp. 8937–8946
2023
Closest in time.
B. Zhang, K. Ma, S. Wu, and Z. Yuan, “Two-stage co-segmentation network based on discriminative representation for recovering human mesh from videos,” in CVPR , 2023, pp. 5662–5670
2023
Closest in time.
X. Shen, Z. Yang, X. Wang, J. Ma, C. Zhou, and Y. Yang, “Global-to-local modeling for video-based 3D human pose and shape estimation,” in CVPR , 2023, pp. 8887–8896
2023
Closest in time.
H. Cho, J. Ahn, Y. Cho, and J. Kim, “Video inference for human mesh recovery with vision transformer,” in FG , 2023, pp. 1–6
2023
Closest in time.
H. Nam, D. S. Jung, Y. Oh, and K. M. Lee, “Cyclic test-time adaptation on monocular video for 3D human mesh reconstruction,” in ICCV , 2023, pp. 14 829–14 839
2023
Closest in time.
V. Ye, G. Pavlakos, J. Malik, and A. Kanazawa, “Decoupling human and camera motion from videos in the wild,” in CVPR , 2023, pp. 21 222–21 232
2023
Closest in time.
Y. Sun, Q. Bao, W. Liu, T. Mei, and M. J. Black, “TRACE: 5D temporal regression of avatars with dynamic cameras in 3D environments,” in CVPR , 2023, pp. 8856–8866
2023
Closest in time.
Z. Shen, Z. Cen, S. Peng, Q. Shuai, H. Bao, and X. Zhou, “Learning human mesh recovery in 3D scenes,” in CVPR , 2023, pp. 17 038–17 047
2023
Closest in time.
W. Wang, Y. Ge, H. Mei, Z. Cai, Q. Sun, Y. Wang, C. Shen, L. Yang, and T. Komura, “Zolly: Zoom focal length correctly for perspective-distorted human mesh reconstruction,” in ICCV , 2023, pp. 3925–3935
2023
Closest in time.
S. Tripathi, L. Müller, C.-H. P. Huang, O. Taheri, M. J. Black, and D. Tzionas, “3D human pose estimation via intuitive physics,” in CVPR , 2023, pp. 4713–4725
2023
Closest in time.
2023
Closest in time.
L. G. Foo, J. Gong, H. Rahmani, and J. Liu, “Distribution-aligned diffusion for human mesh recovery,” in ICCV , 2023, pp. 9221–9232
2023
Closest in time.
H. Cho and J. Kim, “Generative approach for probabilistic human mesh recovery using diffusion models,” in ICCV Workshops , 2023, pp. 4183–4188
2023
Closest in time.
S. Zhang, Q. Ma, Y. Zhang, S. Aliakbarian, D. Cosker, and S. Tang, “Probabilistic human mesh recovery in 3d scenes from egocentric views,” in ICCV , 2023, pp. 7989–8000
2023
Closest in time.
M. J. Black, P. Patel, J. Tesch, and J. Yang, “BEDLAM: A synthetic dataset of bodies exhibiting detailed lifelike animated motion,” in CVPR , 2023, pp. 8726–8737
2023
Closest in time.
Z. Yang, Z. Cai, H. Mei, S. Liu, Z. Chen, W. Xiao, Y. Wei, Z. Qing, C. Wei, B. Dai, W. Wu, C. Qian, D. Lin, Z. Liu, and L. Yang, “SynBody: Synthetic dataset with layered human models for 3D human perception and modeling,” in ICCV , 2023, pp. 20 282–20 292
2023
Closest in time.
Y. Dai, Y. Lin, X. Lin, C. Wen, L. Xu, H. Yi, S. Shen, Y. Ma, and C. Wang, “SLOPER4D: A scene-aware dataset for global 4D human pose estimation in urban environments,” in CVPR , 2023, pp. 682–692
2023
Closest in time.
J. Lin, A. Zeng, S. Lu, Y. Cai, R. Zhang, H. Wang, and L. Zhang, “Motion-X: A large-scale 3D expressive whole-body human motion dataset,” NeurIPS , 2023
2023
Closest in time.
H. Zhang, S. Lin, R. Shao, Y. Zhang, Z. Zheng, H. Huang, Y. Guo, and Y. Liu, “CloSET: Modeling clothed humans on continuous surface with explicit template decomposition,” in CVPR , 2023, pp. 501–511
2023
Closest in time.