Fetching the paper…
Reading the bibliography…
We present a real-time approach for multi-person 3D motion capture at over 30 fps using a single RGB camera.
Performance animation from low-dimensional control signals
Jinxiang Chai and Jessica K Hodgins. 2005 · 2005
Earlier work this paper cites.
Surface capture for performance-based animation
Jonathan Starck and Adrian Hilton. 2007 · 2007
Earlier work this paper cites.
Estimating human shape and pose from a single image. In CVPR . 1381–1388
Peng Guan, A. Weiss, A. O. Bãlan, and M. J. Black. 2009 · 2009
Earlier work this paper cites.
MovieReshape: Tracking and Reshaping of Humans in Videos
Arjun Jain, Thorsten Thormählen, Hans-Peter Seidel, and Christian Theobalt. 2010 · 2010
Earlier work this paper cites.
Clustered Pose and Nonlinear Appearance Models for Human Pose Estimation. In BMVC
Sam Johnson and Mark Everingham. 2010 · 2010
Earlier work this paper cites.
Understanding Motion Capture for Computer Animation, Second Edition (2nd ed.)
Alberto Menache. 2010 · 2010
Earlier work this paper cites.
Humaneva: Synchronized video and motion capture dataset and baseline algorithm for evaluation of articulated human motion
Leonid Sigal, Alexandru O Balan, and Michael J Black. 2010 · 2010
Earlier work this paper cites.
VideoMocap: Modeling Physically Realistic Human Motion from Monocular Video Sequences
Xiaolin Wei and Jinxiang Chai. 2010 · 2010
Earlier work this paper cites.
Learning Effective Human Pose Estimation from Inaccurate Annotation. In CVPR
Sam Johnson and Mark Everingham. 2011 · 2011
Earlier work this paper cites.
Fast articulated motion tracking using a sums of Gaussians body model. In ICCV . 951–958
Carsten Stoll, Nils Hasler, Juergen Gall, Hans-Peter Seidel, and Christian Theobalt. 2011 · 2011
Earlier work this paper cites.
Articulated part-based model for joint object detection and pose estimation. In ICCV . IEEE, 723–730
Min Sun and Silvio Savarese. 2011 · 2011
Earlier work this paper cites.
1 € Filter: A Simple Speed-based Low-pass Filter for Noisy Input in Interactive Systems (CHI ’12) . ACM, New York, NY, USA, 2527–2530
Géry Casiez, Nicolas Roussel, and Daniel Vogel. 2012 · 2012
Earlier work this paper cites.
Articulated people detection and pose estimation: Reshaping the future. In CVPR . IEEE, 3178–3185
Leonid Pishchulin, Arjun Jain, Mykhaylo Andriluka, Thorsten Thormählen, and Bernt Schiele. 2012 · 2012
Earlier work this paper cites.
Reconstructing 3d human pose from 2d image landmarks. In ECCV . Springer, 573–586
Varun Ramakrishna, Takeo Kanade, and Yaser Sheikh. 2012 · 2012
Earlier work this paper cites.
2D Human Pose Estimation: New Benchmark and State of the Art Analysis. In CVPR
Mykhaylo Andriluka, Leonid Pishchulin, Peter Gehler, and Bernt Schiele. 2014 · 2014
Earlier work this paper cites.
Using k-poselets for detecting people and localizing their keypoints. In CVPR . 3582–3589
Georgia Gkioxari, Bharath Hariharan, Ross Girshick, and Jitendra Malik. 2014 · 2014
Earlier work this paper cites.
Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu. 2014 · 2014
Earlier work this paper cites.
3d human pose estimation from monocular images with deep convolutional neural network. In ACCV
Sijin Li and Antoni B Chan. 2014 · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context. In ECCV . Springer, 740–755
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Posebits for monocular human pose estimation. In CVPR . 2337–2344
Gerard Pons-Moll, David J Fleet, and Bodo Rosenhahn. 2014 · 2014
Earlier work this paper cites.
Panoptic Studio: A Massively Multiview System for Social Motion Capture. In ICCV
Lei Tan Lin Gui Bart Nabbe Iain Matthews Takeo Kanade Shohei Nobuhara Hanbyul Joo, Hao Liu and Yaser Sheikh. 2015 · 2015
Earlier work this paper cites.
Panoptic Studio: A Massively Multiview System for Social Motion Capture. In ICCV . 3334–3342
Hanbyul Joo, Hao Liu, Lei Tan, Lin Gui, Bart Nabbe, Iain Matthews, Takeo Kanade, Shohei Nobuhara, and Yaser Sheikh. 2015 · 2015
Earlier work this paper cites.
Maximum-margin structured learning with deep networks for 3d human pose estimation. In ICCV . 2848–2856
Sijin Li, Weichen Zhang, and Antoni B Chan. 2015 · 2015
Earlier work this paper cites.
SMPL: A Skinned Multi-Person Linear Model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J Black. 2015 · 2015
Earlier work this paper cites.
Faster R-CNN: Towards real-time object detection with region proposal networks. In NeurIPS . 91–99
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015 · 2015
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei. 2015 · 2015
Earlier work this paper cites.
Keep it SMPL: Automatic Estimation of 3D Human Pose and Shape from a Single Image. In ECCV
Federica Bogo, Angjoo Kanazawa, Christoph Lassner, Peter Gehler, Javier Romero, and Michael J. Black. 2016 · 2016
Earlier work this paper cites.
Synthesizing Training Images for Boosting Human 3D Pose Estimation. In 3DV
Wenzheng Chen, Huan Wang, Yangyan Li, Hao Su, Zhenhua Wang, Changhe Tu, Dani Lischinski, Daniel Cohen-Or, and Baoquan Chen. 2016 · 2016
Earlier work this paper cites.
MARCOnI - ConvNet-based MARker-less Motion Capture in Outdoor and Indoor Scenes
A. Elhayek, E. Aguiar, A. Jain, J. Tompson, L. Pishchulin, M. Andriluka, C. Bregler, B. Schiele, and C. Theobalt. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition. In CVPR
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and < < 0.5MB model size
Forrest N. Iandola, Song Han, Matthew W. Moskewicz, Khalid Ashraf, William J. Dally, and Kurt Keutzer. 2016 · 2016
Earlier work this paper cites.
Multi-person pose estimation with local joint-to-person associations. In ECCV Workshops . Springer, 627–642
Umar Iqbal and Juergen Gall. 2016 · 2016
Earlier work this paper cites.
DeepCut: Joint Subset Partition and Labeling for Multi Person Pose Estimation. In CVPR
Leonid Pishchulin, Eldar Insafutdinov, Siyu Tang, Bjoern Andres, Mykhaylo Andriluka, Peter Gehler, and Bernt Schiele. 2016 · 2016
Earlier work this paper cites.
EgoCap: Egocentric Marker-less Motion Capture with Two Fisheye Cameras
Helge Rhodin, Christian Richardt, Dan Casas, Eldar Insafutdinov, Mohammad Shafiei, Hans-Peter Seidel, Bernt Schiele, and Christian Theobalt. 2016a · 2016
Cited alongside, same era.
3D Human pose estimation: A review of the literature and analysis of covariates
Nikolaos Sarafianos, Bogdan Boteanu, Bogdan Ionescu, and Ioannis A Kakadiaris. 2016 · 2016
Cited alongside, same era.
Rethinking the inception architecture for computer vision. In Proceedings of the IEEE conference on computer vision and pattern recognition . 2818–2826
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna. 2016 · 2016
Cited alongside, same era.
Structured Prediction of 3D Human Pose with Deep Neural Networks. In BMVC
Bugra Tekin, Isinsu Katircioglu, Mathieu Salzmann, Vincent Lepetit, and Pascal Fua. 2016 · 2016
Cited alongside, same era.
Human Pose Estimation from Video and IMUs
Timo von Marcard, Gerard Pons-Moll, and Bodo Rosenhahn. 2016 · 2016
Cited alongside, same era.
Shufflenet v2: Practical guidelines for efficient cnn architecture design. In Proceedings of the European Conference on Computer Vision (ECCV) . 116–131
Ningning Ma, Xiangyu Zhang, Hai-Tao Zheng, and Jian Sun. 2018 · 2018
Later among the works it cites.
Neural Body Fitting: Unifying Deep Learning and Model Based Human Pose and Shape Estimation. In 3DV
Mohamed Omran, Christop Lassner, Gerard Pons-Moll, Peter Gehler, and Bernt Schiele. 2018 · 2018
Later among the works it cites.
Exploiting temporal information for 3d human pose estimation. In Proceedings of the European Conference on Computer Vision (ECCV) . 68–84
Mir Rayat Imtiaz Hossain and James J Little. 2018 · 2018
Later among the works it cites.
Erfnet: Efficient residual factorized convnet for real-time semantic segmentation
Eduardo Romera, José M Alvarez, Luis M Bergasa, and Roberto Arroyo. 2018 · 2018
Later among the works it cites.
Mobilenetv2: Inverted residuals and linear bottlenecks. In CVPR . IEEE, 4510–4520
Mark Sandler, Andrew Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Binarized convolutional landmark localizers for human pose estimation and face alignment with limited resources. In International Conference on Computer Vision
Adrian Bulat and Georgios Tzimiropoulos. 2017 · 2017
Cited alongside, same era.
Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields. In CVPR
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh. 2017 · 2017
Cited alongside, same era.
Xception: Deep learning with depthwise separable convolutions. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1251–1258
François Chollet. 2017 · 2017
Cited alongside, same era.
Towards accurate marker-less human shape and pose estimation over time. In 2017 International Conference on 3D Vision (3DV) . IEEE, 421–430
Yinghao Huang, Federica Bogo, Christoph Lassner, Angjoo Kanazawa, Peter V Gehler, Javier Romero, Ijaz Akhter, and Michael J Black. 2017a · 2017
Cited alongside, same era.
ArtTrack: Articulated multi-person tracking in the wild. In CVPR
Eldar Insafutdinov, Mykhaylo Andriluka, Leonid Pishchulin, Siyu Tang, Evgeny Levinkov, Bjoern Andres, Bernt Schiele, and Saarland Informatics Campus. 2017 · 2017
Cited alongside, same era.
Unite the People: Closing the Loop Between 3D and 2D Human Representations. In CVPR
Christoph Lassner, Javier Romero, Martin Kiefel, Federica Bogo, Michael J. Black, and Peter V. Gehler. 2017 · 2017
Cited alongside, same era.
A simple yet effective baseline for 3d human pose estimation. In ICCV
Julieta Martinez, Rayat Hossain, Javier Romero, and James J. Little. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Integral human pose regression. In Proceedings of the European Conference on Computer Vision (ECCV) . 529–545
Xiao Sun, Bin Xiao, Fangyin Wei, Shuang Liang, and Yichen Wei. 2018 · 2018
Later among the works it cites.
Recovering Accurate 3D Human Pose in The Wild Using IMUs and a Moving Camera. In ECCV
Timo von Marcard, Roberto Henschel, Michael Black, Bodo Rosenhahn, and Gerard Pons-Moll. 2018 · 2018
Later among the works it cites.
Monoperfcap: Human performance capture from monocular video
Weipeng Xu, Avishek Chatterjee, Michael Zollhöfer, Helge Rhodin, Dushyant Mehta, Hans-Peter Seidel, and Christian Theobalt. 2018 · 2018
Later among the works it cites.
Recovering 3D Planes from a Single Image via Convolutional Neural Networks. In ECCV . 85–100
Fengting Yang and Zihan Zhou. 2018 · 2018
Later among the works it cites.
3d human pose estimation in the wild by adversarial learning. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , Vol. 1
Wei Yang, Wanli Ouyang, Xiaolong Wang, Jimmy Ren, Hongsheng Li, and Xiaogang Wang. 2018 · 2018
Later among the works it cites.
Shufflenet: An extremely efficient convolutional neural network for mobile devices. In Proceedings of the IEEE conference on computer vision and pattern recognition . 6848–6856
Xiangyu Zhang, Xinyu Zhou, Mengxiao Lin, and Jian Sun. 2018 · 2018
Later among the works it cites.
Learning to Reconstruct People in Clothing from a Single RGB Camera. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Thiemo Alldieck, Marcus Magnor, Bharat Lal Bhatnagar, Christian Theobalt, and Gerard Pons-Moll. 2019 · 2019
Closest in time.
Exploiting temporal context for 3D human pose estimation in the wild. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 3395–3404
Anurag Arnab, Carl Doersch, and Andrew Zisserman. 2019 · 2019
Closest in time.
Multi-Garment Net: Learning to Dress 3D People from Images. In IEEE International Conference on Computer Vision (ICCV) . IEEE
Bharat Lal Bhatnagar, Garvita Tiwari, Christian Theobalt, and Gerard Pons-Moll. 2019 · 2019
Closest in time.
OpenPose: Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields
Zhe Cao, Gines Hidalgo Martinez, Tomas Simon, Shih-En Wei, and Yaser A Sheikh. 2019 · 2019
Closest in time.
Multi-Person 3D Human Pose Estimation from Monocular Images. In 2019 International Conference on 3D Vision (3DV) . IEEE, 405–414
Rishabh Dabral, Nitesh B Gundavarapu, Rahul Mitra, Abhishek Sharma, Ganesh Ramakrishnan, and Arjun Jain. 2019 · 2019
Closest in time.
Holopose: Holistic 3d human reconstruction in-the-wild. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 10884–10894
Riza Alp Guler and Iasonas Kokkinos. 2019 · 2019
Closest in time.
What Face and Body Shapes Can Tell Us About Height. In Proceedings of the IEEE International Conference on Computer Vision Workshops . 0–0
Semih Gunel, Helge Rhodin, and Pascal Fua. 2019 · 2019
Closest in time.
LiveCap: Real-time Human Performance Capture from Monocular Video
Marc Habermann, Weipeng Xu, , Michael Zollhoefer, Gerard Pons-Moll, and Christian Theobalt. 2019 · 2019
Closest in time.
Searching for mobilenetv3. In Proceedings of the IEEE International Conference on Computer Vision . 1314–1324
Andrew Howard, Mark Sandler, Grace Chu, Liang-Chieh Chen, Bo Chen, Mingxing Tan, Weijun Wang, Yukun Zhu, Ruoming Pang, Vijay Vasudevan, et al · 2019
Closest in time.
Learning 3D Human Dynamics from Video. In Computer Vision and Pattern Recognition (CVPR)
Angjoo Kanazawa, Jason Y. Zhang, Panna Felsen, and Jitendra Malik. 2019 · 2019
Closest in time.
Camera Distance-aware Top-down Approach for 3D Multi-person Pose Estimation from a Single RGB Image. In The IEEE Conference on International Conference on Computer Vision (ICCV)
Gyeongsik Moon, Juyong Chang, and Kyoung Mu Lee. 2019 · 2019
Closest in time.
3D Human Pose Estimation with 2D Marginal Heatmaps. In WACV
Aiden Nibali, Zhen He, Stuart Morgan, and Luke Prendergast. 2019 · 2019
Closest in time.
Expressive Body Capture: 3D Hands, Face, and Body from a Single Image
Georgios Pavlakos, Vasileios Choutas, Nima Ghorbani, Timo Bolkart, Ahmed A. A. Osman, Dimitrios Tzionas, and Michael J. Black. 2019 · 2019
Closest in time.
Regularized evolution for image classifier architecture search
Esteban Real, Alok Aggarwal, Yanping Huang, and Quoc V Le. 2019 · 2019
Closest in time.
LCR-Net++: Multi-person 2D and 3D Pose Detection in Natural Images
Grégory Rogez, Philippe Weinzaepfel, and Cordelia Schmid. 2019 · 2019
Closest in time.
Efficientnet: Rethinking model scaling for convolutional neural networks
Mingxing Tan and Quoc V Le. 2019 · 2019
Closest in time.
xR-EgoPose: Egocentric 3D Human Pose from an HMD Camera. In Proceedings of the IEEE International Conference on Computer Vision . 7728–7738
Denis Tome, Patrick Peluse, Lourdes Agapito, and Hernan Badino. 2019 · 2019
Closest in time.
Deep High-Resolution Representation Learning for Visual Recognition
Jingdong Wang, Ke Sun, Tianheng Cheng, Borui Jiang, Chaorui Deng, Yang Zhao, Dong Liu, Yadong Mu, Mingkui Tan, Xinggang Wang, Wenyu Liu, and Bin Xiao. 2019 · 2019
Closest in time.
Monocular total capture: Posing face, body, and hands in the wild. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 10965–10974
Donglai Xiang, Hanbyul Joo, and Yaser Sheikh. 2019 · 2019
Closest in time.
VIBE: Video Inference for Human Body Pose and Shape Estimation. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Muhammed Kocabas, Nikos Athanasiou, and Michael J. Black. 2020 · 2020
Closest in time.
Mo 2 Cap 2: Real-time Mobile 3D Motion Capture with a Cap-mounted Fisheye Camera
Weipeng Xu, Avishek Chatterjee, Michael Zollhoefer, Helge Rhodin, Pascal Fua, Hans-Peter Seidel, and Christian Theobalt. 2019 · 2093
Closest in time.