Fetching the paper…
Reading the bibliography…
Videos from edited media like movies are a useful, yet under-explored source of information.
Midwest and its children: The psychological ecology of an american town
Roger G Barker and Herbert F Wright · 1955
Earlier work this paper cites.
Grammar of the film language. 1976
Daniel Arijon · 1976
Earlier work this paper cites.
SCAPE: shape completion and animation of people
Dragomir Anguelov, Praveen Srinivasan, Daphne Koller, Sebastian Thrun, Jim Rodgers, and James Davis · 2005
Earlier work this paper cites.
Estimating human shape and pose from a single image
Peng Guan, Alexander Weiss, Alexandru O Balan, and Michael J Black · 2009
Earlier work this paper cites.
HumanEva: Synchronized video and motion capture dataset and baseline algorithm for evaluation of articulated human motion
Leonid Sigal, Alexandru O Balan, and Michael J Black · 2010
Earlier work this paper cites.
Temporal video segmentation to scenes using high-level audiovisual features
Panagiotis Sidiropoulos, Vasileios Mezaris, Ioannis Kompatsiaris, Hugo Meinedo, Miguel Bugalho, and Isabel Trancoso · 2011
Earlier work this paper cites.
Articulated human detection with flexible mixtures of parts
Yi Yang and Deva Ramanan · 2012
Earlier work this paper cites.
Human3.6m: Large scale datasets and predictive methods for 3D human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2013
Earlier work this paper cites.
From actemes to action: A strongly-supervised representation for detailed action understanding
Weiyu Zhang, Menglong Zhu, and Konstantinos G Derpanis · 2013
Earlier work this paper cites.
2D human pose estimation: New benchmark and state of the art analysis
Mykhaylo Andriluka, Leonid Pishchulin, Peter Gehler, and Bernt Schiele · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
SMPL: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J Black · 2015
Earlier work this paper cites.
Keep it SMPL: Automatic estimation of 3D human pose and shape from a single image
Federica Bogo, Angjoo Kanazawa, Christoph Lassner, Peter Gehler, Javier Romero, and Michael J Black · 2016
Earlier work this paper cites.
RMPE: Regional multi-person pose estimation
Hao-Shu Fang, Shuqin Xie, Yu-Wing Tai, and Cewu Lu · 2017
Earlier work this paper cites.
Towards accurate marker-less human shape and pose estimation over time
Yinghao Huang, Federica Bogo, Christoph Lassner, Angjoo Kanazawa, Peter V Gehler, Javier Romero, Ijaz Akhter, and Michael J Black · 2017
Earlier work this paper cites.
Unite the people: Closing the loop between 3D and 2D human representations
Christoph Lassner, Javier Romero, Martin Kiefel, Federica Bogo, Michael J Black, and Peter V Gehler · 2017
Earlier work this paper cites.
Monocular 3D human pose estimation in the wild using improved cnn supervision
Dushyant Mehta, Helge Rhodin, Dan Casas, Pascal Fua, Oleksandr Sotnychenko, Weipeng Xu, and Christian Theobalt · 2017
Earlier work this paper cites.
Self-supervised learning of motion capture
Hsiao-Yu Tung, Hsiao-Wei Tung, Ersin Yumer, and Katerina Fragkiadaki · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
AVA: A video dataset of spatio-temporally localized atomic visual actions
Chunhui Gu, Chen Sun, David A Ross, Carl Vondrick, Caroline Pantofaru, Yeqing Li, Sudheendra Vijayanarasimhan, George Toderici, Susanna Ricco, Rahul Sukthankar, Cordelia Schmid, and Jitendra Malik · 2018
Cited alongside, same era.
Person search in videos with one portrait through visual and temporal links
Qingqiu Huang, Wentao Liu, and Dahua Lin · 2018
Cited alongside, same era.
Total capture: A 3D deformation model for tracking faces, hands, and bodies
Hanbyul Joo, Tomas Simon, and Yaser Sheikh · 2018
Cited alongside, same era.
End-to-end recovery of human shape and pose
Angjoo Kanazawa, Michael J Black, David W Jacobs, and Jitendra Malik · 2018
Cited alongside, same era.
Neural body fitting: Unifying deep learning and model based human pose and shape estimation
Human mesh recovery from monocular images via a skeleton-disentangled representation
Yu Sun, Yun Ye, Wu Liu, Wenpeng Gao, YiLi Fu, and Tao Mei · 2019
Later among the works it cites.
DenseRac: Joint 3D pose and shape estimation by dense render-and-compare
Yuanlu Xu, Song-Chun Zhu, and Tony Tung · 2019
Later among the works it cites.
Predicting 3D human dynamics from video
Jason Y Zhang, Panna Felsen, Angjoo Kanazawa, and Jitendra Malik · 2019
Later among the works it cites.
Pose2Mesh: Graph convolutional network for 3Dd human pose and mesh recovery from a 2D human pose
Hongsuk Choi, Gyeongsik Moon, and Kyoung Mu Lee · 2020
Closest in time.
Motion capture from internet videos
Junting Dong, Qing Shuai, Yuanqing Zhang, Xian Liu, Xiaowei Zhou, and Hujun Bao · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mohamed Omran, Christoph Lassner, Gerard Pons-Moll, Peter Gehler, and Bernt Schiele · 2018
Cited alongside, same era.
Learning to estimate 3D human pose and shape from a single color image
Georgios Pavlakos, Luyang Zhu, Xiaowei Zhou, and Kostas Daniilidis · 2018
Cited alongside, same era.
Sfv: Reinforcement learning of physical skills from videos
Xue Bin Peng, Angjoo Kanazawa, Jitendra Malik, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Recovering accurate 3D human pose in the wild using IMUs and a moving camera
Timo von Marcard, Roberto Henschel, Michael J Black, Bodo Rosenhahn, and Gerard Pons-Moll · 2018
Cited alongside, same era.
Monocular 3D pose and shape estimation of multiple people in natural scenes-the importance of multiple scene constraints
Andrei Zanfir, Elisabeta Marinoiu, and Cristian Sminchisescu · 2018
Cited alongside, same era.
Exploiting temporal context for 3D human pose estimation in the wild
Anurag Arnab, Carl Doersch, and Andrew Zisserman · 2019
Cited alongside, same era.
OpenPose: realtime multi-person 2D pose estimation using part affinity fields
Zhe Cao, Gines Hidalgo, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2019
Cited alongside, same era.
Georgios Georgakis, Ren Li, Srikrishna Karanam, Terrence Chen, Jana Kosecka, and Ziyan Wu · 2020
Closest in time.
Movienet: A holistic dataset for movie understanding
Qingqiu Huang, Yu Xiong, Anyi Rao, Jiaze Wang, and Dahua Lin · 2020
Closest in time.
Exemplar fine-tuning for 3D human pose fitting towards in-the-wild 3D human pose estimation
Hanbyul Joo, Natalia Neverova, and Andrea Vedaldi · 2020
Closest in time.
VIBE: Video inference for human body pose and shape estimation
Muhammed Kocabas, Nikos Athanasiou, and Michael J Black · 2020
Closest in time.
3D human motion estimation via motion compression and refinement
Zhengyi Luo, S Alireza Golestaneh, and Kris M Kitani · 2020
Closest in time.
Gyeongsik Moon and Kyoung Mu Lee · 2020
Closest in time.
A local-to-global approach to multi-modal movie scene segmentation
Anyi Rao, Linning Xu, Yu Xiong, Guodong Xu, Qingqiu Huang, Bolei Zhou, and Dahua Lin · 2020
Closest in time.
Full-body awareness from partial observations
Chris Rockwell and David F Fouhey · 2020
Closest in time.
STAR: Sparse trained articulated human body regressor
Ahmed AA sman, Timo Bolkart, and Michael J Black · 2020
Closest in time.
Human body model fitting by learned gradient descent
Jie Song, Xu Chen, and Otmar Hilliges · 2020
Closest in time.
BLSM: A bone-level skinned model of the human mesh
Haoyang Wang, Riza Alp Güler, Iasonas Kokkinos, George Papandreou, and Stefanos Zafeiriou · 2020
Closest in time.
GHUM & GHUMl: Generative 3D human shape and articulated pose models
Hongyi Xu, Eduard Gabriel Bazavan, Andrei Zanfir, William T Freeman, Rahul Sukthankar, and Cristian Sminchisescu · 2020
Closest in time.