Fetching the paper…
Reading the bibliography…
In recent times, there has been a growing interest in developing effective perception techniques for combining information from multiple modalities.
Visualizing data using t-sne
Laurens Van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2010
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2011
Earlier work this paper cites.
Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2013
Earlier work this paper cites.
SMPL: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J. Black · 2015
Earlier work this paper cites.
Keep it smpl: Automatic estimation of 3d human pose and shape from a single image
Federica Bogo, Angjoo Kanazawa, Christoph Lassner, Peter Gehler, Javier Romero, and Michael J Black · 2016
Earlier work this paper cites.
Ntu rgb+d: A large scale dataset for 3d human activity analysis
Amir Shahroudy, Jun Liu, Tian-Tsong Ng, and Gang Wang · 2016
Earlier work this paper cites.
A simple yet effective baseline for 3d human pose estimation
Julieta Martinez, Rayat Hossain, Javier Romero, and James J Little · 2017
Earlier work this paper cites.
Monocular 3d human pose estimation in the wild using improved cnn supervision
Dushyant Mehta, Helge Rhodin, Dan Casas, Pascal Fua, Oleksandr Sotnychenko, Weipeng Xu, and Christian Theobalt · 2017
Earlier work this paper cites.
Coarse-to-fine volumetric prediction for single-image 3d human pose
Georgios Pavlakos, Xiaowei Zhou, Konstantinos G Derpanis, and Kostas Daniilidis · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Posetrack: A benchmark for human pose estimation and tracking
Mykhaylo Andriluka, Umar Iqbal, Eldar Insafutdinov, Leonid Pishchulin, Anton Milan, Juergen Gall, and Bernt Schiele · 2018
Earlier work this paper cites.
Vgpn: Voice-guided pointing robot navigation for humans
Jun Hu, Zhongyu Jiang, Xionghao Ding, Taijiang Mu, and Peter Hall · 2018
Earlier work this paper cites.
End-to-end recovery of human shape and pose
Angjoo Kanazawa, Michael J Black, David W Jacobs, and Jitendra Malik · 2018
Earlier work this paper cites.
2d/3d pose estimation and action recognition using multitask deep learning
Diogo C Luvizon, David Picard, and Hedi Tabia · 2018
Earlier work this paper cites.
Neural body fitting: Unifying deep learning and model-based human pose and shape estimation
Mohamed Omran, Christoph Lassner, Gerard Pons-Moll, Peter V. Gehler, and Bernt Schiele · 2018
Earlier work this paper cites.
Learning to estimate 3d human pose and shape from a single color image
Georgios Pavlakos, Luyang Zhu, Xiaowei Zhou, and Kostas Daniilidis · 2018
Earlier work this paper cites.
Recovering accurate 3d human pose in the wild using imus and a moving camera
Timo Von Marcard, Roberto Henschel, Michael J Black, Bodo Rosenhahn, and Gerard Pons-Moll · 2018
Earlier work this paper cites.
Human computer interaction with head pose, eye gaze and body gestures
Kang Wang, Rui Zhao, and Qiang Ji · 2018
Earlier work this paper cites.
Exploiting temporal context for 3d human pose estimation in the wild
Anurag Arnab, Carl Doersch, and Andrew Zisserman · 2019
Earlier work this paper cites.
Sim2real transfer learning for 3d human pose estimation: motion to the rescue
Carl Doersch and Andrew Zisserman · 2019
Cited alongside, same era.
Holopose: Holistic 3d human reconstruction in-the-wild
Riza Alp Guler and Iasonas Kokkinos · 2019
Cited alongside, same era.
Learning 3d human dynamics from video
Angjoo Kanazawa, Jason Y Zhang, Panna Felsen, and Jitendra Malik · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al · 2019
Cited alongside, same era.
3d human pose estimation in video with temporal convolutions and semi-supervised training
Dario Pavllo, Christoph Feichtenhofer, David Grangier, and Michael Auli · 2019
Cited alongside, same era.
Generative modeling by estimating gradients of the data distribution
Deep 3d human pose estimation: A review
Jinbao Wang, Shujie Tan, Xiantong Zhen, Shuo Xu, Feng Zheng, Zhenyu He, and Ling Shao · 2021
Later among the works it cites.
Pymaf: 3d human pose and shape regression with pyramidal mesh alignment feedback loop
Hongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang, Yebin Liu, Limin Wang, and Zhenan Sun · 2021
Later among the works it cites.
3d human pose estimation with spatial and temporal transformers
Ce Zheng, Sijie Zhu, Matias Mendieta, Taojiannan Yang, Chen Chen, and Zhengming Ding · 2021
Later among the works it cites.
Pyskl: Towards good practices for skeleton action recognition
Haodong Duan, Jiaqi Wang, Kai Chen, and Dahua Lin · 2022
Later among the works it cites.
Adaptpose: Cross-dataset adaptation for 3d human pose estimation by learnable motion generation
Mohsen Gholami, Bastian Wandt, Helge Rhodin, Rabab Ward, and Z Jane Wang · 2022
Later among the works it cites.
Audioclip: Extending clip to image, text and audio
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yang Song and Stefano Ermon · 2019
Cited alongside, same era.
Human mesh recovery from monocular images via a skeleton-disentangled representation
Yu Sun, Yun Ye, Wu Liu, Wenpeng Gao, Yili Fu, and Tao Mei · 2019
Cited alongside, same era.
Denserac: Joint 3d pose and shape estimation by dense render-and-compare
Yuanlu Xu, Song-Chun Zhu, and Tony Tung · 2019
Cited alongside, same era.
Semantic graph convolutional networks for 3d human pose regression
Long Zhao, Xi Peng, Yu Tian, Mubbasir Kapadia, and Dimitris N Metaxas · 2019
Cited alongside, same era.
Hierarchical kinematic human mesh recovery
Georgios Georgakis, Ren Li, Srikrishna Karanam, Terrence Chen, Jana Košecká, and Ziyan Wu · 2020
Cited alongside, same era.
Hierarchical kinematic human mesh recovery
Georgios Georgakis, Ren Li, Srikrishna Karanam, Terrence Chen, Jana Košecká, and Ziyan Wu · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Andrey Guzhov, Federico Raue, Jörn Hees, and Andreas Dengel · 2022
Later among the works it cites.
Golfpose: Golf swing analyses with a monocular camera based human pose estimation
Zhongyu Jiang, Haorui Ji, Samuel Menaker, and Jenq-Neng Hwang · 2022
Later among the works it cites.
Frozen clip models are efficient video learners
Ziyi Lin, Shijie Geng, Renrui Zhang, Peng Gao, Gerard de Melo, Xiaogang Wang, Jifeng Dai, Yu Qiao, and Hongsheng Li · 2022
Later among the works it cites.
Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning
Huaishao Luo, Lei Ji, Ming Zhong, Yang Chen, Wen Lei, Nan Duan, and Tianrui Li · 2022
Later among the works it cites.
Putting People in their Place: Monocular Regression of 3D People in Depth
Yu Sun, Wu Liu, Qian Bao, Yili Fu, Tao Mei, and Michael J Black · 2022
Later among the works it cites.
Humannerf: Free-viewpoint rendering of moving people from monocular video
Chung-Yi Weng, Brian Curless, Pratul P Srinivasan, Jonathan T Barron, and Ira Kemelmacher-Shlizerman · 2022
Later among the works it cites.
Vitpose: Simple vision transformer baselines for human pose estimation
Yufei Xu, Jing Zhang, Qiming Zhang, and Dacheng Tao · 2022
Later among the works it cites.
Mixste: Seq2seq mixed spatio-temporal encoder for 3d human pose estimation in video
Jinlu Zhang, Zhigang Tu, Jianyu Yang, Yujin Chen, and Junsong Yuan · 2022
Later among the works it cites.
Wenhao Chai, Zhongyu Jiang, Jenq-Neng Hwang, and Gaoang Wang · 2023
Closest in time.
Gfpose: Learning 3d human pose prior with gradient fields
Hai Ci, Mingdong Wu, Wentao Zhu, Xiaoxuan Ma, Hao Dong, Fangwei Zhong, and Yizhou Wang · 2023
Closest in time.
Imagebind: One embedding space to bind them all
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, and Ishan Misra · 2023
Closest in time.
Back to optimization: Diffusion-based zero-shot 3d human pose estimation
Zhongyu Jiang, Zhuoran Zhou, Lei Li, Wenhao Chai, Cheng-Yen Yang, and Jenq-Neng Hwang · 2023
Closest in time.
Hybrik-x: Hybrid analytical-neural inverse kinematics for whole-body mesh recovery
Jiefeng Li, Siyuan Bian, Chao Xu, Zhicun Chen, Lixin Yang, and Cewu Lu · 2023
Closest in time.
A survey of deep learning in sports applications: Perception, comprehension, and decision
Zhonghan Zhao, Wenhao Chai, Shengyu Hao, Wenhao Hu, Guanhong Wang, Shidong Cao, Mingli Song, Jenq-Neng Hwang, and Gaoang Wang · 2023
Closest in time.
Efficient domain adaptation via generative prior for 3d infant pose estimation, 2023
Zhuoran Zhou, Zhongyu Jiang, Wenhao Chai, Cheng-Yen Yang, Lei Li, and Jenq-Neng Hwang · 2023
Closest in time.