Fetching the paper…
Reading the bibliography…
We present an approach that can reconstruct hands in 3D from monocular input.
Articulated human detection with flexible mixtures of parts
Yi Yang and Deva Ramanan · 2012
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
HICO: A benchmark for recognizing human-object interactions in images
Yu-Wei Chao, Zhan Wang, Yugeng He, Jiaxuan Wang, and Jia Deng · 2015
Earlier work this paper cites.
Panoptic studio: A massively multiview system for social motion capture
Hanbyul Joo, Hao Liu, Lei Tan, Lin Gui, Bart Nabbe, Iain Matthews, Takeo Kanade, Shohei Nobuhara, and Yaser Sheikh · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Embodied hands: Modeling and capturing hands and bodies together
Javier Romero, Dimitris Tzionas, and Michael J Black · 2017
Earlier work this paper cites.
Hand keypoint detection in single images using multiview bootstrapping
Tomas Simon, Hanbyul Joo, Iain Matthews, and Yaser Sheikh · 2017
Earlier work this paper cites.
Learning to estimate 3D hand pose from single RGB images
Christian Zimmermann and Thomas Brox · 2017
Earlier work this paper cites.
Learning to detect human-object interactions
Yu-Wei Chao, Yunfan Liu, Xieyang Liu, Huayi Zeng, and Jia Deng · 2018
Earlier work this paper cites.
Scaling egocentric vision: The EPIC-KITCHENS dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, and Michael Wray · 2018
Earlier work this paper cites.
End-to-end recovery of human shape and pose
Angjoo Kanazawa, Michael J Black, David W Jacobs, and Jitendra Malik · 2018
Earlier work this paper cites.
Pushing the envelope for RGB-based dense 3D hand pose estimation via neural rendering
Seungryul Baek, Kwang In Kim, and Tae-Kyun Kim · 2019
Earlier work this paper cites.
3D hand shape and pose from images in the wild
Adnane Boukhayma, Rodrigo de Bem, and Philip HS Torr · 2019
Earlier work this paper cites.
3D hand shape and pose estimation from a single RGB image
Liuhao Ge, Zhou Ren, Yuncheng Li, Zehao Xue, Yingying Wang, Jianfei Cai, and Junsong Yuan · 2019
Earlier work this paper cites.
Learning joint reconstruction of hands and manipulated objects
Yana Hasson, Gul Varol, Dimitrios Tzionas, Igor Kalevatykh, Michael J Black, Ivan Laptev, and Cordelia Schmid · 2019
Earlier work this paper cites.
Learning to reconstruct 3D human pose and shape via model-fitting in the loop
Nikos Kolotouros, Georgios Pavlakos, Michael J Black, and Kostas Daniilidis · 2019
Earlier work this paper cites.
Monocular total capture: Posing face, body, and hands in the wild
Donglai Xiang, Hanbyul Joo, and Yaser Sheikh · 2019
Earlier work this paper cites.
End-to-end hand mesh recovery from a monocular RGB image
Xiong Zhang, Qiang Li, Hong Mo, Wenbo Zhang, and Wen Zheng · 2019
Earlier work this paper cites.
FreiHAND: A dataset for markerless capture of hand pose and shape from single RGB images
Christian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan Russell, Max Argus, and Thomas Brox · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Pose2Mesh: Graph convolutional network for 3D human pose and mesh recovery from a 2D human pose
Hongsuk Choi, Gyeongsik Moon, and Kyoung Mu Lee · 2020
Earlier work this paper cites.
OpenMMLab pose estimation toolbox and benchmark
MMPose Contributors · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2020
Cited alongside, same era.
HOnnotate: A method for 3D annotation of hand and object poses
Shreyas Hampali, Mahdi Rad, Markus Oberweger, and Vincent Lepetit · 2020
Cited alongside, same era.
Leveraging photometric consistency over time for sparsely supervised hand-object reconstruction
Yana Hasson, Bugra Tekin, Federica Bogo, Ivan Laptev, Marc Pollefeys, and Cordelia Schmid · 2020
Cited alongside, same era.
Whole-body human pose estimation in the wild
Sheng Jin, Lumin Xu, Jin Xu, Can Wang, Wentao Liu, Chen Qian, Wanli Ouyang, and Ping Luo · 2020
Cited alongside, same era.
Keypoint transformer: Solving joint identification in challenging hands and object interactions for accurate 3D pose estimation
Shreyas Hampali, Sayan Deb Sarkar, Mahdi Rad, and Vincent Lepetit · 2022
Later among the works it cites.
Interacting attention graph for single image two-hand reconstruction
Mengcheng Li, Liang An, Hongwen Zhang, Lianpeng Wu, Feng Chen, Tao Yu, and Yebin Liu · 2022
Later among the works it cites.
3D interacting hand pose estimation by hand de-occlusion and removal
Hao Meng, Sheng Jin, Wentao Liu, Chen Qian, Mengxiang Lin, Wanli Ouyang, and Ping Luo · 2022
Later among the works it cites.
HandOccNet: Occlusion-robust 3D hand mesh estimation network
JoonKyu Park, Yeonguk Oh, Gyeongsik Moon, Hongsuk Choi, and Kyoung Mu Lee · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dominik Kulon, Riza Alp Guler, Iasonas Kokkinos, Michael M Bronstein, and Stefanos Zafeiriou · 2020
Cited alongside, same era.
I2L-MeshNet: Image-to-lixel prediction network for accurate 3D human pose and mesh estimation from a single RGB image
Gyeongsik Moon and Kyoung Mu Lee · 2020
Cited alongside, same era.
InterHand2.6M: A dataset and baseline for 3D interacting hand pose estimation from a single RGB image
Gyeongsik Moon, Shoou-I Yu, He Wen, Takaaki Shiratori, and Kyoung Mu Lee · 2020
Cited alongside, same era.
Understanding human hands in contact at internet scale
Dandan Shan, Jiaqi Geng, Michelle Shu, and David F Fouhey · 2020
Cited alongside, same era.
DexYCB: A benchmark for capturing hand grasping of objects
Yu-Wei Chao, Wei Yang, Yu Xiang, Pavlo Molchanov, Ankur Handa, Jonathan Tremblay, Yashraj S Narang, Karl Van Wyk, Umar Iqbal, Stan Birchfield, Jan Kautz, and Dieter Fox · 2021
Cited alongside, same era.
I2UV-HandNet: Image-to-UV prediction network for accurate and high-fidelity 3D hand mesh modeling
Ping Chen, Yujin Chen, Dong Yang, Fangyin Wu, Qin Li, Qingpei Xia, and Yong Tan · 2021
Cited alongside, same era.
End-to-end detection and pose estimation of two interacting hands
Dong Uk Kim, Kwang In Kim, and Seungryul Baek · 2021
Cited alongside, same era.
Assembly101: A large-scale multi-view video dataset for understanding procedural activities
Fadime Sener, Dibyadip Chatterjee, Daniel Shelepov, Kun He, Dipika Singhania, Robert Wang, and Angela Yao · 2022
Later among the works it cites.
Collaborative learning for hand and object reconstruction with attention-guided graph convolution
Tze Ho Elden Tse, Kwang In Kim, Ales Leonardis, and Hyung Jin Chang · 2022
Later among the works it cites.
ViTPose: Simple vision transformer baselines for human pose estimation
Yufei Xu, Jing Zhang, Qiming Zhang, and Dacheng Tao · 2022
Later among the works it cites.
ArtiBoost: Boosting articulated 3D hand-object pose estimation via online exploration and synthesis
Lixin Yang, Kailin Li, Xinyu Zhan, Jun Lv, Wenqiang Xu, Jiefeng Li, and Cewu Lu · 2022
Later among the works it cites.
Towards a richer 2D understanding of hands at scale
Tianyi Cheng, Dandan Shan, Ayda Sultan, Jiaqi Geng, Richard EL Higgins, and David F Fouhey · 2023
Closest in time.
Humans in 4D: Reconstructing and tracking humans with transformers
Shubham Goel, Georgios Pavlakos, Jathushan Rajasegaran, Angjoo Kanazawa, and Jitendra Malik · 2023
Closest in time.
A probabilistic attention model with occlusion-aware texture regression for 3D hand reconstruction from a single RGB image
Zheheng Jiang, Hossein Rahmani, Sue Black, and Bryan M Williams · 2023
Closest in time.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick · 2023
Closest in time.
Bringing inputs to shared domains for 3D interacting hands recovery in the wild
Gyeongsik Moon · 2023
Closest in time.
Recovering 3D hand mesh sequence from a single blurry image: A new dataset and temporal unfolding
Yeonguk Oh, JoonKyu Park, Jaeha Kim, Gyeongsik Moon, and Kyoung Mu Lee · 2023
Closest in time.
AssemblyHands: Towards egocentric activity understanding via 3D hand pose estimation
Takehiko Ohkawa, Kun He, Fadime Sener, Tomas Hodan, Luan Tran, and Cem Keskin · 2023
Closest in time.
OpenAI · 2023
Closest in time.
Decoupled iterative refinement framework for interacting hands reconstruction from a single RGB image
Pengfei Ren, Chao Wen, Xiaozheng Zheng, Zhou Xue, Haifeng Sun, Qi Qi, Jingyu Wang, and Jianxin Liao · 2023
Closest in time.
MeMaHand: Exploiting mesh-mano interaction for single image two-hand reconstruction
Congyi Wang, Feida Zhu, and Shilei Wen · 2023
Closest in time.
ACR: Attention collaboration-based regressor for arbitrary two-hand reconstruction
Zhengdi Yu, Shaoli Huang, Chen Fang, Toby P Breckon, and Jue Wang · 2023
Closest in time.
Reconstructing interacting hands with interaction prior from monocular images
Binghui Zuo, Zimeng Zhao, Wenqian Sun, Wei Xie, Zhou Xue, and Yangang Wang · 2023
Closest in time.