Fetching the paper…
Reading the bibliography…
In this work we explore reconstructing hand-object interactions in the wild.
The theory of affordances
James J Gibson · 1977
Earlier work this paper cites.
Understanding the intentions of others: re-enactment of intended acts by 18-month-old children
Andrew N Meltzoff · 1995
Earlier work this paper cites.
A global geometric framework for nonlinear dimensionality reduction
Joshua B Tenenbaum, Vin De Silva, and John C Langford · 2000
Earlier work this paper cites.
Hands in action: real-time 3d reconstruction of hands in interaction with objects
Javier Romero, Hedvig Kjellström, and Danica Kragic · 2010
Earlier work this paper cites.
Full dof tracking of a hand interacting with an object by modeling occlusions and physical constraints
Iason Oikonomidis, Nikolaos Kyriazis, and Antonis A Argyros · 2011
Earlier work this paper cites.
Parsing ikea objects: Fine pose estimation
Joseph J Lim, Hamed Pirsiavash, and Antonio Torralba · 2013
Earlier work this paper cites.
Interactive markerless articulated hand motion tracking using rgb and depth data
Srinath Sridhar, Antti Oulasvirta, and Christian Theobalt · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
The ycb object and model set: Towards common benchmarks for manipulation research
Berk Calli, Arjun Singh, Aaron Walsman, Siddhartha Srinivasa, Pieter Abbeel, and Aaron M Dollar · 2015
Earlier work this paper cites.
Accurate, robust, and flexible real-time hand tracking
Toby Sharp, Cem Keskin, Duncan Robertson, Jonathan Taylor, Jamie Shotton, David Kim, Christoph Rhemann, Ido Leichter, Alon Vinnikov, Yichen Wei, et al · 2015
Earlier work this paper cites.
Robust articulated-icp for real-time hand tracking
Andrea Tagliasacchi, Matthias Schröder, Anastasia Tkach, Sofien Bouaziz, Mario Botsch, and Mark Pauly · 2015
Earlier work this paper cites.
Amodal instance segmentation
Ke Li and Jitendra Malik · 2016
Earlier work this paper cites.
Real-time joint tracking of a hand manipulating an object from rgb-d input
Srinath Sridhar, Franziska Mueller, Michael Zollhöfer, Dan Casas, Antti Oulasvirta, and Christian Theobalt · 2016
Earlier work this paper cites.
Capturing hands in action using discriminative salient points and physics simulation
Dimitrios Tzionas, Luca Ballan, Abhilash Srikantha, Pablo Aponte, Marc Pollefeys, and Juergen Gall · 2016
Earlier work this paper cites.
Spatial attention deep net with partial pso for hierarchical hybrid hand pose estimation
Qi Ye, Shanxin Yuan, and Tae-Kyun Kim · 2016
Earlier work this paper cites.
Global hypothesis generation for 6d object pose estimation
Frank Michel, Alexander Kirillov, Eric Brachmann, Alexander Krull, Stefan Gumhold, Bogdan Savchynskyy, and Carsten Rother · 2017
Earlier work this paper cites.
Embodied hands: Modeling and capturing hands and bodies together
Javier Romero, Dimitrios Tzionas, and Michael J Black · 2017
Earlier work this paper cites.
Hand keypoint detection in single images using multiview bootstrapping
Tomas Simon, Hanbyul Joo, Iain Matthews, and Yaser Sheikh · 2017
Earlier work this paper cites.
Semantic amodal segmentation
Yan Zhu, Yuandong Tian, Dimitris Metaxas, and Piotr Dollár · 2017
Earlier work this paper cites.
Learning to estimate 3d hand pose from single rgb images
Christian Zimmermann and Thomas Brox · 2017
Earlier work this paper cites.
Weakly-supervised 3d hand pose estimation from monocular rgb images
Yujun Cai, Liuhao Ge, Jianfei Cai, and Junsong Yuan · 2018
Earlier work this paper cites.
Scaling egocentric vision: The epic-kitchens dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al · 2018
Cited alongside, same era.
First-person hand action benchmark with RGB-D videos and 3D hand pose annotations
Guillermo Garcia-Hernando, Shanxin Yuan, Seungryul Baek, and Tae-Kyun Kim · 2018
Cited alongside, same era.
Neural 3d mesh renderer
Hiroharu Kato, Yoshitaka Ushiku, and Tatsuya Harada · 2018
Cited alongside, same era.
3d-rcnn: Instance-level 3d object reconstruction via render-and-compare
Abhijit Kundu, Yin Li, and James M Rehg · 2018
Cited alongside, same era.
V2v-posenet: Voxel-to-voxel prediction network for accurate 3d hand and human pose estimation from a single depth map
Gyeongsik Moon, Ju Yong Chang, and Kyoung Mu Lee · 2018
Cited alongside, same era.
Ganerated hands for real-time 3d hand tracking from monocular rgb
Monocular total capture: Posing face, body, and hands in the wild
Donglai Xiang, Hanbyul Joo, and Yaser Sheikh · 2019
Later among the works it cites.
Disentangling latent hands for image synthesis and pose estimation
Linlin Yang and Angela Yao · 2019
Later among the works it cites.
Freihand: A dataset for markerless capture of hand pose and shape from single rgb images
Christian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan Russell, Max Argus, and Thomas Brox · 2019
Later among the works it cites.
ContactPose: A dataset of grasps with object contact and hand pose
Samarth Brahmbhatt, Chengcheng Tang, Christopher D. Twigg, Charles C. Kemp, and James Hays · 2020
Closest in time.
Ganhand: Predicting human grasp affordances in multi-object scenes
Enric Corona, Albert Pumarola, Guillem Alenya, Francesc Moreno-Noguer, and Grégory Rogez · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Franziska Mueller, Florian Bernard, Oleksandr Sotnychenko, Dushyant Mehta, Srinath Sridhar, Dan Casas, and Christian Theobalt · 2018
Cited alongside, same era.
Sfv: Reinforcement learning of physical skills from videos
Xue Bin Peng, Angjoo Kanazawa, Jitendra Malik, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Pix3d: Dataset and methods for single-image 3d shape modeling
Xingyuan Sun, Jiajun Wu, Xiuming Zhang, Zhoutong Zhang, Chengkai Zhang, Tianfan Xue, Joshua B Tenenbaum, and William T Freeman · 2018
Cited alongside, same era.
Factoring shape, pose, and layout from the 2d image of a 3d scene
Shubham Tulsiani, Saurabh Gupta, David Fouhey, Alexei A. Efros, and Jitendra Malik · 2018
Cited alongside, same era.
Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes
Yu Xiang, Tanner Schmidt, Venkatraman Narayanan, and Dieter Fox · 2018
Cited alongside, same era.
Depth-based 3d hand pose estimation: From current achievements to future goals
Shanxin Yuan, Guillermo Garcia-Hernando, Björn Stenger, Gyeongsik Moon, Ju Yong Chang, Kyoung Mu Lee, Pavlo Molchanov, Jan Kautz, Sina Honari, Liuhao Ge, et al · 2018
Cited alongside, same era.
ContactDB: Analyzing and predicting grasp contact via thermal imaging
Samarth Brahmbhatt, Cusuh Ham, Charles C. Kemp, and James Hays · 2019
Cited alongside, same era.
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Antonino Furnari, Evangelos Kazakos, Jian Ma, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al · 2020
Closest in time.
Honnotate: A method for 3d annotation of hand and object poses
Shreyas Hampali, Mahdi Rad Markus Oberweger, and Vincent Lepetit · 2020
Closest in time.
Leveraging photometric consistency over time for sparsely supervised hand-object reconstruction
Yana Hasson, Bugra Tekin, Federica Bogo, Ivan Laptev, Marc Pollefeys, and Cordelia Schmid · 2020
Closest in time.
Coherent reconstruction of multiple humans from a single image
Wen Jiang, Nikos Kolotouros, Georgios Pavlakos, Xiaowei Zhou, and Kostas Daniilidis · 2020
Closest in time.
Pointrend: Image segmentation as rendering
Alexander Kirillov, Yuxin Wu, Kaiming He, and Ross Girshick · 2020
Closest in time.
Weakly-supervised mesh-convolutional hand reconstruction in the wild
Dominik Kulon, Riza Alp Guler, Iasonas Kokkinos, Michael M. Bronstein, and Stefanos Zafeiriou · 2020
Closest in time.
Mask2cad: 3d shape prediction by learning to segment and retrieve
Weicheng Kuo, Anelia Angelova, Tsung-Yi Lin, and Angela Dai · 2020
Closest in time.
Dexterous robotic grasping with object-centric visual affordances
Priyanka Mandikal and Kristen Grauman · 2020
Closest in time.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
René Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun · 2020
Closest in time.
Frankmocap: Fast monocular 3d hand and body motion capture by regression and integration
Yu Rong, Takaaki Shiratori, and Hanbyul Joo · 2020
Closest in time.
Understanding human hands in contact at internet scale
Dandan Shan, Jiaqi Geng, Michelle Shu, and David Fouhey · 2020
Closest in time.
GRAB: A dataset of whole-body human grasping of objects
Omid Taheri, Nima Ghorbani, Michael J. Black, and Dimitrios Tzionas · 2020
Closest in time.
Perceiving 3d human-object spatial arrangements from a single image in the wild
Jason Y. Zhang, Sam Pepose, Hanbyul Joo, Deva Ramanan, Jitendra Malik, and Angjoo Kanazawa · 2020
Closest in time.
Monocular real-time hand shape and motion capture using multi-modal data
Yuxiao Zhou, Marc Habermann, Weipeng Xu, Ikhsanul Habibie, Christian Theobalt, and Feng Xu · 2020
Closest in time.
State-only imitation learning for dexterous manipulation
Ilija Radosavovic, Xiaolong Wang, Lerrel Pinto, and Jitendra Malik · 2021
Closest in time.