Fetching the paper…
Reading the bibliography…
To enable machines to understand the way humans interact with the physical world in daily life, 3D interaction signals should be captured in natural settings, allowing people to engage with multiple objects in a range of sequential and casual manipulations.
Tracking of humans in action: A 3-D model-based approach
D. Gavrila and LS Davis · 1996
Earlier work this paper cites.
Virtualized reality: Constructing virtual worlds from real scenes
Takeo Kanade, Peter Rander, and P.J. Narayanan · 1997
Earlier work this paper cites.
Image-based visual hulls
Wojciech Matusik, Chris Buehler, Ramesh Raskar, Steven J. Gortler, and Leonard McMillan · 2000
Earlier work this paper cites.
The cmu motion of body (mobo) database
Ralph Gross and Jianbo Shi · 2001
Earlier work this paper cites.
Generation, visualization, and editing of 3d video
T. Matsuyama and T. Takai · 2002
Earlier work this paper cites.
Blue-c: A spatially immersive display and 3d video portal for telepresence
Markus Gross, Stephan Würmlin, Martin Naef, Edouard Lamboray, Christian Spagno, Andreas Kunz, Esther Koller-Meier, Tomas Svoboda, Luc Van Gool, Silke Lang, Kai Strehlke, Andrew Vande Moere, and Oliver Staadt · 2003
Earlier work this paper cites.
Articulated Soft Objects for Multi-View Shape and Motion Capture
Ralf Plankers and Pascal Fua · 2003
Earlier work this paper cites.
Twist based acquisition and tracking of animal and human kinematics
Christoph Bregler, Jitendra Malik, and Katherine Pullen · 2004
Earlier work this paper cites.
Shape-from-silhouette across time part i: Theory and algorithms
Kong Man Cheung, Simon Baker, and Takeo Kanade · 2005
Earlier work this paper cites.
Markerless tracking of complex human motions from multiple views
Roland Kehl and Luc Van Gool · 2006
Earlier work this paper cites.
Performance capture from sparse multi-view video
Edilson de Aguiar, Carsten Stoll, Christian Theobalt, Naveed Ahmed, Hans-Peter Seidel, and Sebastian Thrun · 2008
Earlier work this paper cites.
Dense 3d motion capture from synchronized video streams
Y. Furukawa and J. Ponce · 2008
Earlier work this paper cites.
Motion capture using joint skeleton tracking and surface estimation
Juergen Gall, Carsten Stoll, Edilson De Aguiar, Christian Theobalt, Bodo Rosenhahn, and Hans-Peter Seidel · 2009
Earlier work this paper cites.
Virtualization Gate
Benjamin Petit, Jean-Denis Lesage, Edmond Boyer, and Bruno Raffin · 2009
Earlier work this paper cites.
Combined region and motion-based 3D tracking of rigid and articulated objects
Thomas Brox, Bodo Rosenhahn, Juergen Gall, and Daniel Cremers · 2010
Earlier work this paper cites.
Markerless Motion Capture through Visual Hull, Articulated ICP and Subject Specific Model Generation
Stefano Corazza, Lars Mündermann, Emiliano Gambaretto, Giancarlo Ferrigno, and Thomas P. Andriacchi · 2010
Earlier work this paper cites.
Humaneva: Synchronized video and motion capture dataset and baseline algorithm for evaluation of articulated human motion
Leonid Sigal, Alexandru O Balan, and Michael J Black · 2010
Earlier work this paper cites.
Fast articulated motion tracking using a sums of gaussians body model
Carsten Stoll, Nils Hasler, Juergen Gall, Hans-Peter Seidel, and Christian Theobalt · 2011
Earlier work this paper cites.
Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2013
Earlier work this paper cites.
Automatic generation and detection of highly reliable fiducial markers under occlusion
Sergio Garrido-Jurado, Rafael Muñoz-Salinas, Francisco José Madrid-Cuevas, and Manuel Jesús Marín-Jiménez · 2014
Earlier work this paper cites.
Panoptic studio: A massively multiview system for social motion capture
Hanbyul Joo, Hao Liu, Lei Tan, Lin Gui, Bart Nabbe, Iain Matthews, Takeo Kanade, Shohei Nobuhara, and Yaser Sheikh · 2015
Earlier work this paper cites.
Extending FABRIK with model constraints
Andreas Aristidou, Yiorgos Chrysanthou, and Joan Lasenby · 2016
Earlier work this paper cites.
A deep learning framework for character motion synthesis and editing
Daniel Holden, Jun Saito, and Taku Komura · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L Schonberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Realtime multi-person 2d pose estimation using part affinity fields
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2017
Earlier work this paper cites.
Temporal convolutional networks for action segmentation and detection
Colin Lea, Michael D Flynn, Rene Vidal, Austin Reiter, and Gregory D Hager · 2017
Earlier work this paper cites.
First-person hand action benchmark with rgb-d videos and 3d hand pose annotations
Guillermo Garcia-Hernando, Shanxin Yuan, Seungryul Baek, and Tae-Kyun Kim · 2018
Earlier work this paper cites.
Online optical marker-based hand tracking with deep labels
Shangchen Han, Beibei Liu, Robert Wang, Yuting Ye, Christopher D Twigg, and Kenrick Kin · 2018
Earlier work this paper cites.
An empirical rig for jaw animation
Gaspard Zoss, Derek Bradley, Pascal Bérard, and Thabo Beeler · 2018
Earlier work this paper cites.
Resolving 3d human pose ambiguities with 3d scene constraints
Mohamed Hassan, Vasileios Choutas, Dimitrios Tzionas, and Michael J Black · 2019
Earlier work this paper cites.
Amass: Archive of motion capture as surface shapes
Naureen Mahmood, Nima Ghorbani, Nikolaus F Troje, Gerard Pons-Moll, and Michael J Black · 2019
Earlier work this paper cites.
Expressive body capture: 3d hands, face, and body from a single image
Georgios Pavlakos, Vasileios Choutas, Nima Ghorbani, Timo Bolkart, Ahmed AA Osman, Dimitrios Tzionas, and Michael J Black · 2019
Earlier work this paper cites.
On the continuity of rotation representations in neural networks
Yi Zhou, Connelly Barnes, Lu Jingwan, Yang Jimei, and Li Hao · 2019
Earlier work this paper cites.
Freihand: A dataset for markerless capture of hand pose and shape from single rgb images
Christian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan Russell, Max Argus, and Thomas Brox · 2019
Earlier work this paper cites.
Contactpose: A dataset of grasps with object contact and hand pose
Samarth Brahmbhatt, Chengcheng Tang, Christopher D Twigg, Charles C Kemp, and James Hays · 2020
Earlier work this paper cites.
Long-term human motion prediction with scene context
Zhe Cao, Hang Gao, Karttikeya Mangalam, Qi-Zhi Cai, Minh Vo, and Jitendra Malik · 2020
Earlier work this paper cites.
Ganhand: Predicting human grasp affordances in multi-object scenes
Enric Corona, Albert Pumarola, Guillem Alenya, Francesc Moreno-Noguer, and Grégory Rogez · 2020
Cited alongside, same era.
Hope-net: A graph-based model for hand-object pose estimation
Bardia Doosti, Shujon Naha, Majid Mirbagheri, and David J Crandall · 2020
Cited alongside, same era.
Honnotate: A method for 3d annotation of hand and object poses
Shreyas Hampali, Mahdi Rad, Markus Oberweger, and Vincent Lepetit · 2020
Cited alongside, same era.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2020
Cited alongside, same era.
Grab: A dataset of whole-body human grasping of objects
Omid Taheri, Nima Ghorbani, Michael J Black, and Dimitrios Tzionas · 2020
Cited alongside, same era.
Reconstructing hand-object interactions in the wild
Full-body articulated human-object interaction
Nan Jiang, Tengyu Liu, Zhexuan Cao, Jieming Cui, Zhiyuan Zhang, Yixin Chen, He Wang, Yixin Zhu, and Siyuan Huang · 2023
Later among the works it cites.
Motionscript: Natural language descriptions for expressive 3d human motions
Payam Jome Yazdian, Eric Liu, Li Cheng, and Angelica Lim · 2023
Later among the works it cites.
Emdb: The electromagnetic database of global 3d human pose and shape in the wild
Manuel Kaufmann, Jie Song, Chen Guo, Kaiyue Shen, Tianjian Jiang, Chengcheng Tang, Juan José Zárate, and Otmar Hilliges · 2023
Later among the works it cites.
Segment anything in high quality
Lei Ke, Mingqiao Ye, Martin Danelljan, Yifan Liu, Yu-Wing Tai, Chi-Keung Tang, and Fisher Yu · 2023
Later among the works it cites.
Nifty: Neural object interaction fields for guided human motion synthesis
Nilesh Kulkarni, Davis Rempe, Kyle Genova, Abhijit Kundu, Justin Johnson, David Fouhey, and Leonidas Guibas · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhe Cao, Ilija Radosavovic, Angjoo Kanazawa, and Jitendra Malik · 2021
Cited alongside, same era.
Dexycb: A benchmark for capturing hand grasping of objects
Yu-Wei Chao, Wei Yang, Yu Xiang, Pavlo Molchanov, Ankur Handa, Jonathan Tremblay, Yashraj S Narang, Karl Van Wyk, Umar Iqbal, Stan Birchfield, et al · 2021
Cited alongside, same era.
Contactopt: Optimizing contact to improve grasps
Patrick Grady, Chengcheng Tang, Christopher D Twigg, Minh Vo, Samarth Brahmbhatt, and Charles C Kemp · 2021
Cited alongside, same era.
Human poseitioning system (hps): 3d human pose estimation and self-localization in large scenes from body-mounted sensors
Vladimir Guzov, Aymen Mir, Torsten Sattler, and Gerard Pons-Moll · 2021
Cited alongside, same era.
Stochastic scene-aware motion prediction
Mohamed Hassan, Duygu Ceylan, Ruben Villegas, Jun Saito, Jimei Yang, Yi Zhou, and Michael Black · 2021
Cited alongside, same era.
H2o: Two hands manipulating objects for first person interaction recognition
Taein Kwon, Bugra Tekin, Jan Stühmer, Federica Bogo, and Marc Pollefeys · 2021
Cited alongside, same era.
Semi-supervised 3d hand-object poses estimation with interactions in time
Shaowei Liu, Hanwen Jiang, Jiarui Xu, Sifei Liu, and Xiaolong Wang · 2021
Cited alongside, same era.
Later among the works it cites.
Object motion guided human motion synthesis
Jiaman Li, Jiajun Wu, and C Karen Liu · 2023
Later among the works it cites.
Hybridcap: Inertia-aid monocular capture of challenging human motions
Han Liang, Yannan He, Chengfeng Zhao, Mutian Li, Jingya Wang, Jingyi Yu, and Lan Xu · 2023
Later among the works it cites.
https://www.manus-meta.com/
Manus, 2023 · 2023
Later among the works it cites.
Generating continual human motion in diverse 3d scenes
Aymen Mir, Xavier Puig, Angjoo Kanazawa, and Gerard Pons-Moll · 2023
Later among the works it cites.
https://base.xsens.com/
Movella, 2023 · 2023
Later among the works it cites.
Fusing monocular images and sparse imu signals for real-time human motion capture
Shaohua Pan, Qi Ma, Xinyu Yi, Weifeng Hu, Xiong Wang, Xingkang Zhou, Jijunnan Li, and Feng Xu · 2023
Later among the works it cites.
Hoi-diff: Text-driven synthesis of 3d human-object interactions using diffusion models
Xiaogang Peng, Yiming Xie, Zizhao Wu, Varun Jampani, Deqing Sun, and Huaizu Jiang · 2023
Later among the works it cites.
Object pop-up: Can we infer 3d objects and their poses from human interactions alone?
Ilya A Petrov, Riccardo Marin, Julian Chibane, and Gerard Pons-Moll · 2023
Later among the works it cites.
Interaction replica: Tracking human-object interaction and scene changes from human motion
Gerard Pons-Moll, Vladimir Guzov, Julian Chibane, Riccardo Marin, Yannan He, and Torsten Sattler · 2023
Later among the works it cites.
Flex: Full-body grasping without full-body grasps
Purva Tendulkar, Dídac Surís, and Carl Vondrick · 2023
Later among the works it cites.
Physhoi: Physics-based imitation of dynamic human-object interaction
Yinhuai Wang, Jing Lin, Ailing Zeng, Zhengyi Luo, Jian Zhang, and Lei Zhang · 2023
Later among the works it cites.
Visibility aware human-object interaction tracking from single rgb camera
Xianghui Xie, Bharat Lal Bhatnagar, and Gerard Pons-Moll · 2023
Later among the works it cites.
Interdiff: Generating 3d human-object interactions with physics-informed diffusion
Sirui Xu, Zhengyuan Li, Yu-Xiong Wang, and Liang-Yan Gui · 2023
Later among the works it cites.
Mime: Human-aware 3d scene generation
Hongwei Yi, Chun-Hao P Huang, Shashank Tripathi, Lea Hering, Justus Thies, and Michael J Black · 2023
Later among the works it cites.
https://huggingface.co/black-forest-labs
Black Forest Lab, 2024 · 2024
Closest in time.
Text2hoi: Text-guided 3d motion generation for hand-object interaction
Junuk Cha, Jihyeon Kim, Jae Shin Yoon, and Seungryul Baek · 2024
Closest in time.
Diffh2o: Diffusion-based synthesis of hand-object interactions from textual descriptions
Sammy Christen, Shreyas Hampali, Fadime Sener, Edoardo Remelli, Tomas Hodan, Eric Sauser, Shugao Ma, and Bugra Tekin · 2024
Closest in time.
Scaling up dynamic human-scene interaction modeling
Nan Jiang, Zhiyuan Zhang, Hongjie Li, Xiaoxuan Ma, Zan Wang, Yixin Chen, Tengyu Liu, Yixin Zhu, and Siyuan Huang · 2024
Closest in time.
Mocap everyone everywhere: Lightweight motion capture with smartwatches and a head-mounted camera
Jiye Lee and Hanbyul Joo · 2024
Closest in time.
Controllable human-object interaction synthesis
Jiaman Li, Alexander Clegg, Roozbeh Mottaghi, Jiajun Wu, Xavier Puig, and C Karen Liu · 2024
Closest in time.
Taco: Benchmarking generalizable bimanual tool-action-object understanding
Yun Liu, Haolin Yang, Xu Si, Ling Liu, Zipeng Li, Yuxiang Zhang, Yebin Liu, and Li Yi · 2024
Closest in time.
Grasping diverse objects with simulated humanoids
Zhengyi Luo, Jinkun Cao, Sammy Christen, Alexander Winkler, Kris Kitani, and Weipeng Xu · 2024
Closest in time.
Reconstructing hands in 3d with transformers
Georgios Pavlakos, Dandan Shan, Ilija Radosavovic, Angjoo Kanazawa, David Fouhey, and Jitendra Malik · 2024
Closest in time.
Intertrack: Tracking human object interaction without object templates
Xianghui Xie, Jan Eric Lenssen, and Gerard Pons-Moll · 2024
Closest in time.
Interdreamer: Zero-shot text to 3d dynamic human-object interaction
Sirui Xu, Ziyin Wang, Yu-Xiong Wang, and Liang-Yan Gui · 2024
Closest in time.
F-hoi: Toward fine-grained semantic-aligned 3d human-object interactions
Jie Yang, Xuesong Niu, Nan Jiang, Ruimao Zhang, and Siyuan Huang · 2024
Closest in time.
Generating human interaction motions in scenes with text control
Hongwei Yi, Justus Thies, Michael J. Black, Xue Bin Peng, and Davis Rempe · 2024
Closest in time.
Adl4d: Towards a contextually rich dataset for 4d activities of daily living
Marsil Zakour, Partha Pratim Nath, Ludwig Lohmer, Emre Faik Gökçe, Martin Piccolrovazzi, Constantin Patsch, Yuankai Wu, Rahul Chaudhari, and Eckehard Steinbach · 2024
Closest in time.
Oakink2: A dataset of bimanual hands-object manipulation in complex task completion
Xinyu Zhan, Lixin Yang, Yifei Zhao, Kangrui Mao, Hanlin Xu, Zenan Lin, Kailin Li, and Cewu Lu · 2024
Closest in time.
Gears: Local geometry-aware hand-object interaction synthesis
Keyang Zhou, Bharat Lal Bhatnagar, Jan Eric Lenssen, and Gerard Pons-Moll · 2024
Closest in time.
Omni6d: Large-vocabulary 3d object dataset for category-level 6d object pose estimation
Mengchen Zhang, Tong Wu, Tai Wang, Tengfei Wang, Ziwei Liu, and Dahua Lin · 2025
Closest in time.