Fetching the paper…
Reading the bibliography…
Understanding how humans cooperatively rearrange household objects is critical for VR/AR and human-robot interaction.
Marching cubes: A high resolution 3d surface construction algorithm
William E Lorensen and Harvey E Cline · 1998
Earlier work this paper cites.
Motion capture file formats explained
Maddock Meredith, Steve Maddock, et al · 2001
Earlier work this paper cites.
Optimization based full body control for the atlas robot
Siyuan Feng, Eric Whitman, X Xinjilefu, and Christopher G Atkeson · 2014
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
Angel X Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, et al · 2015
Earlier work this paper cites.
Development of life-sized high-power humanoid robot jaxon for real-world use
Kunio Kojima, Tatsuhi Karasawa, Toyotaka Kozuki, Eisoku Kuroiwa, Sou Yukizaki, Satoshi Iwaishi, Tatsuya Ishikawa, Ryo Koyama, Shintaro Noda, Fumihito Sugai, et al · 2015
Earlier work this paper cites.
Learning structured output representation using deep conditional generative models
Kihyuk Sohn, Honglak Lee, and Xinchen Yan · 2015
Earlier work this paper cites.
Retargeting human-object interaction to virtual avatars
Yeonjoon Kim, Hangil Park, Seungbae Bang, and Sung-Hee Lee · 2016
Earlier work this paper cites.
Pigraphs: learning interaction snapshots from observations
Manolis Savva, Angel X Chang, Pat Hanrahan, Matthew Fisher, and Matthias Nießner · 2016
Earlier work this paper cites.
Malleable embodiment: changing sense of embodiment by spatial-temporal deformation of virtual human body
Shunichi Kasahara, Keina Konno, Richi Owaki, Tsubasa Nishi, Akiko Takeshita, Takayuki Ito, Shoko Kasuga, and Junichi Ushiba · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Deepmimic: Example-guided deep reinforcement learning of physics-based character skills
Xue Bin Peng, Pieter Abbeel, Sergey Levine, and Michiel Van de Panne · 2018
Earlier work this paper cites.
Transferring category-based functional grasping skills by latent space non-rigid registration
Diego Rodriguez and Sven Behnke · 2018
Earlier work this paper cites.
Unitree’s first universal humanoid robot
Unitree · 2018
Earlier work this paper cites.
A multi-sensor dataset of human-human handover
Alessandro Carfì, Francesco Foglino, Barbara Bruno, and Fulvio Mastrogiovanni · 2019
Earlier work this paper cites.
On the choice of grasp type and location when handing over an object
Francesca Cini, V Ortenzi, P Corke, and MJSR Controzzi · 2019
Earlier work this paper cites.
Resolving 3d human pose ambiguities with 3d scene constraints
Mohamed Hassan, Vasileios Choutas, Dimitrios Tzionas, and Michael J Black · 2019
Earlier work this paper cites.
Deepsdf: Learning continuous signed distance functions for shape representation
Jeong Joon Park, Peter Florence, Julian Straub, Richard Newcombe, and Steven Lovegrove · 2019
Earlier work this paper cites.
Expressive body capture: 3d hands, face, and body from a single image
Georgios Pavlakos, Vasileios Choutas, Nima Ghorbani, Timo Bolkart, Ahmed AA Osman, Dimitrios Tzionas, and Michael J Black · 2019
Earlier work this paper cites.
Neural state machine for character-scene interactions
Sebastian Starke, He Zhang, Taku Komura, and Jun Saito · 2019
Earlier work this paper cites.
Convolutional sequence generation for skeleton-based action synthesis
Sijie Yan, Zhizhong Li, Yuanjun Xiong, Huahan Yan, and Dahua Lin · 2019
Earlier work this paper cites.
Long-term human motion prediction with scene context
Zhe Cao, Hang Gao, Karttikeya Mangalam, Qi-Zhi Cai, Minh Vo, and Jitendra Malik · 2020
Earlier work this paper cites.
An affordance and distance minimization based method for computing object orientations for robot human handovers
Wesley P Chan, Matthew KXJ Pan, Elizabeth A Croft, and Masayuki Inaba · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Earlier work this paper cites.
Catch & carry: reusable neural controllers for vision-guided whole-body tasks
Josh Merel, Saran Tunyasuvunakool, Arun Ahuja, Yuval Tassa, Leonard Hasenclever, Vu Pham, Tom Erez, Greg Wayne, and Nicolas Heess · 2020
Earlier work this paper cites.
Grab: A dataset of whole-body human grasping of objects
Omid Taheri, Nima Ghorbani, Michael J Black, and Dimitrios Tzionas · 2020
Earlier work this paper cites.
Predicting camera viewpoint improves cross-dataset generalization for 3d human pose estimation
Zhe Wang, Daeyun Shin, and Charless C Fowlkes · 2020
Earlier work this paper cites.
Tripod: Human trajectory and pose dynamics forecasting in the wild
Vida Adeli, Mahsa Ehsanpour, Ian Reid, Juan Carlos Niebles, Silvio Savarese, Ehsan Adeli, and Hamid Rezatofighi · 2021
Earlier work this paper cites.
Dexycb: A benchmark for capturing hand grasping of objects
Yu-Wei Chao, Wei Yang, Yu Xiang, Pavlo Molchanov, Ankur Handa, Jonathan Tremblay, Yashraj S Narang, Karl Van Wyk, Umar Iqbal, Stan Birchfield, et al · 2021
Earlier work this paper cites.
Gravity-aware monocular 3d human-object reconstruction
Rishabh Dabral, Soshi Shimada, Arjun Jain, Christian Theobalt, and Vladislav Golyanik · 2021
Earlier work this paper cites.
Human poseitioning system (hps): 3d human pose estimation and self-localization in large scenes from body-mounted sensors
Vladimir Guzov, Aymen Mir, Torsten Sattler, and Gerard Pons-Moll · 2021
Earlier work this paper cites.
4d human body capture from egocentric video via 3d scene grounding
Miao Liu, Dexin Yang, Yan Zhang, Zhaopeng Cui, James M Rehg, and Siyu Tang · 2021
Earlier work this paper cites.
Isaac gym: High performance gpu-based physics simulation for robot learning
Viktor Makoviychuk, Lukasz Wawrzyniak, Yunrong Guo, Michelle Lu, Kier Storey, Miles Macklin, David Hoeller, Nikita Rudin, Arthur Allshire, Ankur Handa, et al · 2021
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T Barron, Ravi Ramamoorthi, and Ren Ng · 2021
Earlier work this paper cites.
Humanoid loco-manipulation planning based on graph search and reachability maps
Masaki Murooka, Iori Kumagai, Mitsuharu Morisawa, Fumio Kanehiro, and Abderrahmane Kheddar · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Cpf: Learning a contact potential field to model the hand-object interaction
Lixin Yang, Xinyu Zhan, Kailin Li, Wenqiang Xu, Jiefeng Li, and Cewu Lu · 2021
Cited alongside, same era.
H2o: A benchmark for visual human-human object handover analysis
Ruolin Ye, Wenqiang Xu, Zhendong Xue, Tutian Tang, Yanfeng Wang, and Cewu Lu · 2021
Cited alongside, same era.
Behave: Dataset and method for tracking human object interactions
Bharat Lal Bhatnagar, Xianghui Xie, Ilya A Petrov, Cristian Sminchisescu, Christian Theobalt, and Gerard Pons-Moll · 2022
Cited alongside, same era.
Learning robust real-world dexterous grasping policies via implicit shape augmentation
Zoey Qiuyu Chen, Karl Van Wyk, Yu-Wei Chao, Wei Yang, Arsalan Mousavian, Abhishek Gupta, and Dieter Fox · 2022
Cited alongside, same era.
Yunze Liu, Changxi Chen, and Li Yi · 2023
Later among the works it cites.
Perpetual humanoid control for real-time simulated avatars
Zhengyi Luo, Jinkun Cao, Kris Kitani, Weipeng Xu, et al · 2023
Later among the works it cites.
Online adaptive motion generation for humanoid locomotion on non-flat terrain via template behavior extension
Xiang Meng, Zhangguo Yu, Xuechao Chen, Zelin Huang, Fei Meng, and Qiang Huang · 2023
Later among the works it cites.
It takes two: Learning to plan for human-robot cooperative carrying
Eley Ng, Ziang Liu, and Monroe Kennedy · 2023
Later among the works it cites.
Hoi-diff: Text-driven synthesis of 3d human-object interactions using diffusion models
Xiaogang Peng, Yiming Xie, Zizhao Wu, Varun Jampani, Deqing Sun, and Huaizu Jiang · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D-grasp: Physically plausible dynamic grasp synthesis for hand-object interactions
Sammy Christen, Muhammed Kocabas, Emre Aksan, Jemin Hwangbo, Jie Song, and Otmar Hilliges · 2022
Cited alongside, same era.
Hsc4d: Human-centered 4d scene capture in large-scale indoor-outdoor space using wearable imus and lidar
Yudi Dai, Yitai Lin, Chenglu Wen, Siqi Shen, Lan Xu, Jingyi Yu, Yuexin Ma, and Cheng Wang · 2022
Cited alongside, same era.
Multi-person extreme motion prediction
Wen Guo, Xiaoyu Bie, Xavier Alameda-Pineda, and Francesc Moreno-Noguer · 2022
Cited alongside, same era.
Interaction replica: Tracking human-object interaction and scene changes from human motion
Vladimir Guzov, Julian Chibane, Riccardo Marin, Yannan He, Torsten Sattler, and Gerard Pons-Moll · 2022
Cited alongside, same era.
Mocapdeform: Monocular 3d human motion capture in deformable scenes
Zhi Li, Soshi Shimada, Bernt Schiele, Christian Theobalt, and Vladislav Golyanik · 2022
Cited alongside, same era.
Hoi4d: A 4d egocentric dataset for category-level human-object interaction
Yunze Liu, Yun Liu, Che Jiang, Kangbo Lyu, Weikang Wan, Hao Shen, Boqiang Liang, Zhoujie Fu, He Wang, and Li Yi · 2022
Cited alongside, same era.
Contact-aware human motion forecasting
Wei Mao, Richard I Hartley, Mathieu Salzmann, et al · 2022
Cited alongside, same era.
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Unidexgrasp++: Improving dexterous grasping policy learning via geometry-aware curriculum and iterative generalist-specialist learning
Weikang Wan, Haoran Geng, Yun Liu, Zikang Shan, Yaodong Yang, Li Yi, and He Wang · 2023
Later among the works it cites.
Physhoi: Physics-based imitation of dynamic human-object interaction
Yinhuai Wang, Jing Lin, Ailing Zeng, Zhengyi Luo, Jian Zhang, and Lei Zhang · 2023
Later among the works it cites.
Functional grasp transfer across a category of objects from only one labeled instance
Rina Wu, Tianqiang Zhu, Wanli Peng, Jinglue Hang, and Yi Sun · 2023
Later among the works it cites.
Cimi4d: A large multimodal climbing motion dataset under human-scene interactions
Ming Yan, Xin Wang, Yudi Dai, Siqi Shen, Chenglu Wen, Lan Xu, Yuexin Ma, and Cheng Wang · 2023
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
Tony Z Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn · 2023
Later among the works it cites.
Cams: Canonicalized manipulation spaces for category-level functional hand-object manipulation synthesis
Juntian Zheng, Qingyuan Zheng, Lixing Fang, Yun Liu, and Li Yi · 2023
Later among the works it cites.
Sim-to-real learning for humanoid box loco-manipulation
Jeremy Dao, Helei Duan, and Alan Fern · 2024
Closest in time.
Textim: Part-aware interactive motion synthesis from text
Siyuan Fan, Bo Du, Xiantao Cai, Bo Peng, and Longling Sun · 2024
Closest in time.
Humanplus: Humanoid shadowing and imitation from humans
Zipeng Fu, Qingqing Zhao, Qi Wu, Gordon Wetzstein, and Chelsea Finn · 2024
Closest in time.
Spatial and surface correspondence field for interaction transfer
Zeyu Huang, Honghao Xu, Haibin Huang, Chongyang Ma, Hui Huang, and Ruizhen Hu · 2024
Closest in time.
Generating continual human motion in diverse 3d scenes
Aymen Mir, Xavier Puig, Angjoo Kanazawa, and Gerard Pons-Moll · 2024
Closest in time.
Noitom motion capture systems
INC NOITOM INTERNATIONAL · 2024
Closest in time.
Humanoid locomotion as next token prediction
Ilija Radosavovic, Bike Zhang, Baifeng Shi, Jathushan Rajasegaran, Sarthak Kamat, Trevor Darrell, Koushil Sreenath, and Jitendra Malik · 2024
Closest in time.
Hoianimator: Generating text-prompt human-object animations using novel perceptive diffusion models
Wenfeng Song, Xinyu Zhang, Shuai Li, Yang Gao, Aimin Hao, Xia Hou, Chenglizhao Chen, Ning Li, and Hong Qin · 2024
Closest in time.
Humanmimic: Learning natural locomotion and transitions for humanoid robot via wasserstein adversarial imitation
Annan Tang, Takuma Hiraoka, Naoki Hiraoka, Fan Shi, Kento Kawaharazuka, Kunio Kojima, Kei Okada, and Masayuki Inaba · 2024
Closest in time.
Humans in kitchens: A dataset for multi-person human motion forecasting with scene context
Julian Tanke, Oh-Hun Kwon, Felix B Mueller, Andreas Doering, and Juergen Gall · 2024
Closest in time.
Maskedmimic: Unified physics-based character control through masked motion inpainting
Chen Tessler, Yunrong Guo, Ofir Nabati, Gal Chechik, and Xue Bin Peng · 2024
Closest in time.
A multimodal handover failure detection dataset and baselines
Santosh Thoduka, Nico Hochgeschwender, Juergen Gall, and Paul G Plöger · 2024
Closest in time.
Move as you say interact as you can: Language-guided human motion generation with scene affordance
Zan Wang, Yixin Chen, Baoxiong Jia, Puhao Li, Jinlu Zhang, Jingze Zhang, Tengyu Liu, Yixin Zhu, Wei Liang, and Siyuan Huang · 2024
Closest in time.
Hoh: Markerless multimodal human-object-human handover dataset with large object count
Noah Wiederhold, Ava Megyeri, DiMaggio Paris, Sean Banerjee, and Natasha Banerjee · 2024
Closest in time.
Interdreamer: Zero-shot text to 3d dynamic human-object interaction
Sirui Xu, Ziyin Wang, Yu-Xiong Wang, and Liang-Yan Gui · 2024
Closest in time.
F-hoi: Toward fine-grained semantic-aligned 3d human-object interactions
Jie Yang, Xuesong Niu, Nan Jiang, Ruimao Zhang, and Siyuan Huang · 2024
Closest in time.
Generating human interaction motions in scenes with text control
Hongwei Yi, Justus Thies, Michael J Black, Xue Bin Peng, and Davis Rempe · 2024
Closest in time.
Oakink2: A dataset of bimanual hands-object manipulation in complex task completion
Xinyu Zhan, Lixin Yang, Yifei Zhao, Kangrui Mao, Hanlin Xu, Zenan Lin, Kailin Li, and Cewu Lu · 2024
Closest in time.
Artigrasp: Physically plausible synthesis of bi-manual dexterous grasping and articulation
Hui Zhang, Sammy Christen, Zicong Fan, Luocheng Zheng, Jemin Hwangbo, Jie Song, and Otmar Hilliges · 2024
Closest in time.
I’m hoi: Inertia-aware monocular capture of 3d human-object interactions
Chengfeng Zhao, Juze Zhang, Jiashen Du, Ziwei Shan, Junye Wang, Jingyi Yu, Jingya Wang, and Lan Xu · 2024
Closest in time.
Dreamhoi: Subject-driven generation of 3d human-object interactions with diffusion priors
Thomas Hanwen Zhu, Ruining Li, and Tomas Jakab · 2024
Closest in time.
Himo: A new benchmark for full-body human interacting with multiple objects
Xintao Lv, Liang Xu, Yichao Yan, Xin Jin, Congsheng Xu, Shuwen Wu, Yifan Liu, Lincheng Li, Mengxiao Bi, Wenjun Zeng, et al · 2025
Closest in time.
Generating human interaction motions in scenes with text control
Hongwei Yi, Justus Thies, Michael J Black, Xue Bin Peng, and Davis Rempe · 2025
Closest in time.