Fetching the paper…
Reading the bibliography…
A key component of understanding hand-object interactions is the ability to identify the active object -- the object that is being manipulated by the human hand.
A short introduction to boosting
Yoav Freund, Robert Schapire, and Naoki Abe · 1999
Earlier work this paper cites.
Improving predictive inference under covariate shift by weighting the log-likelihood function
Hidetoshi Shimodaira · 2000
Earlier work this paper cites.
Optimal eye movement strategies in visual search
Jiri Najemnik and Wilson S Geisler · 2005
Earlier work this paper cites.
Depth-encoded hough voting for joint object detection and shape recovery
Min Sun, Gary Bradski, Bing-Xin Xu, and Silvio Savarese · 2010
Earlier work this paper cites.
Imitation learning of hand gestures and its evaluation for humanoid robots
Anand Thobbi and Weihua Sheng · 2010
Earlier work this paper cites.
Understanding egocentric activities
Alireza Fathi, Ali Farhadi, and James M Rehg · 2011
Earlier work this paper cites.
Hand detection using multiple proposals
Arpit Mittal, Andrew Zisserman, and Philip HS Torr · 2011
Earlier work this paper cites.
Full dof tracking of a hand interacting with an object by modeling occlusions and physical constraints
Iason Oikonomidis, Nikolaos Kyriazis, and Antonis A Argyros · 2011
Earlier work this paper cites.
The role of observers’ gaze behaviour when watching object manipulation tasks: predicting and evaluating the consequences of action
J Randall Flanagan, Gerben Rotman, Andreas F Reichelt, and Roland S Johansson · 2013
Earlier work this paper cites.
An attention-based activity recognition for egocentric video
Kenji Matsuo, Kentaro Yamada, Satoshi Ueno, and Sei Naito · 2014
Earlier work this paper cites.
Lending a hand: Detecting hands and recognizing activities in complex egocentric interactions
Sven Bambach, Stefan Lee, David J Crandall, and Chen Yu · 2015
Earlier work this paper cites.
Active object localization with deep reinforcement learning
Juan C Caicedo and Svetlana Lazebnik · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Cited alongside, same era.
Tree-structured reinforcement learning for sequential object localization
Zequn Jie, Xiaodan Liang, Jiashi Feng, Xiaojie Jin, Wen Lu, and Shuicheng Yan · 2016
Cited alongside, same era.
Going deeper into first-person activity recognition
Minghuang Ma, Haoqi Fan, and Kris M Kitani · 2016
Cited alongside, same era.
Reinforcement learning for visual object detection
Stefan Mathe, Aleksis Pirinen, and Cristian Sminchisescu · 2016
Cited alongside, same era.
Robust hand pose estimation during the interaction with an unknown object
Chiho Choi, Sang Ho Yoon, Chin-Ning Chen, and Karthik Ramani · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
H+ o: Unified egocentric recognition of 3d hand-object poses and interactions
Bugra Tekin, Federica Bogo, and Marc Pollefeys · 2019
Later among the works it cites.
Xingyi Zhou, Dequan Wang, and Philipp Krähenbühl · 2019
Later among the works it cites.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Later among the works it cites.
Hope-net: A graph-based model for hand-object pose estimation
Bardia Doosti, Shujon Naha, Majid Mirbagheri, and David J Crandall · 2020
Later among the works it cites.
Leveraging photometric consistency over time for sparsely supervised hand-object reconstruction
Yana Hasson, Bugra Tekin, Federica Bogo, Ivan Laptev, Marc Pollefeys, and Cordelia Schmid · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes
Yu Xiang, Tanner Schmidt, Venkatraman Narayanan, and Dieter Fox · 2017
Cited alongside, same era.
Encoder-decoder with atrous separable convolution for semantic image segmentation
Liang-Chieh Chen, Yukun Zhu, George Papandreou, Florian Schroff, and Hartwig Adam · 2018
Cited alongside, same era.
Detecting and recognizing human-object interactions
Georgia Gkioxari, Ross Girshick, Piotr Dollár, and Kaiming He · 2018
Cited alongside, same era.
Centernet: Keypoint triplets for object detection
Kaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi, Qingming Huang, and Qi Tian · 2019
Cited alongside, same era.
Pay attention to them: deep reinforcement learning-based cascade object detection
Songtao Liu, Di Huang, and Yunhong Wang · 2019
Cited alongside, same era.
Pvnet: Pixel-wise voting network for 6dof pose estimation
Sida Peng, Yuan Liu, Qixing Huang, Xiaowei Zhou, and Hujun Bao · 2019
Cited alongside, same era.
Ppdm: Parallel point detection and matching for real-time human-object interaction detection
Yue Liao, Si Liu, Fei Wang, Yanjie Chen, Chen Qian, and Jiashi Feng · 2020
Later among the works it cites.
Understanding human hands in contact at internet scale
Dandan Shan, Jiaqi Geng, Michelle Shu, and David F Fouhey · 2020
Later among the works it cites.
Efficient object detection in large images using deep reinforcement learning
Burak Uzkent, Christopher Yeh, and Stefano Ermon · 2020
Later among the works it cites.
Joint hand-object 3d reconstruction from a single image with cross-branch feature fusion
Yujin Chen, Zhigang Tu, Di Kang, Ruizhi Chen, Linchao Bao, Zhengyou Zhang, and Junsong Yuan · 2021
Closest in time.
Dirv: Dense interaction region voting for end-to-end human-object interaction detection
Hao-Shu Fang, Yichen Xie, Dian Shao, and Cewu Lu · 2021
Closest in time.
Hotr: End-to-end human-object interaction detection with transformers
Bumsoo Kim, Junhyun Lee, Jaewoo Kang, Eun-Sol Kim, and Hyunwoo J Kim · 2021
Closest in time.
Semi-supervised 3d hand-object poses estimation with interactions in time
Shaowei Liu, Hanwen Jiang, Jiarui Xu, Sifei Liu, and Xiaolong Wang · 2021
Closest in time.
The meccano dataset: Understanding human-object interactions from egocentric videos in an industrial-like domain
Francesco Ragusa, Antonino Furnari, Salvatore Livatino, and Giovanni Maria Farinella · 2021
Closest in time.