Fetching the paper…
Reading the bibliography…
Human-Object Interaction (HOI) Detection is an important problem to understand how humans interact with objects.
Unsupervised discovery of action classes
Y. Wang, H. Jiang, Mark. S. Drew, Z.-N. Li, and G. Mori · 2006
Earlier work this paper cites.
Recognizing actions from still images
N. Ikizler, R. G. Cinbis, S. Pehlivan, and P. Duygulu · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Recognizing human actions in still images: a study of bag-of-features and part-based representations
V. Delaitre, I. Laptev, and J. Sivic · 2010
Earlier work this paper cites.
Recognizing human actions from still images with latent poses
W. Yang, Y. Wang, and G. Mori · 2010
Earlier work this paper cites.
Recognition using visual phrases
M. A. Sadeghi and A. Farhadi · 2012
Earlier work this paper cites.
Predicting the location of “interactees” in novel human-object interactions
C.-Y. Chen and K. Grauman · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Activitynet: A large-scale video benchmark for human activity understanding
F. Caba Heilbron, V. Escorcia, B. Ghanem, and J. Carlos Niebles · 2015
Earlier work this paper cites.
Hico: A benchmark for recognizing human-object interactions in images
Y. W. Chao, Z. Wang, Y. He, J. Wang, and J. Deng · 2015
Earlier work this paper cites.
Fast r-cnn
R. Girshick · 2015
Earlier work this paper cites.
S. Gupta and J. Malik · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Visual relationship detection with language priors
C. Lu, R. Krishna, M. Bernstein, and L. Fei-Fei · 2016
Cited alongside, same era.
Learning models for actions and person-object interactions with transfer to question answering
A. Mallya and S. Lazebnik · 2016
Cited alongside, same era.
Situation recognition: Visual semantic role labeling for image understanding
M. Yatskar, L. Zettlemoyer, and A. Farhadi · 2016
Cited alongside, same era.
RMPE: Regional multi-person pose estimation
H.-S. Fang, S. Xie, Y.-W. Tai, and C. Lu · 2017
Cited alongside, same era.
Detecting and recognizing human-object interactions
G. Gkioxari, R. Girshick, P. Dollár, and K. He · 2017
Cited alongside, same era.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
ican: Instance-centric attention network for human-object interaction detection
C. Gao, Y. Zou, and J.-B. Huang · 2018
Closest in time.
Detectron
R. Girshick, I. Radosavovic, G. Gkioxari, P. Dollár, and K. He · 2018
Closest in time.
Learning human-object interactions by graph parsing neural networks
S. Qi, W. Wang, B. Jia, J. Shen, and S.-C. Zhu · 2018
Closest in time.
Scaling human-object interaction recognition through zero-shot learning
L. Shen, S. Yeung, J. Hoffman, G. Mori, and L. Fei Fei · 2018
Closest in time.
Zoom-net: Mining deep feature interactions for visual relationship recognition
G. Yin, L. Sheng, B. Liu, N. Yu, X. Wang, J. Shao, and C. C. Loy · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, et al · 2017
Cited alongside, same era.
Feature pyramid networks for object detection
T.-Y. Lin, P. Dollár, R. B. Girshick, K. He, B. Hariharan, and S. J. Belongie · 2017
Cited alongside, same era.
Scene graph generation by iterative message passing
D. Xu, Y. Zhu, C. B. Choy, and L. Fei-Fei · 2017
Cited alongside, same era.
Visual translation embedding network for visual relation detection
H. Zhang, Z. Kyaw, S.-F. Chang, and T.-S. Chua · 2017
Cited alongside, same era.
Care about you: towards large-scale human-centric visual relationship detection
B. Zhuang, Q. Wu, C. Shen, I. Reid, and A. v. d. Hengel · 2017
Cited alongside, same era.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár and R. Girshick · 2017
Cited alongside, same era.
Learning to detect human-object interactions
Y. W. Chao, Y. Liu, X. Liu, H. Zeng, and J. Deng · 2018
Cited alongside, same era.
H. Fang, Y. Xu, W. Wang, X. Liu, and S.-C. Zhu · 2018
Closest in time.
SRDA: Generating Instance Segmentation Annotation via Scanning, Reasoning and Domain Adaptation
W. Xu, Y. Li and C. Lu · 2018
Closest in time.
Pose Flow: Efficient Online Pose Tracking
Y. Xiu, J. Li, H. Wang, Y. Fang and C. Lu · 2018
Closest in time.
Weakly and semi supervised human body part parsing via pose-guided knowledge transfer
H. S. Fang, G. Lu, X. Fang, J. Xie, Y. W. Tai and C. Lu · 2018
Closest in time.
Beyond holistic object recognition: Enriching image understanding with part states
C. Lu, H. Su, Y. L. Li, Y. Lu, L. Yi, C. K. Tang and L. J. Leonidas · 2018
Closest in time.
Deep RNN Framework for Visual Sequential Applications
B. Pang, K. Zha, H. Cao, S. Chen and C. Lu · 2018
Closest in time.
CrowdPose: Efficient Crowded Scenes Pose Estimation and A New Benchmark
J. Li, C. Wang, H. Zhu, Y. Mao, H. S. Fang and C. Lu · 2018
Closest in time.
Graph r-cnn for scene graph generation
L. Yang, L. Lu, S. Lee. D. Batra and D. Parikh · 2018
Closest in time.