Spatial pyramid pooling in deep convolutional networks for visual recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2014
Cited alongside, same era.
TUHOI: trento universal human object interaction dataset
D. Le, R. Bernardi, and J. Uijlings · 2014
Cited alongside, same era.
Microsoft COCO: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Cited alongside, same era.
HICO: A benchmark for recognizing human-object interactions in images
Y.-W. Chao, Z. Wang, Y. He, J. Wang, and J. Deng · 2015
Cited alongside, same era.
Fast R-CNN
R. Girshick · 2015
Cited alongside, same era.
Contextual action recognition with R*CNN
G. Gkioxari, R. Girshick, and J. Malik · 2015
Cited alongside, same era.
Visual semantic role labeling
S. Gupta and J. Malik · 2015
Cited alongside, same era.
Learning semantic relationships for better action retrieval in images
V. Ramanathan, C. Li, J. Deng, W. Han, Z. Li, K. Gu, Y. Song, S. Bengio, C. Rossenberg, and L. Fei-Fei · 2015
Cited alongside, same era.
Faster R-CNN: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
Describing common human visual actions in images
M. R. Ronchi and P. Perona · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.