Fetching the paper…
Reading the bibliography…
We present Open Images V4, a dataset of 9.2M images with unified annotations for image classification, object detection and visual relationship detection.
Qian N (1999) On the momentum term in gradient descent learning algorithms. Neural Networks 12(1):145–151
1999
Earlier work this paper cites.
Fei-Fei L, Fergus R, Perona P (2006) One-shot learning of object categories. IEEE Trans on PAMI 28(4):594–611
2006
Earlier work this paper cites.
Griffin G, Holub A, Perona P (2007) The Caltech-256. Tech. rep., Caltech
2007
Earlier work this paper cites.
Deng J, Dong W, Socher R, Li LJ, Li K, Fei-fei L (2009) ImageNet: A large-scale hierarchical image database. In: CVPR
2009
Earlier work this paper cites.
Gupta A, Kembhavi A, Davis L (2009) Observing human-object interactions: Using spatial and functional compatibility for recognition. In: IEEE Trans. on PAMI
2009
Earlier work this paper cites.
Krizhevsky A (2009) Learning multiple layers of features from tiny images. Tech. rep., University of Toronto
2009
Earlier work this paper cites.
Alexe B, Deselaers T, Ferrari V (2010) What is an object? In: CVPR
2010
Earlier work this paper cites.
Everingham M, Van Gool L, Williams CKI, Winn J, Zisserman A (2010) The PASCAL Visual Object Classes (VOC) Challenge. IJCV
2010
Earlier work this paper cites.
Yao B, Fei-Fei L (2010) Modeling mutual context of object and human pose in human-object interaction activities. In: CVPR
2010
Earlier work this paper cites.
Alexe B, Deselaers T, Ferrari V (2012) Measuring the objectness of image windows. IEEE Trans on PAMI
2012
Earlier work this paper cites.
Everingham M, Van Gool L, Williams CKI, Winn J, Zisserman A (2012) The PASCAL Visual Object Classes Challenge 2012 (VOC2012) Results. http://www.pascal-network.org/challenges/VOC/voc2012/workshop/index.html
2012
Earlier work this paper cites.
Krizhevsky A, Sutskever I, Hinton GE (2012) Imagenet classification with deep convolutional neural networks. In: NeurIPS
2012
Earlier work this paper cites.
Prest A, Schmid C, Ferrari V (2012) Weakly supervised learning of interactions between humans and objects. IEEE Trans on PAMI
2012
Earlier work this paper cites.
Su H, Deng J, Fei-Fei L (2012) Crowdsourcing annotations for visual object detection. In: AAAI Human Computation Workshop
2012
Earlier work this paper cites.
Mikolov T, Sutskever I, Chen K, Corrado GS, Dean J (2013) Distributed representations of words and phrases and their compositionality. In: NeurIPS
2013
Earlier work this paper cites.
Uijlings JRR, van de Sande KEA, Gevers T, Smeulders AWM (2013) Selective search for object recognition. IJCV
2013
Earlier work this paper cites.
Girshick R, Donahue J, Darrell T, Malik J (2014) Rich feature hierarchies for accurate object detection and semantic segmentation. In: CVPR
2014
Earlier work this paper cites.
Hinton GE, Vinyals O, Dean J (2014) Distilling the knowledge in a neural network. In: NeurIPS
2014
Earlier work this paper cites.
Lin TY, Maire M, Belongie S, Bourdev L, Girshick R, Hays J, Perona P, Ramanan D, Zitnick CL, Dollár P (2014) Microsoft COCO: Common objects in context. In: ECCV
2014
Cited alongside, same era.
Everingham M, Eslami S, van Gool L, Williams C, Winn J, Zisserman A (2015) The PASCAL visual object classes challenge: A retrospective. IJCV
2015
Cited alongside, same era.
Girshick R (2015) Fast R-CNN. In: ICCV
2015
Cited alongside, same era.
Gupta S, Malik J (2015) Visual semantic role labeling. arXiv preprint arXiv:150504474
2015
Cited alongside, same era.
Ioffe S, Szegedy C (2015) Batch normalization: Accelerating deep network training by reducing internal covariate shift. In: ICML
2015
Cited alongside, same era.
Li Y, Ouyang W, Wang X, Tang X (2017) ViP-CNN: Visual phrase guided convolutional neural network. In: CVPR
2017
Later among the works it cites.
Liang X, Lee L, Xing EP (2017) Deep variation-structured reinforcement learning for visual relationship and attribute detection. In: CVPR
2017
Later among the works it cites.
Lin T, Goyal P, Girshick R, He K, Dollar P (2017) Focal loss for dense object detection. In: ICCV
2017
Later among the works it cites.
Papadopoulos DP, Uijlings JR, Keller F, Ferrari V (2017) Extreme clicking for efficient object annotation. In: ICCV
2017
Later among the works it cites.
Peyre J, Laptev I, Schmid C, Sivic J (2017) Weakly-supervised learning of visual relations. In: CVPR
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ren S, He K, Girshick R, Sun J (2015) Faster R-CNN: Towards real-time object detection with region proposal networks. In: NeurIPS
2015
Cited alongside, same era.
Russakovsky O, Deng J, Su H, Krause J, Satheesh S, Ma S, Huang Z, Karpathy A, Khosla A, Bernstein M, Berg A, Fei-Fei L (2015) ImageNet large scale visual recognition challenge. IJCV
2015
Cited alongside, same era.
Szegedy C, Liu W, Jia Y, Sermanet P, Reed S, Anguelov D, Erhan D, Vanhoucke V, Rabinovich A (2015) Going deeper with convolutions. In: CVPR
2015
Cited alongside, same era.
He K, Zhang X, Ren S, Sun J (2016) Deep residual learning for image recognition. In: CVPR
2016
Cited alongside, same era.
Liu W, Anguelov D, Erhan D, Szegedy C, Reed S, Fu CY, Berg AC (2016) SSD: Single shot multibox detector. In: ECCV
2016
Cited alongside, same era.
Lu C, Krishna R, Bernstein M, Fei-Fei L (2016) Visual relationship detection with language priors. In: European Conference on Computer Vision
2016
Cited alongside, same era.
Papadopoulos DP, Uijlings JRR, Keller F, Ferrari V (2016) We don’t need no bounding-boxes: Training object class detectors using only human verification. In: CVPR
2016
Cited alongside, same era.
2017
Later among the works it cites.
Sun C, Shrivastava A, Singh S, Gupta A (2017) Revisiting unreasonable effectiveness of data in deep learning era. In: ICCV
2017
Later among the works it cites.
Szegedy C, Ioffe S, Vanhoucke V, Alemi A (2017) Inception-v4, inception-resnet and the impact of residual connections on learning. In: AAAI
2017
Later among the works it cites.
Veit A, Alldrin N, Chechik G, Krasin I, Gupta A, Belongie S (2017) Learning from noisy large-scale datasets with minimal supervision. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 839–847, URL http://openaccess.thecvf.com/content_cvpr_2017/papers/Veit_Learning_From_Noisy_CVPR_2017_paper.pdf
2017
Later among the works it cites.
Xu D, Zhu Y, Choy C, Fei-Fei L (2017) Scene graph generation by iterative message passing. In: Computer Vision and Pattern Recognition (CVPR)
2017
Later among the works it cites.
Gao C, Zou Y, Huang JB (2018) iCAN: Instance-centric attention network for human-object interaction detection. In: BMVC
2018
Closest in time.
Gkioxari G, Girshick R, Dollár P, He K (2018) Detecting and recognizing human-object interactions. CVPR
2018
Closest in time.
Kolesnikov A, Kuznetsova A, Lampert C, Ferrari V (2018) Detecting visual relationships using box attention. arXiv 1807.02136
2018
Closest in time.
Liang K, Guo Y, Chang H, Chen X (2018) Visual relationship detection with deep structural ranking. In: AAAI
2018
Closest in time.
Sandler M, Howard AG, Zhu M, Zhmoginov A, Chen L (2018) Mobilenetv2: Inverted residuals and linear bottleneck. In: CVPR
2018
Closest in time.
Uijlings J, Popov S, Ferrari V (2018) Revisiting knowledge transfer for training object class detectors. In: CVPR
2018
Closest in time.
Zellers R, Yatskar M, Thomson S, Choi Y (2018) Neural motifs: Scene graph parsing with global context. In: CVPR
2018
Closest in time.