Fetching the paper…
Reading the bibliography…
Current state-of-the-art object detection and segmentation methods work well under the closed-world assumption.
Normalized cuts and image segmentation
Jianbo Shi and J. Malik · 2000
Earlier work this paper cites.
A database of human segmented natural images and its application to evaluating segmentation algorithms and measuring ecological statistics
D. Martin, C. Fowlkes, D. Tal, and J. Malik · 2001
Earlier work this paper cites.
One-shot learning of object categories
Li Fei-Fei, R. Fergus, and P. Perona · 2006
Earlier work this paper cites.
Object retrieval with large vocabularies and fast spatial matching
J. Philbin, O. Chum, M. Isard, J. Sivic, and A. Zisserman · 2007
Earlier work this paper cites.
Learning spatial context: Using stuff to find things
Geremy Heitz and Daphne Koller · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
What is an object?
Bogdan Alexe, Thomas Deselaers, and Vittorio Ferrari · 2010
Earlier work this paper cites.
Efficient hierarchical graph-based video segmentation
M. Grundmann, V. Kwatra, M. Han, and I. Essa · 2010
Earlier work this paper cites.
Segmentation as selective search for object recognition
K. E. A. van de Sande, J. R. R. Uijlings, T. Gevers, and A. W. M. Smeulders · 2011
Earlier work this paper cites.
Category-independent object proposals with diverse ranking
I. Endres and D. Hoiem · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, M. Maire, Serge J. Belongie, James Hays, P. Perona, D. Ramanan, Piotr Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Edge boxes: Locating object proposals from edges
C. L. Zitnick and Piotr Dollár · 2014
Earlier work this paper cites.
Towards open world recognition
Abhijit Bendale and Terrance Boult · 2015
Earlier work this paper cites.
The pascal visual object classes challenge: A retrospective
M. Everingham, S. M. A. Eslami, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman · 2015
Earlier work this paper cites.
Learning to segment object candidates
Pedro O. Pinheiro, Ronan Collobert, and Piotr Dollár · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Learning spatiotemporal features with 3d convolutional networks
Du Tran, Lubomir Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri · 2015
Earlier work this paper cites.
Object instance search in videos via spatio-temporal trajectory discovery
J. Meng, J. Yuan, J. Yang, G. Wang, and Y. Tan · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
F. Perazzi, J. Pont-Tuset, B. McWilliams, L. Van Gool, M. Gross, and A. Sorkine-Hornung · 2016
Earlier work this paper cites.
You only look once: Unified, real-time object detection
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi · 2016
Cited alongside, same era.
Libsvx: A supervoxel library and benchmark for early video processing
Chenliang Xu and Jason Corso · 2016
Cited alongside, same era.
Quo vadis, action recognition? a new model and the kinetics dataset
J. Carreira and A. Zisserman · 2017
Cited alongside, same era.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Cited alongside, same era.
The kinetics human action video dataset
Will Kay, Joao Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, Mustafa Suleyman, and Andrew Zisserman · 2017
Cited alongside, same era.
The 2017 davis challenge on video object segmentation
Slowfast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Later among the works it cites.
Lvis: A dataset for large vocabulary instance segmentation
A. Gupta, P. Dollár, and R. Girshick · 2019
Later among the works it cites.
Large-scale long-tailed recognition in an open world
Ziwei Liu, Zhongqi Miao, Xiaohang Zhan, Jiayun Wang, Boqing Gong, and Stella X Yu · 2019
Later among the works it cites.
Video classification with channel-separated convolutional networks
Du Tran, Heng Wang, Lorenzo Torresani, and Matt Feiszli · 2019
Later among the works it cites.
Mots: Multi-object tracking and segmentation
Paul Voigtlaender, Michael Krause, Aljosa Osep, Jonathon Luiten, Berin Balachandar Gnana Sekar, Andreas Geiger, and Bastian Leibe · 2019
Later among the works it cites.
Detectron2
Yuxin Wu, Alexander Kirillov, Francisco Massa, Wan-Yen Lo, and Ross Girshick · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbelaez, Alex Sorkine-Hornung, and Luc Van Gool · 2017
Cited alongside, same era.
End-to-end video classification with knowledge graphs
Fang Yuan, Zhe Wang, Jie Lin, Luis Fernando D’Haro, Kim Jung Jae, Z. Zeng, and V. Chandrasekhar · 2017
Cited alongside, same era.
Deep free-form deformation network for object-mask registration
Haoyang Zhang and Xuming He · 2017
Cited alongside, same era.
Scene parsing through ade20k dataset
B. Zhou, H. Zhao, X. Puig, S. Fidler, A. Barriuso, and A. Torralba · 2017
Cited alongside, same era.
Unsupervised learning of depth and ego-motion from video
Tinghui Zhou, M. Brown, Noah Snavely, and D. Lowe · 2017
Cited alongside, same era.
Coco-stuff: Thing and stuff classes in context
Holger Caesar, Jasper Uijlings, and Vittorio Ferrari · 2018
Cited alongside, same era.
Scaling egocentric vision: The epic-kitchens dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, and Michael Wray · 2018
Cited alongside, same era.
Pixel objectness: Learning to segment generic objects automatically in images and videos
B. Xiong, S. Jain, and K. Grauman · 2019
Later among the works it cites.
Video instance segmentation
L. Yang, Y. Fan, and N. Xu · 2019
Later among the works it cites.
Tao: A large-scale benchmark for tracking any object
Achal Dave, Tarasha Khurana, Pavel Tokmakov, Cordelia Schmid, and Deva Ramanan · 2020
Later among the works it cites.
The overlooked elephant of object detection: Open set
A. R. Dhamija, M. Günther, J. Ventura, and T. E. Boult · 2020
Later among the works it cites.
All about knowledge graphs for actions, 2020
Pallabi Ghosh, Nirat Saini, Larry S. Davis, and Abhinav Shrivastava · 2020
Later among the works it cites.
Space-time memory networks for video object segmentation with user guidance
S. W. Oh, J. Y. Lee, N. Xu, and S. J. Kim · 2020
Later among the works it cites.
Don’t judge an object by its context: Learning to overcome contextual bias
Krishna Kumar Singh, Dhruv Mahajan, Kristen Grauman, Yong Jae Lee, Matt Feiszli, and Deepti Ghadiyaram · 2020
Later among the works it cites.
https://modelcards.withgoogle.com/object-detection
Google cloud model cards object detection · 2021
Closest in time.
pytorchvideo
Facebook AI · 2021
Closest in time.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Closest in time.
Class-agnostic object detection
Ayush Jaiswal, Yue Wu, Pradeep Natarajan, and Premkumar Natarajan · 2021
Closest in time.
Supervoxel attention graphs for long-range video modeling
Yang Wang, Gedas Bertasius, Tae-Hyun Oh, Abhinav Gupta, Minh Hoai, and Lorenzo Torresani · 2021
Closest in time.