Fetching the paper…
Reading the bibliography…
Humans recognize the visual world at multiple levels: we effortlessly categorize scenes and detect objects inside, while also identifying the textures and surfaces of the objects along with their different compositional parts.
Batch renormalization: Towards reducing minibatch dependence in batch-normalized models
Ioffe, S.: · 1950
Earlier work this paper cites.
Integrated segmentation and recognition of hand-printed numerals
Keeler, J.D., Rumelhart, D.E., Leow, W.K.: · 1991
Earlier work this paper cites.
An expectation maximization approach to the synergy between image segmentation and object categorization
Kokkinos, I., Maragos, P.: · 2005
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: · 2009
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
Everingham, M., Van Gool, L., Williams, C.K., Winn, J., Zisserman, A.: · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, V., Hinton, G.E.: · 2010
Earlier work this paper cites.
Object detection and segmentation from joint embedding of parts and pixels
Maire, M., Stella, X.Y., Perona, P.: · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., Hinton, G.E.: · 2012
Earlier work this paper cites.
Opensurfaces: A richly annotated catalog of surface appearance
Bell, S., Upchurch, P., Snavely, N., Bala, K.: · 2013
Earlier work this paper cites.
What is network science?
Brandes, U., Robins, G., McCranie, A., Wasserman, S.: · 2013
Earlier work this paper cites.
Describing textures in the wild
Cimpoi, M., Maji, S., Kokkinos, I., Mohamed, S., Vedaldi, A.: · 2014
Earlier work this paper cites.
Semantic image segmentation with deep convolutional nets and fully connected crfs
Chen, L.C., Papandreou, G., Kokkinos, I., Murphy, K., Yuille, A.L.: · 2014
Earlier work this paper cites.
The role of context for object detection and semantic segmentation in the wild
Mottaghi, R., Chen, X., Liu, X., Cho, N.G., Lee, S.W., Fidler, S., Urtasun, R., Yuille, A.: · 2014
Earlier work this paper cites.
Detect what you can: Detecting and representing objects using holistic models and body parts
Chen, X., Mottaghi, R., Liu, X., Fidler, S., Urtasun, R., Yuille, A.: · 2014
Earlier work this paper cites.
Learning deep features for scene recognition using places database
Zhou, B., Lapedriza, A., Xiao, J., Torralba, A., Oliva, A.: · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Zeiler, M.D., Fergus, R.: · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K., Zisserman, A.: · 2015
Cited alongside, same era.
Going deeper with convolutions, Cvpr (2015)
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., Rabinovich, A., et al.: · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
Long, J., Shelhamer, E., Darrell, T.: · 2015
Cited alongside, same era.
Learning deconvolution network for semantic segmentation
Noh, H., Hong, S., Han, B.: · 2015
Cited alongside, same era.
Convolutional models for joint object categorization and pose estimation
Elhoseiny, M., El-Gaaly, T., Bakry, A., Elgammal, A.: · 2015
Chen, L.C., Papandreou, G., Kokkinos, I., Murphy, K., Yuille, A.L.: · 2016
Later among the works it cites.
Scene parsing through ade20k dataset
Zhou, B., Zhao, H., Puig, X., Fidler, S., Barriuso, A., Torralba, A.: · 2017
Later among the works it cites.
Learning to segment every thing
Hu, R., Dollár, P., He, K., Darrell, T., Girshick, R.: · 2017
Later among the works it cites.
Refinenet: Multi-path refinement networks for high-resolution semantic segmentation
Lin, G., Milan, A., Shen, C., Reid, I.: · 2017
Later among the works it cites.
Pyramid scene parsing network
Zhao, H., Shi, J., Qi, X., Wang, X., Jia, J.: · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture
Eigen, D., Fergus, R.: · 2015
Cited alongside, same era.
Object detectors emerge in deep scene cnns
Zhou, B., Khosla, A., Lapedriza, A., Oliva, A., Torralba, A.: · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S., Szegedy, C.: · 2015
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., Sun, J.: · 2016
Cited alongside, same era.
Multi-scale context aggregation by dilated convolutions
Yu, F., Koltun, V.: · 2016
Cited alongside, same era.
The cityscapes dataset for semantic urban scene understanding
Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., Schiele, B.: · 2016
Cited alongside, same era.
Ubernet: Training a universal convolutional neural network for low-, mid-, and high-level vision using diverse datasets and limited memory
Kokkinos, I.: · 2017
Later among the works it cites.
Network dissection: Quantifying interpretability of deep visual representations
Bau, D., Zhou, B., Khosla, A., Oliva, A., Torralba, A.: · 2017
Later among the works it cites.
Feature pyramid networks for object detection
Lin, T.Y., Dollár, P., Girshick, R., He, K., Hariharan, B., Belongie, S.: · 2017
Later among the works it cites.
Megdet: A large mini-batch object detector
Peng, C., Xiao, T., Li, Z., Jiang, Y., Zhang, X., Jia, K., Yu, G., Sun, J.: · 2017
Later among the works it cites.
Aggregated residual transformations for deep neural networks
Xie, S., Girshick, R., Dollár, P., Tu, Z., He, K.: · 2017
Later among the works it cites.
Segnet: A deep convolutional encoder-decoder architecture for image segmentation
Badrinarayanan, V., Kendall, A., Cipolla, R.: · 2017
Later among the works it cites.
Mscoco challenge 2017: stuff segmentation, team fair
Kirillov, A., He, K., Girshick, R., Dollár, P.: · 2017
Later among the works it cites.
Interpreting deep visual representations via network dissection
Zhou, B., Bau, D., Oliva, A., Torralba, A.: · 2018
Closest in time.
Adaptive deconvolutional networks for mid and high level feature learning
Zeiler, M.D., Taylor, G.W., Fergus, R.: · 2025
Closest in time.