Fetching the paper…
Reading the bibliography…
Parsing urban scene images benefits many applications, especially self-driving.
Geometric context from a single image
D. Hoiem, A. A. Efros, and M. Hebert · 2005
Earlier work this paper cites.
Semantic object classes in video: A high-definition ground truth database
G. J. Brostow, J. Fauqueur, and R. Cipolla · 2008
Earlier work this paper cites.
Putting objects in perspective
D. Hoiem, A. A. Efros, and M. Hebert · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Combining appearance and structure from motion features for road scene understanding
P. Sturgess, K. Alahari, L. Ladicky, and P. H. Torr · 2009
Earlier work this paper cites.
What, where and how many? combining object detectors and crfs
L. Ladický, P. Sturgess, K. Alahari, C. Russell, and P. H. S. Torr · 2010
Earlier work this paper cites.
Efficient inference in fully connected crfs with gaussian edge potentials
P. Krähenbühl and V. Koltun · 2011
Earlier work this paper cites.
Efficient inference for fully-connected crfs with stationarity
Y. Zhang and T. Chen · 2012
Earlier work this paper cites.
Learning hierarchical features for scene labeling
C. Farabet, C. Couprie, L. Najman, and Y. LeCun · 2013
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
D. Eigen, C. Puhrsch, and R. Fergus · 2014
Earlier work this paper cites.
Caffe: Convolutional architecture for fast feature embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Earlier work this paper cites.
Pulling things out of perspective
L. Ladicky, J. Shi, and M. Pollefeys · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Recurrent models of visual attention
V. Mnih, N. Heess, A. Graves, et al · 2014
Cited alongside, same era.
Neural decision forests for semantic image labelling
S. Rota Bulo and P. Kontschieder · 2014
Cited alongside, same era.
Segnet: A deep convolutional encoder-decoder architecture for robust semantic pixel-wise labelling
V. Badrinarayanan, A. Handa, and R. Cipolla · 2015
Cited alongside, same era.
Boxsup: Exploiting bounding boxes to supervise convolutional networks for semantic segmentation
J. Dai, K. He, and J. Sun · 2015
Cited alongside, same era.
Fast r-cnn
R. Girshick · 2015
Cited alongside, same era.
Semantic image segmentation with deep convolutional nets and fully connected crfs
C. Liang-Chieh, G. Papandreou, I. Kokkinos, K. Murphy, and A. Yuille · 2015
Conditional random fields as recurrent neural networks
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. H. Torr · 2015
Later among the works it cites.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille · 2016
Later among the works it cites.
Attention to scale: Scale-aware semantic image segmentation
L.-C. Chen, Y. Yang, J. Wang, W. Xu, and A. L. Yuille · 2016
Later among the works it cites.
The cityscapes dataset for semantic urban scene understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Later among the works it cites.
Unsupervised cnn for single view depth estimation: Geometry to the rescue
R. Garg, G. Carneiro, and I. Reid · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Efficient piecewise training of deep structured models for semantic segmentation
G. Lin, C. Shen, I. Reid, et al · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Feedforward semantic segmentation with zoom-out features
M. Mostajabi, P. Yadollahpour, and G. Shakhnarovich · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
Action recognition using visual attention
S. Sharma, R. Kiros, and R. Salakhutdinov · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, K. Cho, A. C. Courville, R. Salakhutdinov, R. S. Zemel, and Y. Bengio · 2015
Cited alongside, same era.
Laplacian pyramid reconstruction and refinement for semantic segmentation
G. Ghiasi and C. C. Fowlkes · 2016
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
Deepgender: Occlusion and low resolution robust facial gender classification via progressively trained convolutional neural networks with attention
F. Juefei-Xu, E. Verma, P. Goel, A. Cherodian, and M. Savvides · 2016
Later among the works it cites.
Exploring context with deep structured models for semantic segmentation
G. Lin, C. Shen, A. v. d. Hengel, and I. Reid · 2016
Later among the works it cites.
Learning to refine object segments
P. O. Pinheiro, T.-Y. Lin, R. Collobert, and P. Dollár · 2016
Later among the works it cites.
Dag-recurrent neural networks for scene labeling
B. Shuai, Z. Zuo, B. Wang, and G. Wang · 2016
Later among the works it cites.
Zoom better to see clearer: Human and object parsing with hierarchical auto-zoom net
F. Xia, P. Wang, L.-C. Chen, and A. L. Yuille · 2016
Later among the works it cites.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2016
Later among the works it cites.