Fetching the paper…
Reading the bibliography…
Visual understanding of complex urban street scenes is an enabling factor for a wide range of applications.
Modeling the shape of the scene: A holistic representation of the spatial envelope
A. Oliva and A. Torralba · 2001
Earlier work this paper cites.
In-factory calibration of multiocular camera systems
L. Krüger, C. Wöhler, A. Würz-Wessel, and F. Stein · 2004
Earlier work this paper cites.
Dynamic 3D scene analysis from a moving vehicle
B. Leibe, N. Cornelis, K. Cornelis, and L. Van Gool · 2007
Earlier work this paper cites.
Robust object detection with interleaved categorization and segmentation
B. Leibe, A. Leonardis, and B. Schiele · 2008
Earlier work this paper cites.
LabelMe: A database and web-based tool for image annotation
B. C. Russell, A. Torralba, K. P. Murphy, and W. T. Freeman · 2008
Earlier work this paper cites.
Semantic object classes in video: A high-definition ground truth database
G. J. Brostow, J. Fauqueur, and R. Cipolla · 2009
Earlier work this paper cites.
Monocular pedestrian detection: Survey and experiments
M. Enzweiler and D. M. Gavrila · 2009
Earlier work this paper cites.
Object detection with discriminatively trained part-based models
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan · 2010
Earlier work this paper cites.
Object detection and segmentation from joint embedding of parts and pixels
M. Maire, S. X. Yu, and P. Perona · 2011
Earlier work this paper cites.
Pedestrian detection at 100 frames per second
R. Benenson, M. Mathias, R. Timofte, and L. Van Gool · 2012
Earlier work this paper cites.
Pedestrian detection: An evaluation of the state of the art
P. Dollár, C. Wojek, B. Schiele, and P. Perona · 2012
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Hough regions for joining instance localization and segmentation
H. Riemenschneider, S. Sternig, M. Donoser, P. M. Roth, and H. Bischof · 2012
Earlier work this paper cites.
Automatic dense visual semantic mapping from street-level imagery
S. Sengupta, P. Sturgess, L. Ladicky, and P. H. S. Torr · 2012
Earlier work this paper cites.
Describing the scene as a whole: Joint object detection, scene classification and semantic segmentation
J. Yao, S. Fidler, and R. Urtasun · 2012
Earlier work this paper cites.
Making Bertha see
U. Franke, D. Pfeiffer, C. Rabe, C. Knöppel, M. Enzweiler, F. Stein, and R. G. Herrtwich · 2013
Earlier work this paper cites.
Toward automated driving in cities using close-to-market sensors: An overview of the V-Charge project
P. Furgale, U. Schwesinger, M. Rufli, W. Derendarz, H. Grimmett, P. Mühlfellner, S. Wonneberger, B. Li, et al · 2013
Earlier work this paper cites.
Vision meets robotics: The KITTI dataset
A. Geiger, P. Lenz, C. Stiller, and R. Urtasun · 2013
Earlier work this paper cites.
Nonparametric semantic segmentation for 3D street scenes
H. Hu and B. Upcroft · 2013
Earlier work this paper cites.
Exploiting the power of stereo confidences
D. Pfeiffer, S. K. Gehrig, and N. Schneider · 2013
Earlier work this paper cites.
Efficient multi-cue scene segmentation
T. Scharwächter, M. Enzweiler, U. Franke, and S. Roth · 2013
Earlier work this paper cites.
Semantic modelling of urban scenes
S. Sengupta, E. Greveson, A. Shahrokni, and P. H. S. Torr · 2013
Earlier work this paper cites.
Superparsing
J. Tighe and S. Lazebnik · 2013
Earlier work this paper cites.
Selective search for object recognition
J. R. R. Uijlings, K. E. A. van de Sande, T. Gevers, and A. W. M. Smeulders · 2013
Earlier work this paper cites.
Information fusion on oversegmented images: An application for urban scene understanding
P. Xu, F. Davoine, J.-B. Bordes, H. Zhao, and T. Denoeux · 2013
Earlier work this paper cites.
Multiscale combinatorial grouping
P. Arbelaez, J. Pont-Tuset, J. Barron, F. Marqués, and J. Malik · 2014
Earlier work this paper cites.
3D traffic scene understanding from movable platforms
A. Geiger, M. Lauer, C. Wojek, C. Stiller, and R. Urtasun · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
Simultaneous detection and segmentation
B. Hariharan, P. Arbeláez, R. B. Girshick, and J. Malik · 2014
Cited alongside, same era.
An exemplar-based CRF for multi-instance object segmentation
X. He and S. Gould · 2014
Cited alongside, same era.
Joint semantic segmentation and 3D reconstruction from monocular video
A. Kundu, Y. Li, F. Dellaert, F. Li, and J. Rehg · 2014
Cited alongside, same era.
Pulling things out of perspective
L. Ladicky, J. Shi, and M. Pollefeys · 2014
Cited alongside, same era.
Microsoft COCO: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Cited alongside, same era.
The role of context for object detection and semantic segmentation in the wild
R. Mottaghi, X. Chen, X. Liu, N.-G. Cho, S.-W. Lee, S. Fidler, R. Urtasun, and A. Yuille · 2014
Cited alongside, same era.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Later among the works it cites.
Parsenet: Looking wider to see better
W. Liu, A. Rabinovich, and A. C. Berg · 2015
Later among the works it cites.
Semantic image segmentation via deep parsing network
Z. Liu, X. Li, P. Luo, C. C. Loy, and X. Tang · 2015
Later among the works it cites.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Later among the works it cites.
Watch and learn: Semi-supervised learning for object detectors from video
I. Misra, A. Shrivastava, and M. Hebert · 2015
Later among the works it cites.
Feedforward semantic segmentation with zoom-out features
M. Mostajabi, P. Yadollahpour, and G. Shakhnarovich · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Recurrent convolutional neural networks for scene parsing
P. H. Pinheiro and R. Collobert · 2014
Cited alongside, same era.
Learning where to classify in multi-view semantic segmentation
H. Riemenschneider, A. Bódis-Szomorú, J. Weissenberg, and L. Van Gool · 2014
Cited alongside, same era.
Stixmantics: A medium-level model for real-time semantic scene understanding
T. Scharwächter, M. Enzweiler, U. Franke, and S. Roth · 2014
Cited alongside, same era.
OverFeat: Integrated recognition, localization and detection using convolutional networks
P. Sermanet, D. Eigen, X. Zhang, M. Mathieu, R. Fergus, and Y. LeCun · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Learning deep features for scene recognition using places database
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva · 2014
Cited alongside, same era.
Is object localization for free? Weakly-supervised learning with convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2015
Later among the works it cites.
Weakly- and semi-supervised learning of a DCNN for semantic image segmentation
G. Papandreou, L.-C. Chen, K. Murphy, and A. L. Yuille · 2015
Later among the works it cites.
Constrained convolutional neural networks for weakly supervised segmentation
D. Pathak, P. Kraehenbuehl, and T. Darrell · 2015
Later among the works it cites.
Fully convolutional multi-class multiple instance learning
D. Pathak, E. Shelhamer, J. Long, and T. Darrell · 2015
Later among the works it cites.
From image-level to pixel-level labeling with convolutional networks
P. H. Pinheiro and R. Collobert · 2015
Later among the works it cites.
Boosting object proposals: From Pascal to COCO
J. Pont-Tuset and L. Van Gool · 2015
Later among the works it cites.
Faster R-CNN: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Later among the works it cites.
Vision-based offline-online perception paradigm for autonomous driving
G. Ros, S. Ramos, M. Granados, D. Vazquez, and A. M. Lopez · 2015
Later among the works it cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Later among the works it cites.
Fully connected deep structured networks
A. Schwing and R. Urtasun · 2015
Later among the works it cites.
Deep hierarchical parsing for semantic segmentation
A. Sharma, O. Tuzel, and D. W. Jacobs · 2015
Later among the works it cites.
Sun RGB-D: A RGB-D scene understanding benchmark suite
S. Song, S. P. Lichtenberg, and J. Xiao · 2015
Later among the works it cites.
Scene parsing with object instance inference using regions and per-exemplar detectors
J. Tighe, M. Niethammer, and S. Lazebnik · 2015
Later among the works it cites.
Incremental dense semantic stereo fusion for large-scale semantic scene reconstruction
V. Vineet, O. Miksik, M. Lidegaard, M. Niessner, S. Golodetz, V. A. Prisacariu, O. Kahler, D. W. Murray, S. Izadi, P. Perez, and P. H. S. Torr · 2015
Later among the works it cites.
STC: A simple to complex framework for weakly-supervised semantic segmentation
Y. Wei, X. Liang, Y. Chen, X. Shen, M.-M. Cheng, Y. Zhao, and S. Yan · 2015
Later among the works it cites.
Learning to segment under various forms of weak supervision
J. Xu, A. G. Schwing, and R. Urtasun · 2015
Later among the works it cites.
Monocular object instance segmentation and depth ordering with CNNs
Z. Zhang, A. Schwing, S. Fidler, and R. Urtasun · 2015
Later among the works it cites.
Conditional random fields as recurrent neural networks
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. H. S. Torr · 2015
Later among the works it cites.
Efficient piecewise training of deep structured models for semantic segmentation
G. Lin, C. Shen, A. van den Hengel, and I. Reid · 2016
Closest in time.
Semantic instance annotation of street scenes by 3D to 2D label transfer
J. Xie, M. Kiefel, M.-T. Sun, and A. Geiger · 2016
Closest in time.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2016
Closest in time.