Fetching the paper…
Reading the bibliography…
Despite enormous progress in object detection and classification, the problem of incorporating expected contextual relationships among object instances into modern recognition systems remains a key challenge.
Metropolis N, Rosenbluth AW, Rosenbluth MN, Teller AH, Teller E (1953) “Equation of State Calculations by Fast Computing Machines”. Journal of Chemical Physics 21:1087–1092
1953
Earlier work this paper cites.
Hastings WK (1970) “Monte Carlo sampling methods using Markov chains and their applications”. Biometrika 57(1):97–109
1970
Earlier work this paper cites.
Mode CJ (1971) “Multitype branching processes; theory and applications”. American Elsevier Pub. Co New York
1971
Earlier work this paper cites.
Dempster AP, Laird NM, Rubin DB (1977) “Maximum likelihood from incomplete data via the EM algorithm”. J of the Royal Stat Soc Series B (Methodological) pp 1–38
1977
Earlier work this paper cites.
Celeux G, Diebolt J (1985) “The SEM Algorithm: A probabilistic teacher algorithm derived from the EM algorithm for the mixture problem”. Computational Statistics Quarterly 2:73–82
1985
Earlier work this paper cites.
Homma T, Atlas LE, Marks II RJ (1988) “An Artificial Neural Network for Spatio-Temporal Bipolar Patterns: Application to Phoneme Classification”. In: Neural Information Processing Systems, pp 31–40
1988
Earlier work this paper cites.
1990
Earlier work this paper cites.
Geman D, Jedynak B (1996) “An active testing model for tracking roads in satellite images”. IEEE Transactions on Pattern Analysis and Machine Intelligence 18(1):1–14
1996
Earlier work this paper cites.
Lecun Y, Bottou L, Bengio Y, Haffner P (1998) “Gradient-based learning applied to document recognition”. Proceedings of the IEEE 86(11):2278–2324
1998
Earlier work this paper cites.
Reynolds JH, Chelazzi L, Desimone R (1999) “Competitive mechanisms subserve attention in macaque areas V2 and V4”. Journal of Neuroscience 19:1736–1753
1999
Earlier work this paper cites.
Roberts GO, Rosenthal JS (2001) Optimal scaling for various Metropolis-Hastings algorithms. Statist Sci 16(4):351–367
2001
Earlier work this paper cites.
Geman S, Potter DF, Chi Z (2002) “Composition systems”. Quarterly of Applied Mathematics pp 707–736
2002
Earlier work this paper cites.
Ma Y, Soatto S, Kosecka J, Sastry S (2003) An Invitation to 3D Vision: From Images to Geometric Models. Springer Verlag
2003
Earlier work this paper cites.
Hartley R, Zisserman A (2004) “Multiple View Geometry in Computer Vision”, 2nd edn. Cambridge
2004
Earlier work this paper cites.
Serences JT, Yantis S (2006) “Selective visual attention and perceptual coherence”. Trends in Cognitive Sciences 10(1):38–45, DOI http://dx.doi.org/10.1016/j.tics.2005.11.008
2005
Earlier work this paper cites.
Hinton GE, Osindero S, Teh YW (2006) “A Fast Learning Algorithm for Deep Belief Nets”. Neural Comput 18(7):1527–1554
2006
Earlier work this paper cites.
Bengio Y, Lamblin P, Popovici D, Larochelle H (2007) “Greedy Layer-Wise Training of Deep Networks”. In: Advances in Neural Information Processing Systems, MIT Press, pp 153–160
2007
Earlier work this paper cites.
Hoiem D, Efros AA, Hebert M (2007) “Recovering Surface Layout from an Image”. Int J Comput Vision 75(1):151–172
2007
Cited alongside, same era.
Rabinovich A, Vedaldi A, Galleguillos C, Wiewiora E, Belongie S (2007) “Objects in context”. In: ICCV
2007
Cited alongside, same era.
Ranzato M, Poultney C, Chopra S, LeCun Y (2007) “Efficient Learning of Sparse Representations with an Energy-Based Model”. In: Advances in Neural Information Processing Systems, MIT Press, pp 1137–1144
2007
Cited alongside, same era.
Russell BC, Torralba A, Murphy KP, Freeman WT (2008) “LabelMe: A Database and Web-Based Tool for Image Annotation”. Int J Comput Vision 77(1-3):157–173
2008
Cited alongside, same era.
Deng J, Dong W, Socher R, Li LJ, Li K, Fei-Fei L (2009) “ImageNet: A Large-Scale Hierarchical Image Database”. In: CVPR09
2009
Cited alongside, same era.
Silberman N, Hoiem D, Kohli P, Fergus R (2012) “Indoor Segmentation and Support Inference from RGBD Images”. In: ECCV
2012
Later among the works it cites.
Sznitman R, Richa R, Taylor RH, Jedynak B, Hager GD (2013) “Unified Detection and Tracking of Instruments during Retinal Microsurgery”. IEEE Transactions on Pattern Analysis and Machine Intelligence 35(5):1263–1273
2013
Later among the works it cites.
Uijlings J, van de Sande K, Gevers T, Smeulders A (2013) “Selective Search for Object Recognition”. International Journal of Computer Vision
2013
Later among the works it cites.
Branson S, Van Horn G, Wah C, Perona P, Belongie S (2014) The ignorant led by the blind: A hybrid human–machine vision system for fine-grained categorization. International Journal of Computer Vision 108(1-2):3–29
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nemirovski A, Juditsky A, Lan G, Shapiro A (2009) “Robust stochastic approximation approach to stochastic programming”. SIAM Journal on Optimization 19(4):1574–1609
2009
Cited alongside, same era.
Saxena A, Sun M, Ng AY (2009) “Make3D: Learning 3D Scene Structure from a Single Still Image”. IEEE Trans Pattern Anal Mach Intell 31(5):824–840
2009
Cited alongside, same era.
Bao SY, Sun M, Savarese S (2010) “Toward Coherent Object Detection And Scene Layout Understanding”. In: CVPR
2010
Cited alongside, same era.
Felzenszwalb PF, Girshick RB, McAllester D, Ramanan D (2010) “Object Detection with Discriminatively Trained Part-Based Models”. IEEE Trans Pattern Anal Mach Intell 32(9):1627–1645
2010
Cited alongside, same era.
Lee DC, Gupta A, Hebert M, Kanade T (2010) “Estimating Spatial Layout of Rooms using Volumetric Reasoning about Objects and Surfaces”. In: NIPS
2010
Cited alongside, same era.
Porway J, Wang K, Zhu SC (2010) “A Hierarchical and Contextual Model for Aerial Image Understanding”. Int’l Journal of Computer Vision 88(2):254–283
2010
Cited alongside, same era.
Sznitman R, Jedynak B (2010) “Active Testing for Face Detection and Localization”. IEEE Transactions on Pattern Analysis and Machine Intelligence 32(10):1914–1920, DOI http://doi.ieeecomputersociety.org/10.1109/TPAMI.2010.106
2010
Cited alongside, same era.
2014
Later among the works it cites.
Hoai M, Zisserman A (2014) “Talking Heads: Detecting Humans and Recognizing Their Interactions”. In: CVPR
2014
Later among the works it cites.
Jia Y, Shelhamer E, Donahue J, Karayev S, Long J, Girshick R, Guadarrama S, Darrell T (2014) “Caffe: Convolutional Architecture for Fast Feature Embedding”. arXiv preprint arXiv:14085093
2014
Later among the works it cites.
Liu X, Zhao Y, Zhu S (2014) “Single-View 3D Scene Parsing by Attributed Grammar”. In: CVPR
2014
Later among the works it cites.
Mottaghi R, Chen X, Liu X, Fidler S, Urtasun R, Yuille A (2014) “The Role of Context for Object Detection and Semantic Segmentation in the Wild”. In: CVPR
2014
Later among the works it cites.
2014
Later among the works it cites.
Sun M, Kim B, Kohli P, Savarese S (2014) “Relating Things and Stuff via ObjectProperty Interactions”. IEEE Trans Pattern Anal Mach Intell 36(7):1370–1383
2014
Later among the works it cites.
Geman D, Geman S, Hallonquist N, Younes L (2015) Visual turing test for computer vision systems. Proceedings of the National Academy of Sciences 112(12):3618–3623
2015
Later among the works it cites.
Minka TP (2012) “The Fastfit Matlab toolbox”. http://research.microsoft.com/en-us/um/people/minka/software/fastfit/ , [Online; accessed 15-Dec-2015]
2015
Later among the works it cites.
Ren S, He K, Girshick R, Sun R (2015) Faster R-CNN: Towards real-time object detection with region proposal networks. In: Advances in Neural Information Processing Systems (NIPS)
2015
Later among the works it cites.
Russakovsky O, Deng J, Su H, Krause J, Satheesh S, Ma S, Huang Z, Karpathy A, Khosla A, Bernstein M, Berg AC, Fei-Fei L (2015) “ImageNet Large Scale Visual Recognition Challenge”. International Journal of Computer Vision (IJCV) 115(3):211–252, DOI 10.1007/s11263-015-0816-y
2015
Later among the works it cites.
Girshick R, Donahue J, Darrell T, Malik J (2016) “Region-Based Convolutional Networks for Accurate Object Detection and Segmentation”. IEEE Transactions on Pattern Analysis and Machine Intelligence 38(1):142–158
2016
Later among the works it cites.
Jahangiri E (2016) On efficient bayesian scene interpretation. PhD thesis, Johns Hopkins University
2016
Later among the works it cites.