Fetching the paper…
Reading the bibliography…
This is an opinion paper about the strengths and weaknesses of Deep Nets for vision.
1911
Earlier work this paper cites.
1912
Earlier work this paper cites.
Mu J, Qiu W, Hager GD, Yuille AL (2019) Learning from synthetic animals. CoRR abs/1912.08265
1912
Earlier work this paper cites.
Zhu L, Chen Y, Torralba A, Freeman WT, Yuille AL (2010) Part and appearance sharing: Recursive compositional models for multi-view. In: CVPR, IEEE Computer Society, pp 1919–1926
1926
Earlier work this paper cites.
Shen Z, Liu Z, Li J, Jiang Y, Chen Y, Xue X (2017b) DSOD: learning deeply supervised object detectors from scratch. In: ICCV, IEEE Computer Society, pp 1937–1945
1945
Earlier work this paper cites.
Green DM, Swets JA (1966) Signal Detection Theory and Psychophysics. John Wiley
1966
Earlier work this paper cites.
Guzmán A (1968) Decomposition of a visual scene into three-dimensional bodies. In: Proceedings of the December 9-11, 1968, fall joint computer conference, part I, pp 291–304
1968
Earlier work this paper cites.
Julesz B (1971) Foundations of cyclopean perception. U. Chicago Press
1971
Earlier work this paper cites.
Gregory RL (1973) Eye and brain: The psychology of seeing. McGraw-Hill
1973
Earlier work this paper cites.
Fukushima K, Miyake S (1982) Neocognitron: A self-organizing neural network model for a mechanism of visual pattern recognition. In: Competition and cooperation in neural nets, Springer, pp 267–285
1982
Earlier work this paper cites.
Marr D (1982) Vision: A computational investigation into the human representation and processing of visual information, henry holt and co. Inc, New York, NY 2(4.2)
1982
Earlier work this paper cites.
Canny JF (1986) A computational approach to edge detection. IEEE Trans Pattern Anal Mach Intell 8(6):679–698
1986
Earlier work this paper cites.
Gibson JJ (1986) The Ecological Approach to Visual Perception. Psychology Press
1986
Earlier work this paper cites.
Rumelhart DE, Hinton GE, Williams RJ (1986) Learning representations by back-propagating errors. nature 323(6088):533–536
1986
Earlier work this paper cites.
Biederman I (1987) Recognition-by-components: a theory of human image understanding. Psychological review 94(2):115
1987
Earlier work this paper cites.
Cybenko G (1989) Approximation by superpositions of a sigmoidal function. MCSS 2(4):303–314
1989
Earlier work this paper cites.
Hornik K, Stinchcombe MB, White H (1989) Multilayer feedforward networks are universal approximators. Neural Networks 2(5):359–366
1989
Earlier work this paper cites.
LeCun Y, Boser BE, Denker JS, Henderson D, Howard RE, Hubbard WE, Jackel LD (1989) Backpropagation applied to handwritten zip code recognition. Neural Computation 1(4):541–551
1989
Earlier work this paper cites.
Pearl J (1989) Probabilistic reasoning in intelligent systems - networks of plausible inference. Morgan Kaufmann series in representation and reasoning, Morgan Kaufmann
1989
Earlier work this paper cites.
Grenander U (1993) General pattern theory-A mathematical study of regular structures. Clarendon Press
1993
Earlier work this paper cites.
Mumford D (1994) Pattern theory: a unifying perspective. In: First European congress of mathematics, Springer, pp 187–224
1994
Earlier work this paper cites.
Xu L, Krzyzak A, Yuille AL (1994) On radial basis function nets and kernel regression: Statistical consistency, convergence rates, and receptive field size. Neural Networks 7(4):609–628
1994
Earlier work this paper cites.
Liu Z, Knill DC, Kersten D (1995) Object classification for human and ideal observers. Vision research 35(4):549–568
1995
Earlier work this paper cites.
Smirnakis SM, Yuille AL (1995) Neural implementation of bayesian vision theories by unsupervised learning. In: The Neurobiology of Computation, Springer, pp 427–432
1995
Earlier work this paper cites.
Tjan BS, Braje WL, Legge GE, Kersten D (1995) Human efficiency for recognizing 3-d objects in luminance noise. Vision research 35(21):3053–3069
1995
Earlier work this paper cites.
Barlow H, Tripathy SP (1997) Correspondence noise and signal pooling in the detection of coherent visual motion. Journal of Neuroscience 17(20):7954–7966
1997
Earlier work this paper cites.
Rensink RA, O’Regan JK, Clark JJ (1997) To see or not to see: The need for attention to perceive changes in scenes. Psychological science 8(5):368–373
1997
Earlier work this paper cites.
Bowyer KW, Kranenburg C, Dougherty S (1999) Edge detector evaluation using empirical ROC curves. In: CVPR, IEEE Computer Society, pp 1354–1359
1999
Earlier work this paper cites.
Gopnik A, Meltzoff AN, Kuhl PK (1999) The scientist in the crib: Minds, brains, and how children learn. William Morrow & Co
1999
Earlier work this paper cites.
Konishi S, Yuille AL, Coughlan JM, Zhu SC (1999) Fundamental bounds on edge detection: An information theoretic evaluation of different edge cues. In: CVPR, IEEE Computer Society, pp 1573–1579
1999
Earlier work this paper cites.
Riesenhuber M, Poggio T (1999) Hierarchical models of object recognition in cortex. Nature neuroscience 2(11):1019
1999
Earlier work this paper cites.
Simons DJ, Chabris CF (1999) Gorillas in our midst: Sustained inattentional blindness for dynamic events. perception 28(9):1059–1074
1999
Earlier work this paper cites.
Poirazi P, Mel BW (2001) Impact of active dendrites and structural plasticity on the memory capacity of neural tissue. Neuron 29(3):779–796
2001
Earlier work this paper cites.
Hoffman J, Tzeng E, Park T, Zhu J, Isola P, Saenko K, Efros AA, Darrell T (2018) Cycada: Cycle-consistent adversarial domain adaptation. In: ICML, PMLR, Proceedings of Machine Learning Research, vol 80, pp 1994–2003
2003
Earlier work this paper cites.
Konishi S, Yuille AL, Coughlan JM, Zhu SC (2003) Statistical edge detection: Learning and evaluating edge cues. IEEE Trans Pattern Anal Mach Intell 25(1):57–74
2003
Earlier work this paper cites.
2003
Earlier work this paper cites.
Lee TS, Mumford D (2003) Hierarchical bayesian inference in the visual cortex. JOSA A 20(7):1434–1448
2003
Earlier work this paper cites.
2003
Earlier work this paper cites.
Tu Z, Chen X, Yuille AL, Zhu SC (2003) Image parsing: Unifying segmentation, detection, and recognition. In: ICCV, IEEE Computer Society, pp 18–25
2003
Earlier work this paper cites.
2003
Earlier work this paper cites.
Gopnik A, Glymour C, Sobel DM, Schulz LE, Kushnir T, Danks D (2004) A theory of causal learning in children: causal maps and bayes nets. Psychological review 111(1):3
2004
Earlier work this paper cites.
Rother C, Kolmogorov V, Blake A (2004) ”grabcut”: interactive foreground extraction using iterated graph cuts. ACM Trans Graph 23(3):309–314
2004
Earlier work this paper cites.
2004
Earlier work this paper cites.
Boyden ES, Zhang F, Bamberg E, Nagel G, Deisseroth K (2005) Millisecond-timescale, genetically targeted optical control of neural activity. Nature neuroscience 8(9):1263
2005
Earlier work this paper cites.
Lu H, Yuille AL (2005) Ideal observers for detecting motion: Correspondence noise. In: NIPS, pp 827–834
2005
Earlier work this paper cites.
Smith L, Gasser M (2005) The development of embodied cognition: Six lessons from babies. Artificial life 11(1-2):13–29
2005
Earlier work this paper cites.
Yuille A, Kersten D (2006) Vision as bayesian inference: analysis by synthesis? Trends in cognitive sciences 10(7):301–308
2006
Earlier work this paper cites.
Zhu S, Mumford D (2006) A stochastic grammar of images. Foundations and Trends in Computer Graphics and Vision 2(4):259–362
2006
Earlier work this paper cites.
Chen Y, Zhu L, Lin C, Yuille AL, Zhang H (2007) Rapid inference on a novel AND/OR graph for object detection, segmentation and parsing. In: NIPS, Curran Associates, Inc., pp 289–296
2007
Earlier work this paper cites.
Geman S (2007) Compositionality in vision. In: The grammar of vision: probabilistic grammar-based models for visual scene understanding and object categorization
2007
Earlier work this paper cites.
Penn DC, Holyoak KJ, Povinelli DJ (2008) Darwin’s mistake: Explaining the discontinuity between human and nonhuman minds. Behavioral and Brain Sciences 31(2):109–130
2008
Earlier work this paper cites.
Yamane Y, Carlson ET, Bowman KC, Wang Z, Connor CE (2008) A neural code for three-dimensional object shape in macaque inferotemporal cortex. Nature neuroscience 11(11):1352–1360
2008
Earlier work this paper cites.
Deng J, Dong W, Socher R, Li L, Li K, Li F (2009) Imagenet: A large-scale hierarchical image database. In: CVPR, IEEE Computer Society, pp 248–255
2009
Earlier work this paper cites.
Pearl J (2009) Causality. Cambridge university press
2009
Cited alongside, same era.
Bashford A, Levine P (2010) The Oxford handbook of the history of eugenics. OUP USA
2010
Cited alongside, same era.
Changizi M (2010) The vision revolution: How the latest research overturns everything we thought we knew about human vision. Benbella books
2010
Cited alongside, same era.
Everingham M, Gool LJV, Williams CKI, Winn JM, Zisserman A (2010) The pascal visual object classes (VOC) challenge. International Journal of Computer Vision 88(2):303–338
2010
Cited alongside, same era.
Felzenszwalb PF, Girshick RB, McAllester DA, Ramanan D (2010) Object detection with discriminatively trained part-based models. IEEE Trans Pattern Anal Mach Intell 32(9):1627–1645
2010
Cited alongside, same era.
Mengistu H, Huizinga J, Mouret J, Clune J (2016) The evolutionary origins of hierarchy. PLoS Computational Biology 12(6)
2016
Later among the works it cites.
Noroozi M, Favaro P (2016) Unsupervised learning of visual representations by solving jigsaw puzzles. In: ECCV (6), Springer, Lecture Notes in Computer Science, vol 9910, pp 69–84
2016
Later among the works it cites.
Qiu W, Yuille AL (2016) Unrealcv: Connecting computer vision to unreal engine. In: ECCV Workshops (3), Lecture Notes in Computer Science, vol 9915, pp 909–916
2016
Later among the works it cites.
Vinyals O, Blundell C, Lillicrap T, Kavukcuoglu K, Wierstra D (2016) Matching networks for one shot learning. In: NIPS, pp 3630–3638
2016
Later among the works it cites.
Wang P, Yuille AL (2016) DOC: deep occlusion estimation from a single image. In: ECCV (1), Springer, Lecture Notes in Computer Science, vol 9905, pp 545–561
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mumford D, Desolneux A (2010) Pattern theory: the stochastic analysis of real-world signals. CRC Press
2010
Cited alongside, same era.
Russell SJ, Norvig P (2010) Artificial Intelligence - A Modern Approach, Third International Edition. Pearson Education
2010
Cited alongside, same era.
Geisler WS (2011) Contributions of ideal observer theory to vision research. Vision research 51(7):771–781
2011
Cited alongside, same era.
McManus JN, Li W, Gilbert CD (2011) Adaptive shape processing in primary visual cortex. Proceedings of the National Academy of Sciences 108(24):9739–9746
2011
Cited alongside, same era.
Torralba A, Efros AA (2011) Unbiased look at dataset bias. In: CVPR, IEEE Computer Society, pp 1521–1528
2011
Cited alongside, same era.
Achanta R, Shaji A, Smith K, Lucchi A, Fua P, Süsstrunk S (2012) SLIC superpixels compared to state-of-the-art superpixel methods. IEEE Trans Pattern Anal Mach Intell 34(11):2274–2282
2012
Cited alongside, same era.
Hoiem D, Chodpathumwan Y, Dai Q (2012) Diagnosing error in object detectors. In: ECCV (3), Springer, Lecture Notes in Computer Science, vol 7574, pp 340–353
2012
Cited alongside, same era.
2016
Later among the works it cites.
Xia F, Wang P, Chen L, Yuille AL (2016) Zoom better to see clearer: Human and object parsing with hierarchical auto-zoom net. In: ECCV (5), Springer, Lecture Notes in Computer Science, vol 9909, pp 648–663
2016
Later among the works it cites.
Yuille AL, Mottaghi R (2016) Complexity of representation and inference in compositional models with part sharing. Journal of Machine Learning Research 17:11:1–11:28
2016
Later among the works it cites.
Zhang R, Isola P, Efros AA (2016) Colorful image colorization. In: ECCV (3), Springer, Lecture Notes in Computer Science, vol 9907, pp 649–666
2016
Later among the works it cites.
Zitnick CL, Agrawal A, Antol S, Mitchell M, Batra D, Parikh D (2016) Measuring machine intelligence through visual question answering. AI Magazine 37(1):63–72
2016
Later among the works it cites.
George D, Lehrach W, Kansky K, Lázaro-Gredilla M, Laan C, Marthi B, Lou X, Meng Z, Liu Y, Wang H, et al. (2017) A generative vision model that trains with high data efficiency and breaks text-based captchas. Science 358(6368):eaag2612
2017
Later among the works it cites.
Guu K, Pasupat P, Liu EZ, Liang P (2017) From language to programs: Bridging reinforcement learning and maximum marginal likelihood. In: ACL (1), Association for Computational Linguistics, pp 1051–1062
2017
Later among the works it cites.
Jégou S, Drozdzal M, Vázquez D, Romero A, Bengio Y (2017) The one hundred layers tiramisu: Fully convolutional densenets for semantic segmentation. In: CVPR Workshops, IEEE Computer Society, pp 1175–1183
2017
Later among the works it cites.
Kokkinos I (2017) Ubernet: Training a universal convolutional neural network for low-, mid-, and high-level vision using diverse datasets and limited memory. In: CVPR, IEEE Computer Society, pp 5454–5463
2017
Later among the works it cites.
Lin X, Wang H, Li Z, Zhang Y, Yuille AL, Lee TS (2017) Transfer of view-manifold learning to similarity perception of novel objects. In: International Conference on Learning Representations
2017
Later among the works it cites.
2017
Later among the works it cites.
Ren Z, Yan J, Ni B, Liu B, Yang X, Zha H (2017) Unsupervised deep learning for optical flow estimation. In: AAAI, AAAI Press, pp 1495–1501
2017
Later among the works it cites.
Sabour S, Frosst N, Hinton GE (2017) Dynamic routing between capsules. In: NIPS, pp 3856–3866
2017
Later among the works it cites.
Tzeng E, Hoffman J, Saenko K, Darrell T (2017) Adversarial discriminative domain adaptation. In: CVPR, pp 2962–2971, DOI 10.1109/CVPR.2017.316
2017
Later among the works it cites.
Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser L, Polosukhin I (2017) Attention is all you need. In: NIPS, pp 5998–6008
2017
Later among the works it cites.
Wen H, Shi J, Zhang Y, Lu KH, Cao J, Liu Z (2017) Neural encoding and decoding with deep learning for dynamic natural vision. Cerebral Cortex pp 1–25
2017
Later among the works it cites.
Xie C, Wang J, Zhang Z, Zhou Y, Xie L, Yuille AL (2017) Adversarial examples for semantic segmentation and object detection. In: ICCV, IEEE Computer Society, pp 1378–1387
2017
Later among the works it cites.
Xie L, Yuille AL (2017) Genetic CNN. In: ICCV, IEEE Computer Society, pp 1388–1397
2017
Later among the works it cites.
Zhou T, Brown M, Snavely N, Lowe DG (2017) Unsupervised learning of depth and ego-motion from video. In: CVPR, IEEE Computer Society, pp 6612–6619
2017
Later among the works it cites.
Zhu Z, Xie L, Yuille AL (2017) Object recognition with and without objects. In: IJCAI, ijcai.org, pp 3609–3615
2017
Later among the works it cites.
Zoph B, Le QV (2017) Neural architecture search with reinforcement learning. In: ICLR, OpenReview.net
2017
Later among the works it cites.
Athalye A, Carlini N, Wagner DA (2018) Obfuscated gradients give a false sense of security: Circumventing defenses to adversarial examples. In: ICML, JMLR.org, JMLR Workshop and Conference Proceedings, vol 80, pp 274–283
2018
Closest in time.
Buolamwini J, Gebru T (2018) Gender shades: Intersectional accuracy disparities in commercial gender classification. In: Conference on fairness, accountability and transparency, pp 77–91
2018
Closest in time.
Chen L, Papandreou G, Kokkinos I, Murphy K, Yuille AL (2018) Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs. IEEE Trans Pattern Anal Mach Intell 40(4):834–848
2018
Closest in time.
Darwiche A (2018) Human-level intelligence or animal-like abilities? Commun ACM 61(10):56–67
2018
Closest in time.
Liu C, Zoph B, Neumann M, Shlens J, Hua W, Li L, Fei-Fei L, Yuille AL, Huang J, Murphy K (2018) Progressive neural architecture search. In: ECCV (1), Springer, Lecture Notes in Computer Science, vol 11205, pp 19–35
2018
Closest in time.
Marcus G (2018) Deep learning: A critical appraisal. CoRR abs/1801.00631
2018
Closest in time.
Murez Z, Kolouri S, Kriegman DJ, Ramamoorthi R, Kim K (2018) Image to image translation for domain adaptation. In: CVPR, pp 4500–4509, DOI 10.1109/CVPR.2018.00473
2018
Closest in time.
Pham H, Guan MY, Zoph B, Le QV, Dean J (2018) Efficient neural architecture search via parameter sharing. In: ICML, PMLR, Proceedings of Machine Learning Research, vol 80, pp 4092–4101
2018
Closest in time.
Qiao S, Liu C, Shen W, Yuille AL (2018) Few-shot image recognition by predicting parameters from activations. In: CVPR, IEEE Computer Society, pp 7229–7238
2018
Closest in time.
Rosenfeld A, Zemel RS, Tsotsos JK (2018) The elephant in the room. CoRR abs/1808.03305
2018
Closest in time.
Santoro A, Hill F, Barrett DGT, Morcos AS, Lillicrap TP (2018) Measuring abstract reasoning in neural networks. In: ICML, JMLR.org, JMLR Workshop and Conference Proceedings, vol 80, pp 4477–4486
2018
Closest in time.
Uesato J, O’Donoghue B, Kohli P, van den Oord A (2018) Adversarial risk and the dangers of evaluating against weak attacks. In: ICML, PMLR, Proceedings of Machine Learning Research, vol 80, pp 5032–5041
2018
Closest in time.
Wang J, Zhang Z, Xie C, Zhou Y, Premachandran V, Zhu J, Xie L, Yuille A (2018) Visual concepts and compositional voting. Annals of Mathematical Sciences and Applications 2(3):4
2018
Closest in time.
Wu Z, Xiong Y, Yu SX, Lin D (2018) Unsupervised feature learning via non-parametric instance discrimination. In: CVPR, IEEE Computer Society, pp 3733–3742
2018
Closest in time.
Xie C, Wang J, Zhang Z, Ren Z, Yuille AL (2018) Mitigating adversarial effects through randomization. In: International Conference on Learning Representations
2018
Closest in time.
Zhang Y, Qiu W, Chen Q, Hu X, Yuille AL (2018) Unrealstereo: Controlling hazardous factors to analyze stereo vision. In: 3DV, IEEE Computer Society, pp 228–237
2018
Closest in time.
Alcorn MA, Li Q, Gong Z, Wang C, Mai L, Ku W, Nguyen A (2019) Strike (with) a pose: Neural networks are easily fooled by strange poses of familiar objects. In: CVPR, Computer Vision Foundation / IEEE, pp 4845–4854
2019
Closest in time.
Liu R, Liu C, Bai Y, Yuille AL (2019) Clevr-ref+: Diagnosing visual reasoning with referring expressions. In: CVPR, Computer Vision Foundation / IEEE, pp 4185–4194
2019
Closest in time.
Tsipras D, Santurkar S, Engstrom L, Turner A, Madry A (2019) Robustness may be at odds with accuracy. In: ICLR (Poster), OpenReview.net
2019
Closest in time.
Wang T, Zhao J, Yatskar M, Chang K, Ordonez V (2019) Balanced datasets are not enough: Estimating and mitigating gender bias in deep image representations. In: ICCV, IEEE, pp 5309–5318
2019
Closest in time.
Zhou Z, Firestone C (2019) Humans can decipher adversarial images. Nature communications 10(1):1–9
2019
Closest in time.
Zhu H, Tang P, Yuille AL, Park S, Park J (2019) Robustness of object recognition under extreme occlusion in humans and computational models. In: CogSci, cognitivesciencesociety.org, pp 3213–3219
2019
Closest in time.
Firestone C (2020) Performance vs. competence in human-machine comparisons. Proceedings of the National Academy of Sciences In Press
2020
Closest in time.
Kaushik D, Hovy EH, Lipton ZC (2020) Learning the difference that makes A difference with counterfactually-augmented data. In: ICLR, OpenReview.net
2020
Closest in time.
Shu M, Liu C, Qiu W, Yuille AL (2020) Identifying model weakness with adversarial examiner. In: AAAI, AAAI Press, pp 11998–12006
2020
Closest in time.
Zhang Z, Shen W, Qiao S, Wang Y, Wang B, Yuille AL (2020) Robust face detection via learning small faces on hard images. In: WACV, IEEE, pp 1350–1359
2020
Closest in time.
Zendel O, Murschitz M, Humenberger M, Herzner W (2015) CV-HAZOP: introducing test data validation for computer vision. In: ICCV, IEEE Computer Society, pp 2066–2074
2074
Closest in time.