Fetching the paper…
Reading the bibliography…
Convolutional neural networks (CNNs) are inherently limited to model geometric transformations due to the fixed geometric structures in its building modules.
Representation of local geometry in the visual system
J. J. Koenderink and A. J. van Doom · 1987
Earlier work this paper cites.
A real-time algorithm for signal analysis with the help of the wavelet transform
M. Holschneider, R. Kronland-Martinet, J. Morlet, and P. Tchamitchian · 1989
Earlier work this paper cites.
The design and use of steerable filters
W. T. Freeman and E. H. Adelson · 1991
Earlier work this paper cites.
Convolutional networks for images, speech, and time series
Y. LeCun and Y. Bengio · 1995
Earlier work this paper cites.
Deformable kernels for early vision
P. Perona · 1995
Earlier work this paper cites.
Object recognition from local scale-invariant features
D. G. Lowe · 1999
Earlier work this paper cites.
Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories
S. Lazebnik, C. Schmid, and J. Ponce · 2006
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
A theoretical analysis of feature pooling in visual recognition
Y.-L. Boureau, J. Ponce, and Y. LeCun · 2010
Earlier work this paper cites.
The PASCAL Visual Object Classes (VOC) Challenge
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Object detection with discriminatively trained part-based models
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan · 2010
Earlier work this paper cites.
Semantic contours from inverse detectors
B. Hariharan, P. Arbeláez, L. Bourdev, S. Maji, and J. Malik · 2011
Earlier work this paper cites.
Orb: an efficient alternative to sift or surf
E. Rublee, V. Rabaud, K. Konolige, and G. Bradski · 2011
Earlier work this paper cites.
Beyond spatial pyramids: Receptive field learning for pooled image features
Y. Jia, C. Huang, and T. Darrell · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Learning invariant representations with local transformations
K. Sohn and H. Lee · 2012
Earlier work this paper cites.
Invariant scattering convolution networks
J. Bruna and S. Mallat · 2013
Earlier work this paper cites.
Deep symmetry networks
R. Gens and P. M. Domingos · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Cited alongside, same era.
Deformable part models are convolutional neural networks
R. Girshick, F. Iandola, T. Darrell, and J. Malik · 2014
Cited alongside, same era.
Simultaneous detection and segmentation
B. Hariharan, P. Arbeláez, R. Girshick, and J. Malik · 2014
Cited alongside, same era.
Spatial pyramid pooling in deep convolutional networks for visual recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2014
Cited alongside, same era.
Locally scale-invariant convolutional neural networks
A. Kanazawa, A. Sharma, and D. Jacobs · 2014
Cited alongside, same era.
Microsoft COCO: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
The cityscapes dataset for semantic urban scene understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Later among the works it cites.
R-fcn: Object detection via region-based fully convolutional networks
J. Dai, Y. Li, K. He, and J. Sun · 2016
Later among the works it cites.
Exploiting cyclic symmetry in convolutional neural networks
S. Dieleman, J. D. Fauw, and K. Kavukcuoglu · 2016
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
Speed/accuracy trade-offs for modern convolutional object detectors
J. Huang, V. Rathod, C. Sun, M. Zhu, A. Korattikara, A. Fathi, I. Fischer, Z. Wojna, Y. Song, S. Guadarrama, and K. Murphy · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Scalable, high-quality object detection
C. Szegedy, S. Reed, D. Erhan, and D. Anguelov · 2014
Cited alongside, same era.
Semantic image segmentation with deep convolutional nets and fully connected crfs
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille · 2015
Cited alongside, same era.
Object detection via a multi-region & semantic segmentation-aware cnn model
S. Gidaris and N. Komodakis · 2015
Cited alongside, same era.
Fast R-CNN
R. Girshick · 2015
Cited alongside, same era.
Spatial transformer networks
M. Jaderberg, K. Simonyan, A. Zisserman, and K. Kavukcuoglu · 2015
Cited alongside, same era.
Transformation-invariant convolutional jungles
D. Laptev and J. M.Buhmann · 2015
Cited alongside, same era.
Structured receptive fields in cnns
J.-H. Jacobsen, J. van Gemert, Z. Lou, and A. W.M.Smeulders · 2016
Later among the works it cites.
Ti-pooling: transformation-invariant pooling for feature learning in convolutional neural networks
D. Laptev, N. Savinov, J. M. Buhmann, and M. Pollefeys · 2016
Later among the works it cites.
Inverse compositional spatial transformer networks
C.-H. Lin and S. Lucey · 2016
Later among the works it cites.
Ssd: Single shot multibox detector
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, and S. Reed · 2016
Later among the works it cites.
You only look once: Unified, real-time object detection
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi · 2016
Later among the works it cites.
Faster R-CNN: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2016
Later among the works it cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
C. Szegedy, S. Ioffe, V. Vanhoucke, and A. Alemi · 2016
Later among the works it cites.
Harmonic networks: Deep translation and rotation equivariance
D. E. Worrall, S. J. Garbin, D. Turmukhambetov, and G. J. Brostow · 2016
Later among the works it cites.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2016
Later among the works it cites.
Active convolution: Learning the shape of convolution for image classification
Y. Jeon and J. Kim · 2017
Closest in time.
Feature pyramid networks for object detection
T.-Y. Lin, P. Dollár, R. Girshick, K. He, B. Hariharan, and S. Belongie · 2017
Closest in time.
Understanding the effective receptive field in deep convolutional neural networks
W. Luo, Y. Li, R. Urtasun, and R. Zemel · 2017
Closest in time.
Dilated residual networks
F. Yu, V. Koltun, and T. Funkhouser · 2017
Closest in time.