Fetching the paper…
Reading the bibliography…
While 360{\deg} cameras offer tremendous new possibilities in vision, graphics, and augmented reality, the spherical images they produce make core feature extraction non-trivial.
Curvilinear perspective, 1987
A. Barre, A. Flocon, and R. Hansen · 1987
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Squaring the circle in panoramas
L. Zelnik-Manor, G. Peters, and P. Perona · 2005
Earlier work this paper cites.
Model compression
C. Buciluǎ, R. Caruana, and A. Niculescu-Mizil · 2006
Earlier work this paper cites.
Scale-invariant features on the sphere
P. Hansen, P. Corke, W. Boles, and K. Daniilidis · 2007
Earlier work this paper cites.
Scale invariant feature matching with wide angle images
P. Hansen, P. Corket, W. Boles, and K. Daniilidis · 2007
Earlier work this paper cites.
Imagenet: a large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Earlier work this paper cites.
Recognizing scene viewpoint using panoramic place representation
J. Xiao, K. A. Ehinger, A. Oliva, and A. Torralba · 2012
Earlier work this paper cites.
Do deep nets really need to be deep?
J. Ba and R. Caruana · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
Caffe: Convolutional architecture for fast feature embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2014
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Panocontext: A whole-room 3d context model for panoramic scene understanding
Y. Zhang, S. Song, P. Tan, and J. Xiao · 2014
Cited alongside, same era.
Learning deep features for scene recognition using places database
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva · 2014
Cited alongside, same era.
The pascal visual object classes challenge: A retrospective
M. Everingham, S. M. A. Eslami, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
G. Hinton, O. Vinyals, and J. Dean · 2015
Cited alongside, same era.
Spatial transformer networks
M. Jaderberg, K. Simonyan, A. Zisserman, et al · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Pano2vid: Automatic cinematography for watching 360° videos
Y.-C. Su, D. Jayaraman, and K. Grauman · 2016
Later among the works it cites.
Learning to learn: Model regression networks for easy small sample learning
Y.-X. Wang and M. Hebert · 2016
Later among the works it cites.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2016
Later among the works it cites.
Convolutional networks for spherical signals
T. Cohen, M. Geiger, J. Köhler, and M. Welling · 2017
Closest in time.
Deformable convolutional networks
J. Dai, H. Qi, Y. Xiong, Y. Li, G. Zhang, H. Hu, and Y. Wei · 2017
Closest in time.
Affine covariant features for fisheye distortion local modeling
A. Furnari, G. M. Farinella, A. R. Bruna, and S. Battiato · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
Fitnets: Hints for thin deep nets
A. Romero, N. Ballas, S. E. Kahou, A. Chassang, C. Gatta, and Y. Bengio · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.
Learning spatiotemporal features with 3d convolutional networks
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2015
Cited alongside, same era.
Convolutional two-stream network fusion for video action recognition
C. Feichtenhofer, A. Pinz, and A. Zisserman · 2016
Cited alongside, same era.
Cross modal distillation for supervision transfer
S. Gupta, J. Hoffman, and J. Malik · 2016
Cited alongside, same era.
Closest in time.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Closest in time.
Deep 360 pilot: Learning a deep agent for piloting through 360 ° 360\degree sports video
H.-N. Hu, Y.-C. Lin, M.-Y. Liu, H.-T. Cheng, Y.-J. Chang, and M. Sun · 2017
Closest in time.
Fusionseg: Learning to combine motion and appearance for fully automatic segmentation of generic objects in video
S. Jain, B. Xiong, and K. Grauman · 2017
Closest in time.
Active convolution: Learning the shape of convolution for image classification
Y. Jeon and J. Kim · 2017
Closest in time.
Graph-based classification of omnidirectional images
R. Khasanova and P. Frossard · 2017
Closest in time.
Semantic-driven generation of hyperlapse from 360° video
W.-S. Lai, Y. Huang, N. Joshi, C. Buehler, M.-H. Yang, and S. B. Kang · 2017
Closest in time.
Making 360° video watchable in 2d: Learning videography for click free viewing
Y.-C. Su and K. Grauman · 2017
Closest in time.