A training algorithm for optimal margin classifiers
Bernhard E Boser, Isabelle M Guyon, and Vladimir N Vapnik · 1992
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Local aggregation for unsupervised learning of visual embeddings
Chengxu Zhuang, Alex Lin Zhai, and Daniel Yamins · 2002
Earlier work this paper cites.
Visual categorization with bags of keypoints
Gabriella Csurka, Christopher Dance, Lixin Fan, Jutta Willamowski, and Cédric Bray · 2004
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
David G Lowe · 2004
Earlier work this paper cites.
Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories
Svetlana Lazebnik, Cordelia Schmid, and Jean Ponce · 2006
Earlier work this paper cites.
Video google: Efficient visual search of videos
Josef Sivic and Andrew Zisserman · 2006
Earlier work this paper cites.
Fisher kernels on visual vocabularies for image categorization
Florent Perronnin and Christopher Dance · 2007
Earlier work this paper cites.
Evaluating bag-of-visual-words representations in scene classification
Jun Yang, Yu-Gang Jiang, Alexander G Hauptmann, and Chong-Wah Ngo · 2007
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
Pascal Vincent, Hugo Larochelle, Yoshua Bengio, and Pierre-Antoine Manzagol · 2008
Earlier work this paper cites.
On the burstiness of visual elements
Hervé Jégou, Matthijs Douze, and Cordelia Schmid · 2009
Earlier work this paper cites.
Packing bag-of-features
Hervé Jégou, Matthijs Douze, and Cordelia Schmid · 2009
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
M. Everingham, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation
Hervé Jégou, Matthijs Douze, Cordelia Schmid, and Patrick Pérez · 2010
Earlier work this paper cites.
Improving the fisher kernel for large-scale image classification
Florent Perronnin, Jorge Sánchez, and Thomas Mensink · 2010
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Original
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
To aggregate or not to aggregate: Selective match kernels for image search
Giorgos Tolias, Yannis Avrithis, and Hervé Jégou · 2013
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
David Eigen, Christian Puhrsch, and Rob Fergus · 2014
Earlier work this paper cites.
Multi-scale orderless pooling of deep convolutional activation features
Yunchao Gong, Liwei Wang, Ruiqi Guo, and Svetlana Lazebnik · 2014
Earlier work this paper cites.
Learning deep features for scene recognition using places database
Bolei Zhou, Agata Lapedriza, Jianxiong Xiao, Antonio Torralba, and Aude Oliva · 2014
Earlier work this paper cites.
Unsupervised visual representation learning by context prediction
Carl Doersch, Abhinav Gupta, and Alexei A Efros · 2015
Earlier work this paper cites.
Discriminative unsupervised feature learning with exemplar convolutional neural networks
Alexey Dosovitskiy, Philipp Fischer, Jost Tobias Springenberg, Martin Riedmiller, and Thomas Brox · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Original
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Skip-thought vectors
Ryan Kiros, Yukun Zhu, Ruslan R Salakhutdinov, Richard Zemel, Raquel Urtasun, Antonio Torralba, and Sanja Fidler · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.