Fetching the paper…
Reading the bibliography…
We present a simple and effective architecture for fine-grained visual recognition called Bilinear Convolutional Neural Networks (B-CNNs).
Object recognition from local scale-invariant features
D. G. Lowe · 1999
Earlier work this paper cites.
A parametric texture model based on joint statistics of complex wavelet coefficients
J. Portilla and E. Simoncelli · 2000
Earlier work this paper cites.
Separating style and content with bilinear models
J. B. Tenenbaum and W. T. Freeman · 2000
Earlier work this paper cites.
Representing and recognizing the visual appearance of materials using three-dimensional textons
T. Leung and J. Malik · 2001
Earlier work this paper cites.
Learning with kernels: support vector machines, regularization, optimization, and beyond
B. Scholkopf and A. J. Smola · 2001
Earlier work this paper cites.
Visual categorization with bags of keypoints
G. Csurka, C. R. Dance, L. Dan, J. Willamowski, and C. Bray · 2004
Earlier work this paper cites.
Class-specific material categorisation
B. Caputo, E. Hayman, and P. Mallikarjuna · 2005
Earlier work this paper cites.
The PASCAL visual obiect classes challenge 2007 (VOC2007) results
M. Everingham, A. Zisserman, C. Williams, and L. V. Gool · 2007
Earlier work this paper cites.
Fisher kernels on visual vocabularies for image categorization
F. Perronnin and C. R. Dance · 2007
Earlier work this paper cites.
Bilinear classifiers for visual recognition
H. Pirsiavash, D. Ramanan, and C. C. Fowlkes · 2009
Earlier work this paper cites.
Recognizing indoor scenes
A. Quattoni and A. Torralba · 2009
Earlier work this paper cites.
Material perceprion: What can you see in a brief glance?
L. Sharan, R. Rosenholtz, and E. H. Adelson · 2009
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation
H. Jégou, M. Douze, C. Schmid, and P. Pérez · 2010
Earlier work this paper cites.
Improving the Fisher kernel for large-scale image classification
F. Perronnin, J. Sánchez, and T. Mensink · 2010
Earlier work this paper cites.
VLFeat: an open and portable library of computer vision algorithms
A. Vedaldi and B. Fulkerson · 2010
Earlier work this paper cites.
Describing people: A poselet-based approach to attribute classification
L. Bourdev, S. Maji, and J. Malik · 2011
Earlier work this paper cites.
Birdlets: Subordinate categorization using volumetric primitives and pose-normalized appearance
R. Farrell, O. Oza, N. Zhang, V. I. Morariu, T. Darrell, and L. S. Davis · 2011
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 Dataset
C. Wah, S. Branson, P. Welinder, P. Perona, and S. Belongie · 2011
Earlier work this paper cites.
Semantic segmentation with second-order pooling
J. Carreira, R. Caseiro, J. Batista, and C. Sminchisescu · 2012
Earlier work this paper cites.
Discriminative learning of sum-product networks
R. Gens and P. Domingos · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Pose pooling kernels for sub-category recognition
N. Zhang, R. Farrell, and T. Darrell · 2012
Earlier work this paper cites.
Decaf: A deep convolutional activation feature for generic visual recognition
J. Donahue, Y. Jia, O. Vinyals, J. Hoffman, N. Zhang, E. Tzeng, and T. Darrell · 2013
Earlier work this paper cites.
3d object representations for fine-grained categorization
J. Krause, M. Stark, J. Deng, and L. Fei-Fei · 2013
Earlier work this paper cites.
Fine-grained visual classification of aircraft
S. Maji, E. Rahtu, J. Kannala, M. Blaschko, and A. Vedaldi · 2013
Cited alongside, same era.
Fast and scalable polynomial kernels via explicit feature maps
N. Pham and R. Pagh · 2013
Cited alongside, same era.
Bird species categorization using pose normalized deep convolutional nets
S. Branson, G. V. Horn, S. Belongie, and P. Perona · 2014
Cited alongside, same era.
Return of the devil in the details: Delving deep into convolutional nets
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Cited alongside, same era.
Describing textures in the wild
M. Cimpoi, S. Maji, I. Kokkinos, S. Mohamed, and A. Vedaldi · 2014
Cited alongside, same era.
Rich feature hierarchies for accurate object detection and semantic segmentation
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Closest in time.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Closest in time.
Building a bird recognition app and large scale dataset with citizen scientists: The fine print in fine-grained dataset collection
G. Van Horn, S. Branson, R. Farrell, S. Haber, J. Barry, P. Ipeirotis, P. Perona, and S. Belongie · 2015
Closest in time.
MatConvNet – Convolutional Neural Networks for MATLAB
A. Vedaldi and K. Lenc · 2015
Closest in time.
NetVLAD: CNN architecture for weakly supervised place recognition
R. Arandjelović, P. Gronat, A. Torii, T. Pajdla, and J. Sivic · 2016
Closest in time.
Deep filter banks for texture recognition, description, and segmentation
M. Cimpoi, S. Maji, I. Kokkinos, and A. Vedaldi · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. B. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Cited alongside, same era.
Multi-scale orderless pooling of deep convolutional activation features
Y. Gong, L. Wang, R. Guo, and S. Lazebnik · 2014
Cited alongside, same era.
Revisiting the fisher vector for fine-grained classification
P.-H. Gosselin, N. Murray, H. Jégou, and F. Perronnin · 2014
Cited alongside, same era.
Recurrent models of visual attention
V. Mnih, N. Heess, A. Graves, and K. Kavukcuoglu · 2014
Cited alongside, same era.
A compact and discriminative face track descriptor
O. M. Parkhi, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Cited alongside, same era.
CNN features off-the-shelf: An astounding baseline for recognition
A. S. Razavin, H. Azizpour, J. Sullivan, and S. Carlsson · 2014
Cited alongside, same era.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Closest in time.
Convolutional two-stream network fusion for video action recognition
C. Feichtenhofer, A. Pinz, and A. Zisserman · 2016
Closest in time.
Multimodal compact bilinear pooling for visual question answering and visual grounding
A. Fukui, D. H. Park, D. Yang, A. Rohrbach, T. Darrell, and M. Rohrbach · 2016
Closest in time.
Compact bilinear pooling
Y. Gao, O. Beijbom, N. Zhang, and T. Darrell · 2016
Closest in time.
Image style transfer using convolutional neural networks
L. A. Gatys, A. S. Ecker, and M. Bethge · 2016
Closest in time.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Closest in time.
The unreasonable effectiveness of noisy data for fine-grained recognition
J. Krause, B. Sapp, A. Howard, H. Zhou, A. Toshev, T. Duerig, J. Philbin, and L. Fei-Fei · 2016
Closest in time.
Visualizing and understanding deep texture representations
T.-Y. Lin and S. Maji · 2016
Closest in time.
Cross-convolutional-layer pooling for image recognition
L. Liu, C. Shen, and A. van den Hengel · 2016
Closest in time.
Ssd: Single shot multibox detector
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.-Y. Fu, and A. C. Berg · 2016
Closest in time.
Visualizing deep convolutional neural networks using natural pre-images
A. Mahendran and A. Vedaldi · 2016
Closest in time.
Boosted convolutional neural networks
M. Moghimi, S. Belongie, M. Saberian, J. Yang, N. Vasconcelos, and L.-J. Li · 2016
Closest in time.
Texture synthesis using shallow convolutional networks with random filters
I. Ustyuzhaninov, W. Brendel, L. A. Gatys, and M. Bethge · 2016
Closest in time.
Mining discriminative triplets of patches for fine-grained classification
Y. Wang, J. Choi, V. Morariu, and L. S. Davis · 2016
Closest in time.
SPDA-CNN: unifying semantic part detection and abstraction for fine-grained recognition
H. Zhang, T. Xu, M. Elhoseiny, X. Huang, S. Zhang, A. Elgammal, and D. Metaxas · 2016
Closest in time.
Fine-grained pose prediction, normalization, and recognition
N. Zhang, E. Shelhamer, Y. Gao, and T. Darrell · 2016
Closest in time.
Picking deep filter responses for fine-grained image recognition
X. Zhang, H. Xiong, W. Zhou, W. Lin, and Q. Tian · 2016
Closest in time.
Accessed: 2017-03-28
2017
Closest in time.
Accessed: 2017-03-30
2017
Closest in time.