Rich feature hierarchies for accurate object detection and semantic segmentation
Girshick, R., Donahue, J., Darrell, T., Malik, J.: · 2014
Closest in time.
Joint training of a convolutional network and a graphical model for human pose estimation
Tompson, J., Jain, A., LeCun, Y., Bregler, C.: · 2014
Closest in time.
Spatial pyramid pooling in deep convolutional networks for visual recognition
He, K., Zhang, X., Ren, S., Sun, J.: · 2014
Closest in time.
Deeply-supervised nets
Original
Lee, C.Y., Xie, S., Gallagher, P., Zhang, Z., Tu, Z.: · 2014
Closest in time.
Very deep convolutional networks for large-scale image recognition
Original
Simonyan, K., Zisserman, A.: · 2014
Closest in time.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., van Merrienboer, B., Gulcehre, C., Bougares, F., Schwenk, H., Bahdanau, D., Bengio, Y.: · 2014
Closest in time.
Coherent multi-sentence video description with variable level of detail
Rohrbach, A., Rohrbach, M., Qiu, W., Friedrich, A., Amin, S., Andriluka, M., Pinkal, M., Schiele, B.: · 2014
Closest in time.
Grounded compositional semantics for finding and describing images with sentences
Socher, R., Karpathy, A., Le, Q., Manning, C., Ng, A.: · 2014
Closest in time.
Deep fragment embeddings for bidirectional image sentence mapping
Karpathy, A., Joulin, A., Fei-Fei, L.: · 2014
Closest in time.
Explain images with multimodal recurrent neural networks
Original
Mao, J., Xu, W., Yang, Y., Wang, J., Yuille, A.L.: · 2014
Closest in time.
What are you talking about? text-to-image coreference
Kong, C., Lin, D., Bansal, M., Urtasun, R., Fidler, S.: · 2014
Closest in time.
Semantic parsing via paraphrasing
Berant, J., Liang, P.: · 2014
Closest in time.
A neural network for factoid question answering over paragraphs
Iyyer, M., Boyd-Graber, J., Claudino, L., Socher, R., III, H.D.: · 2014
Closest in time.
Open question answering over curated and extracted knowledge bases
Fader, A., Zettlemoyer, L., Etzioni, O.: · 2014
Closest in time.
A multi-world approach to question answering about real-world scenes based on uncertain input
Malinowski, M., Fritz, M.: · 2014
Closest in time.
LSDA: Large scale detection through adaptation
Hoffman, J., Guadarrama, S., Tzeng, E., Hu, R., Donahue, J., Girshick, R., Darrell, T., Saenko, K.: · 2014
Closest in time.
Low-dimensional embeddings of logic
Rocktäschel, T., Bosnjak, M., Singh, S., Riedel, S.: · 2014
Closest in time.
Combining formal and distributional models of temporal and intensional semantics
Lewis, M., Steedman, M.: · 2014
Closest in time.
Imagenet large scale visual recognition challenge
Original
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., Berg, A.C., Fei-Fei, L.: · 2014
Closest in time.
Improving image-sentence embeddings using large weakly annotated photo collections
Gong, Y., Wang, L., Hodosh, M., Hockenmaier, J., Lazebnik, S.: · 2014
Closest in time.
Glove: Global vectors for word representation
Pennington, J., Socher, R., Manning, C.D.: · 2014
Closest in time.