Fetching the paper…
Reading the bibliography…
We present novel method for image-text multi-modal representation learning.
Bidirectional recurrent neural networks
Schuster, Mike and Paliwal, Kuldip K · 1997
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, Jia, Dong, Wei, Socher, Richard, Li, Li-Jia, Li, Kai, and Fei-Fei, Li · 2009
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, Vinod and Hinton, Geoffrey E · 2010
Earlier work this paper cites.
Multimodal learning with deep boltzmann machines
Srivastava, Nitish and Salakhutdinov, Ruslan R · 2012
Earlier work this paper cites.
Devise: A deep visual-semantic embedding model
Frome, Andrea, Corrado, Greg S, Shlens, Jon, Bengio, Samy, Dean, Jeff, Mikolov, Tomas, et al · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Mikolov, Tomas, Sutskever, Ilya, Chen, Kai, Corrado, Greg S, and Dean, Jeff · 2013
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, Ian, Pouget-Abadie, Jean, Mirza, Mehdi, Xu, Bing, Warde-Farley, David, Ozair, Sherjil, Courville, Aaron, and Bengio, Yoshua · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Kim, Yoon · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, Diederik P. and Ba, Jimmy · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, Tsung-Yi, Maire, Michael, Belongie, Serge, Hays, James, Perona, Pietro, Ramanan, Deva, Dollár, Piotr, and Zitnick, C Lawrence · 2014
Cited alongside, same era.
Deep captioning with multimodal recurrent neural networks (m-rnn)
Mao, Junhua, Xu, Wei, Yang, Yi, Wang, Jiang, Huang, Zhiheng, and Yuille, Alan · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, Karen and Zisserman, Andrew · 2014
Cited alongside, same era.
Improved multimodal deep learning with variation of information
Sohn, Kihyuk, Shang, Wenling, and Lee, Honglak · 2014
Cited alongside, same era.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, Nitish, Hinton, Geoffrey E, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan · 2014
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Later among the works it cites.
Deep visual-semantic alignments for generating image descriptions
Karpathy, Andrej and Fei-Fei, Li · 2015
Later among the works it cites.
Associating neural word embeddings with deep image representations using fisher vectors
Klein, Benjamin, Lev, Guy, Sadeh, Gil, and Wolf, Lior · 2015
Later among the works it cites.
Multimodal convolutional neural networks for matching image and sentence
Ma, Lin, Lu, Zhengdong, Shang, Lifeng, and Li, Hang · 2015
Later among the works it cites.
Unsupervised representation learning with deep convolutional generative adversarial networks
Radford, Alec, Metz, Luke, and Chintala, Soumith · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Accelerating t-sne using tree-based algorithms
Van Der Maaten, Laurens · 2014
Cited alongside, same era.
A probabilistic covariate shift assumption for domain adaptation
Adel, Tameem and Wong, Alexander · 2015
Cited alongside, same era.
Unsupervised domain adaptation by backpropagation
Ganin, Yaroslav and Lempitsky, Victor · 2015
Cited alongside, same era.
Fast r-cnn
Girshick, Ross · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
Vinyals, Oriol, Toshev, Alexander, Bengio, Samy, and Erhan, Dumitru · 2015
Later among the works it cites.
Generative adversarial text to image synthesis
Reed, Scott E., Akata, Zeynep, Yan, Xinchen, Logeswaran, Lajanugen, Schiele, Bernt, and Lee, Honglak · 2016
Closest in time.
Learning deep structure-preserving image-text embeddings
Wang, Liwei, Li, Yin, and Lazebnik, Svetlana · 2016
Closest in time.