Fetching the paper…
Reading the bibliography…
Generating natural questions from an image is a semantic task that requires using vision and language modalities to learn multimodal representations.
Large automatic learning, rule extraction, and generalization
J. Denker, D. Schwartz, B. Wittner, S. Solla, R. Howard, L. Jackel, and J. Hopfield · 1987
Earlier work this paper cites.
Consistent inference of probabilities in layered networks: Predictions and generalization
N. Tishby, E. Levin, and S. A. Solla · 1989
Earlier work this paper cites.
Bayesian back-propagation
W. L. Buntine and A. S. Weigend · 1991
Earlier work this paper cites.
Transforming neural-net output levels to probability distributions
J. S. Denker and Y. Lecun · 1991
Earlier work this paper cites.
Bayesian interpolation
D. J. MacKay · 1992
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
G. E. Hinton and D. Van Camp · 1993
Earlier work this paper cites.
Bayesian learning via stochastic dynamics
R. M. Neal · 1993
Earlier work this paper cites.
Ensemble learning in bayesian neural networks
D. Barber and C. M. Bishop · 1998
Earlier work this paper cites.
Optimal model inference for bayesian mixture of experts
N. Ueda and Z. Ghahramani · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu · 2002
Earlier work this paper cites.
N. de freitas, d
K. Barnard, P. Duygulu, and D. Forsyth · 2003
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
C.-Y. Lin · 2004
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
S. Banerjee and A. Lavie · 2005
Earlier work this paper cites.
Every picture tells a story: Generating sentences from images
A. Farhadi, M. Hejrati, M. A. Sadeghi, P. Young, C. Rashtchian, J. Hockenmaier, and D. Forsyth · 2010
Earlier work this paper cites.
Practical variational inference for neural networks
A. Graves · 2011
Earlier work this paper cites.
Baby talk: Understanding and generating image descriptions
G. Kulkarni, V. Premraj, S. Dhar, S. Li, Y. Choi, A. C. Berg, and T. L. Berg · 2011
Earlier work this paper cites.
Bayesian learning for neural networks
R. M. Neal · 2012
Earlier work this paper cites.
Twenty years of mixture of experts
S. E. Yuksel, J. N. Wilson, and P. D. Gader · 2012
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
A multi-world approach to question answering about real-world scenes based on uncertain input
M. Malinowski and M. Fritz · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Grounded compositional semantics for finding and describing images with sentences
R. Socher, A. Karpathy, Q. V. Le, C. D. Manning, and A. Y. Ng · 2014
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Cited alongside, same era.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Cited alongside, same era.
Weight uncertainty in neural networks
C. Blundell, J. Cornebise, K. Kavukcuoglu, and D. Wierstra · 2015
Cited alongside, same era.
Mind’s eye: A recurrent visual representation for image caption generation
X. Chen and C. Lawrence Zitnick · 2015
Cited alongside, same era.
From captions to visual concepts and back
H. Fang, S. Gupta, F. Iandola, R. Srivastava, L. Deng, P. Dollár, J. Gao, X. He, M. Mitchell, J. Platt, et al · 2015
Generating natural questions about an image
N. Mostafazadeh, I. Misra, J. Devlin, M. Mitchell, X. He, and L. Vanderwende · 2016
Later among the works it cites.
Image question answering using convolutional neural network with dynamic parameter prediction
H. Noh, P. Hongsuck Seo, and B. Han · 2016
Later among the works it cites.
X-cnn: Cross-modal convolutional neural networks for sparse datasets
P. Veličković, D. Wang, N. D. Lane, and P. Liò · 2016
Later among the works it cites.
Diverse beam search: Decoding diverse solutions from neural sequence models
A. K. Vijayakumar, M. Cogswell, R. R. Selvaraju, Q. Sun, S. Lee, D. Crandall, and D. Batra · 2016
Later among the works it cites.
Attribute2image: Conditional image generation from visual attributes
X. Yan, J. Yang, K. Sohn, and H. Lee · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Bayesian convolutional neural networks with bernoulli approximate variational inference
Y. Gal and Z. Ghahramani · 2015
Cited alongside, same era.
Deep visual-semantic alignments for generating image descriptions
A. Karpathy and L. Fei-Fei · 2015
Cited alongside, same era.
A. Kendall, V. Badrinarayanan, and R. Cipolla · 2015
Cited alongside, same era.
Exploring models and data for image question answering
M. Ren, R. Kiros, and R. Zemel · 2015
Cited alongside, same era.
Cider: Consensus-based image description evaluation
R. Vedantam, L. Zitnick, and D. Parikh · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Cited alongside, same era.
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola · 2016
Later among the works it cites.
Colorful image colorization
R. Zhang, P. Isola, and A. A. Efros · 2016
Later among the works it cites.
Visual7w: Grounded question answering in images
Y. Zhu, O. Groth, M. Bernstein, and L. Fei-Fei · 2016
Later among the works it cites.
Cross-modal scene networks
Y. Aytar, L. Castrejon, C. Vondrick, H. Pirsiavash, and A. Torralba · 2017
Later among the works it cites.
Evaluating visual conversational agents via cooperative human-ai games
P. Chattopadhyay, D. Yadav, V. Prabhu, A. Chandrasekaran, A. Das, S. Lee, D. Batra, and D. Parikh · 2017
Later among the works it cites.
Visual Dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. Moura, D. Parikh, and D. Batra · 2017
Later among the works it cites.
What’s in a question: Using visual questions as a form of supervision
S. Ganju, O. Russakovsky, and A. Gupta · 2017
Later among the works it cites.
Creativity: Generating diverse questions using variational autoencoders
U. Jain, Z. Zhang, and A. G. Schwing · 2017
Later among the works it cites.
Hadamard Product for Low-rank Bilinear Pooling
J.-H. Kim, K. W. On, W. Lim, J. Kim, J.-W. Ha, and B.-T. Zhang · 2017
Later among the works it cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, et al · 2017
Later among the works it cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra · 2017
Later among the works it cites.
Places: A 10 million image database for scene recognition
B. Zhou, A. Lapedriza, A. Khosla, A. Oliva, and A. Torralba · 2017
Later among the works it cites.
Multi-task learning using uncertainty to weigh losses for scene geometry and semantics
A. Kendall, Y. Gal, and R. Cipolla · 2018
Later among the works it cites.
Predictive uncertainty estimation via prior networks
A. Malinin and M. Gales · 2018
Later among the works it cites.
Multimodal differential network for visual question generation
B. N. Patro, S. Kumar, V. K. Kurmi, and V. Namboodiri · 2018
Later among the works it cites.
Attending to discriminative certainty for domain adaptation
V. K. Kurmi, S. Kumar, and V. P. Namboodiri · 2019
Later among the works it cites.
U-cam: Visual explanation using uncertainty based class activation maps
B. N. Patro, M. Lunayach, S. Patel, and V. P. Namboodiri · 2019
Later among the works it cites.