Fetching the paper…
Reading the bibliography…
For stability and reliability of real-world applications, the robustness of DNNs in unimodal tasks has been evaluated.
Selective question answering under domain shift
Kamath, A.; Jia, R.; and Liang, P. 2020 · 2006
Earlier work this paper cites.
Visualizing data using t-SNE
Maaten, L. v. d.; and Hinton, G. 2008 · 2008
Earlier work this paper cites.
Active learning literature survey
Settles, B. 2009 · 2009
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Goodfellow, I. J.; Shlens, J.; and Szegedy, C. 2014 · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y.; Maire, M.; Belongie, S.; Hays, J.; Perona, P.; Ramanan, D.; Dollár, P.; and Zitnick, C. L. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Pennington, J.; Socher, R.; and Manning, C. 2014 · 2014
Earlier work this paper cites.
Vqa: Visual question answering
Antol, S.; Agrawal, A.; Lu, J.; Mitchell, M.; Batra, D.; Lawrence Zitnick, C.; and Parikh, D. 2015 · 2015
Earlier work this paper cites.
Deep visual-semantic alignments for generating image descriptions
Karpathy, A.; and Fei-Fei, L. 2015 · 2015
Earlier work this paper cites.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images
Nguyen, A.; Yosinski, J.; and Clune, J. 2015 · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Ren, S.; He, K.; Girshick, R.; and Sun, J. 2015 · 2015
Earlier work this paper cites.
Multimodal compact bilinear pooling for visual question answering and visual grounding
Fukui, A.; Park, D. H.; Yang, D.; Rohrbach, A.; Darrell, T.; and Rohrbach, M. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Question Relevance in VQA: Identifying Non-Visual And False-Premise Questions
Ray, A.; Christie, G.; Bansal, M.; Batra, D.; and Parikh, D. 2016 · 2016
Earlier work this paper cites.
Making the V in VQA matter: Elevating the role of image understanding in Visual Question Answering
Goyal, Y.; Khot, T.; Summers-Stay, D.; Batra, D.; and Parikh, D. 2017 · 2017
Cited alongside, same era.
On calibration of modern neural networks
Guo, C.; Pleiss, G.; Sun, Y.; and Weinberger, K. Q. 2017 · 2017
Cited alongside, same era.
A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks
Hendrycks, D.; and Gimpel, K. 2017 · 2017
Cited alongside, same era.
The Promise of Premise: Harnessing Question Premises in Visual Question Answering
Mahendru, A.; Prabhu, V.; Mohapatra, A.; Batra, D.; and Lee, S. 2017 · 2017
Cited alongside, same era.
PixelCNN++: A PixelCNN Implementation with Discretized Logistic Mixture Likelihood and Other Modifications
Salimans, T.; Karpathy, A.; Chen, X.; and Kingma, D. P. 2017 · 2017
Cited alongside, same era.
A simple unified framework for detecting out-of-distribution samples and adversarial attacks
Lee, K.; Lee, K.; Lee, H.; and Shin, J. 2018 · 2018
Later among the works it cites.
Enhancing The Reliability of Out-of-distribution Image Detection in Neural Networks
Liang, S.; Li, Y.; and Srikant, R. 2018 · 2018
Later among the works it cites.
Beyond bilinear: Generalized multimodal factorized high-order pooling for visual question answering
Yu, Z.; Yu, J.; Xiang, C.; Fan, J.; and Tao, D. 2018 · 2018
Later among the works it cites.
Why Does a Visual Question Have Different Answers?
Bhattacharya, N.; Li, Q.; and Gurari, D. 2019 · 2019
Later among the works it cites.
Why ReLU networks yield high-confidence predictions far away from the training data and how to mitigate the problem
Hein, M.; Andriushchenko, M.; and Bitterwolf, J. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multi-modal factorized bilinear pooling with co-attention learning for visual question answering
Yu, Z.; Yu, J.; Fan, J.; and Tao, D. 2017 · 2017
Cited alongside, same era.
Attention with sparsity regularization for neural machine translation and summarization
Zhang, J.; Zhao, Y.; Li, H.; and Zong, C. 2018 · 2017
Cited alongside, same era.
Bottom-up and top-down attention for image captioning and visual question answering
Anderson, P.; He, X.; Buehler, C.; Teney, D.; Johnson, M.; Gould, S.; and Zhang, L. 2018 · 2018
Cited alongside, same era.
Deep attention neural tensor network for visual question answering
Bai, Y.; Fu, J.; Zhao, T.; and Mei, T. 2018 · 2018
Cited alongside, same era.
Waic, but why? generative ensembles for robust anomaly detection
Choi, H.; Jang, E.; and Alemi, A. A. 2018 · 2018
Cited alongside, same era.
Vizwiz grand challenge: Answering visual questions from blind people
Gurari, D.; Li, Q.; Stangl, A. J.; Guo, A.; Lin, C.; Grauman, K.; Luo, J.; and Bigham, J. P. 2018 · 2018
Cited alongside, same era.
Bilinear Attention Networks
Kim, J.-H.; Jun, J.; and Zhang, B.-T. 2018 · 2018
Cited alongside, same era.
Hendrycks, D.; Mazeika, M.; and Dietterich, T. 2019 · 2019
Later among the works it cites.
Towards neural networks that provably know when they don’t know
Meinke, A.; and Hein, M. 2019 · 2019
Later among the works it cites.
Likelihood ratios for out-of-distribution detection
Ren, J.; Liu, P. J.; Fertig, E.; Snoek, J.; Poplin, R.; Depristo, M.; Dillon, J.; and Lakshminarayanan, B. 2019 · 2019
Later among the works it cites.
Towards VQA Models That Can Read
Singh, A.; Natarjan, V.; Shah, M.; Jiang, Y.; Chen, X.; Parikh, D.; and Rohrbach, M. 2019 · 2019
Later among the works it cites.
Deep modular co-attention networks for visual question answering
Yu, Z.; Yu, J.; Cui, Y.; Tao, D.; and Tian, Q. 2019 · 2019
Later among the works it cites.
https://www.bespecular.com
BeSpecular. 2020 · 2020
Closest in time.
Query-Driven Multi-Instance Learning
Hsu, Y.-C.; Hong, C.-Y.; Lee, M.-S.; and Liu, T.-L. 2020 · 2020
Closest in time.