Fetching the paper…
Reading the bibliography…
We propose a technique for producing "visual explanations" for decisions from a large class of CNN-based models, making them more transparent.
Introduction to Expert Systems
P. Jackson · 1998
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Visualizing Higher-layer Features of a Deep Network
D. Erhan, Y. Bengio, A. Courville, and P. Vincent · 2009
Earlier work this paper cites.
The PASCAL Visual Object Classes Challenge 2007 (VOC2007) Results
M. Everingham, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman · 2009
Earlier work this paper cites.
Diagnosing Error in Object Detectors
D. Hoiem, Y. Chodpathumwan, and Q. Dai · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Y. Bengio, A. Courville, and P. Vincent · 2013
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
K. Simonyan, A. Vedaldi, and A. Zisserman · 2013
Earlier work this paper cites.
HOGgles: Visualizing Object Detection Features
C. Vondrick, A. Khosla, T. Malisiewicz, and A. Torralba · 2013
Earlier work this paper cites.
Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
Caffe: Convolutional Architecture for Fast Feature Embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Earlier work this paper cites.
What I learned from competing against a ConvNet on ImageNet
A. Karpathy · 2014
Earlier work this paper cites.
Network in network
M. Lin, Q. Chen, and S. Yan · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Learning and transferring mid-level image representations using convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2014
Earlier work this paper cites.
Striving for Simplicity: The All Convolutional Net
J. T. Springenberg, A. Dosovitskiy, T. Brox, and M. A. Riedmiller · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2014
Earlier work this paper cites.
Object detectors emerge in deep scene cnns
B. Zhou, A. Khosla, À. Lapedriza, A. Oliva, and A. Torralba · 2014
Earlier work this paper cites.
CloudCV: Large Scale Distributed Computer Vision as a Cloud Service
H. Agrawal, C. S. Mathialagan, Y. Goyal, N. Chavali, P. Banik, A. Mohapatra, A. Osman, and D. Batra · 2015
Earlier work this paper cites.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. Lawrence Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Microsoft COCO captions: Data Collection and Evaluation Server
X. Chen, H. Fang, T.-Y. Lin, R. Vedantam, S. Gupta, P. Dollár, and C. L. Zitnick · 2015
Earlier work this paper cites.
Inverting Convolutional Networks with Convolutional Networks
A. Dosovitskiy and T. Brox · 2015
Cited alongside, same era.
From Captions to Visual Concepts and Back
H. Fang, S. Gupta, F. Iandola, R. K. Srivastava, L. Deng, P. Dollár, J. Gao, X. He, M. Mitchell, J. C. Platt, et al · 2015
Cited alongside, same era.
Devnet: A deep event network for multimedia event detection and evidence recounting
C. Gan, N. Wang, Y. Yang, D.-Y. Yeung, and A. G. Hauptmann · 2015
Cited alongside, same era.
Are You Talking to a Machine? Dataset and Methods for Multilingual Image Question Answering
H. Gao, J. Mao, J. Zhou, Z. Huang, L. Wang, and W. Xu · 2015
Cited alongside, same era.
Explaining and harnessing adversarial examples
I. J. Goodfellow, J. Shlens, and C. Szegedy · 2015
Cited alongside, same era.
Becoming the Expert - Interactive Multi-Class Machine Teaching
E. Johns, O. Mac Aodha, and G. J. Brostow · 2015
DenseCap: Fully Convolutional Localization Networks for Dense Captioning
J. Johnson, A. Karpathy, and L. Fei-Fei · 2016
Closest in time.
Seed, expand and constrain: Three principles for weakly-supervised image segmentation
A. Kolesnikov and C. H. Lampert · 2016
Closest in time.
The Mythos of Model Interpretability
Z. C. Lipton · 2016
Closest in time.
Hierarchical question-image co-attention for visual question answering
J. Lu, J. Yang, D. Batra, and D. Parikh · 2016
Closest in time.
Salient deconvolutional networks
A. Mahendran and A. Vedaldi · 2016
Closest in time.
Visualizing deep convolutional neural networks using natural pre-images
A. Mahendran and A. Vedaldi · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep visual-semantic alignments for generating image descriptions
A. Karpathy and L. Fei-Fei · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Deeper LSTM and normalized CNN Visual Question Answering model
J. Lu, X. Lin, D. Batra, and D. Parikh · 2015
Cited alongside, same era.
Ask your neurons: A neural-based approach to answering questions about images
M. Malinowski, M. Rohrbach, and M. Fritz · 2015
Cited alongside, same era.
Is object localization for free? – weakly-supervised learning with convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2015
Cited alongside, same era.
From image-level to pixel-level labeling with convolutional networks
P. O. Pinheiro and R. Collobert · 2015
Cited alongside, same era.
"Why Should I Trust You?": Explaining the Predictions of Any Classifier
M. T. Ribeiro, S. Singh, and C. Guestrin · 2016
Closest in time.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Closest in time.
Rethinking the inception architecture for computer vision
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna · 2016
Closest in time.
Top-down Neural Attention by Excitation Backprop
J. Zhang, Z. Lin, J. Brandt, X. Shen, and S. Sclaroff · 2016
Closest in time.
Learning Deep Features for Discriminative Localization
B. Zhou, A. Khosla, L. A., A. Oliva, and A. Torralba · 2016
Closest in time.
Network dissection: Quantifying interpretability of deep visual representations
D. Bau, B. Zhou, A. Khosla, A. Oliva, and A. Torralba · 2017
Closest in time.
Visual Dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. Moura, D. Parikh, and D. Batra · 2017
Closest in time.
Learning cooperative visual dialog agents with deep reinforcement learning
A. Das, S. Kottur, J. M. Moura, S. Lee, and D. Batra · 2017
Closest in time.
Guesswhat?! visual object discovery through multi-modal dialogue
H. de Vries, F. Strub, S. Chandar, O. Pietquin, H. Larochelle, and A. C. Courville · 2017
Closest in time.
Iqa: Visual question answering in interactive environments
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi · 2017
Closest in time.
Places: A 10 million image database for scene recognition
B. Zhou, A. Lapedriza, A. Khosla, A. Oliva, and A. Torralba · 2017
Closest in time.
Embodied Question Answering
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra · 2018
Closest in time.
Choose your neuron: Incorporating domain knowledge through neuron-importance
R. R. Selvaraju, P. Chattopadhyay, M. Elhoseiny, T. Sharma, D. Batra, D. Parikh, and S. Lee · 2018
Closest in time.
Taking a hint: Leveraging explanations to make vision and language models more grounded
R. R. Selvaraju, S. Lee, Y. Shen, H. Jin, S. Ghosh, L. Heck, D. Batra, and D. Parikh · 2019
Closest in time.