Fetching the paper…
Reading the bibliography…
Neural networks have recently achieved human-level performance on various challenging natural language processing (NLP) tasks, but it is notoriously difficult to understand why a neural network produced a particular prediction.
Multitask learning
Caruana, R · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Earlier work this paper cites.
The pascal recognising textual entailment challenge
Dagan, I., Glickman, O., and Magnini, B · 2005
Earlier work this paper cites.
Modeling annotators: A generative approach to learning from annotator rationales
Zaidan, O. and Eisner, J · 2008
Earlier work this paper cites.
How to explain individual classification decisions
Baehrens, D., Schroeter, T., Harmeling, S., Kawanabe, M., Hansen, K., and MÞller, K.-R · 2010
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Maas, A. L., Daly, R. E., Pham, P. T., Huang, D., Ng, A. Y., and Potts, C · 2011
Earlier work this paper cites.
The winograd schema challenge
Levesque, H., Davis, E., and Morgenstern, L · 2012
Earlier work this paper cites.
Generating sequences with recurrent neural networks
Graves, A · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y · 2014
Earlier work this paper cites.
A convolutional neural network for modelling sentences
Kalchbrenner, N., Grefenstette, E., and Blunsom, P · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q. V · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Bowman, S. R., Angeli, G., Potts, C., and Manning, C. D · 2015
Earlier work this paper cites.
Image-based recommendations on styles and substitutes
McAuley, J., Targett, C., Shi, Q., and Van Den Hengel, A · 2015
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A · 2015
Earlier work this paper cites.
Show, attend and tell: Neural image caption generation with visual attention
Xu, K., Ba, J., Kiros, R., Cho, K., Courville, A., Salakhudinov, R., Zemel, R., and Bengio, Y · 2015
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering
He, R. and McAuley, J · 2016
Earlier work this paper cites.
Synthetic and natural noise both break neural machine translation
Belinkov, Y. and Bisk, Y · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Doshi-Velez, F. and Kim, B · 2017
Earlier work this paper cites.
Input switched affine networks: an rnn architecture designed for interpretability
Foerster, J. N., Gilmer, J., Sohl-Dickstein, J., Chorowski, J., and Sussillo, D · 2017
Cited alongside, same era.
Adversarial examples for evaluating reading comprehension systems
Jia, R. and Liang, P · 2017
Cited alongside, same era.
Online and linear-time attention by enforcing monotonic alignments
Raffel, C., Luong, M.-T., Liu, P. J., Weiss, R. J., and Eck, D · 2017
Cited alongside, same era.
An overview of multi-task learning in deep neural networks
Ruder, S · 2017
Cited alongside, same era.
Smoothgrad: removing noise by adding noise
Smilkov, D., Thorat, N., Kim, B., Viégas, F., and Wattenberg, M · 2017
Cited alongside, same era.
Kudo, T. and Richardson, J · 2018
Later among the works it cites.
Hallucinations in neural machine translation
Lee, K., Firat, O., Agarwal, A., Fannjiang, C., and Sussillo, D · 2018
Later among the works it cites.
Deep contextualized word representations
Peters, M. E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., and Zettlemoyer, L · 2018
Later among the works it cites.
A call for clarity in reporting bleu scores
Post, M · 2018
Later among the works it cites.
Adafactor: Adaptive learning rates with sublinear memory cost
Shazeer, N. and Stern, M · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Axiomatic attribution for deep networks
Sundararajan, M., Taly, A., and Yan, Q · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Williams, A., Nangia, N., and Bowman, S. R · 2017
Cited alongside, same era.
e-snli: Natural language inference with natural language explanations
Camburu, O.-M., Rocktäschel, T., Lukasiewicz, T., and Blunsom, P · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2018
Cited alongside, same era.
A survey of methods for explaining black box models
Guidotti, R., Monreale, A., Ruggieri, S., Turini, F., Giannotti, F., and Pedreschi, D · 2018
Cited alongside, same era.
Evaluating feature importance estimates
Hooker, S., Erhan, D., Kindermans, P.-J., and Kim, B · 2018
Cited alongside, same era.
Later among the works it cites.
Visual interpretability for deep learning: a survey
Zhang, Q. and Zhu, S · 2018
Later among the works it cites.
Eraser: A benchmark to evaluate rationalized nlp models
DeYoung, J., Jain, S., Rajani, N. F., Lehman, E., Xiong, C., Socher, R., and Wallace, B. C · 2019
Later among the works it cites.
Jain, S. and Wallace, B. C · 2019
Later among the works it cites.
The (un) reliability of saliency methods
Kindermans, P.-J., Hooker, S., Adebayo, J., Alber, M., Schütt, K. T., Dähne, S., Erhan, D., and Kim, B · 2019
Later among the works it cites.
ALBERT: A lite bert for self-supervised learning of language representations
Lan, Z., Chen, M., Goodman, S., Gimpel, K., Sharma, P., and Soricut, R · 2019
Later among the works it cites.
Multi-task deep neural networks for natural language understanding
Liu, X., He, P., Chen, W., and Gao, J · 2019
Later among the works it cites.
Interpretable machine learning
Molnar, C · 2019
Later among the works it cites.
Adversarial nli: A new benchmark for natural language understanding
Nie, Y., Williams, A., Dinan, E., Bansal, M., Weston, J., and Kiela, D · 2019
Later among the works it cites.
Learning to deceive with attention-based explanations
Pruthi, D., Gupta, M., Dhingra, B., Neubig, G., and Lipton, Z. C · 2019
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2019
Later among the works it cites.
Explain yourself! leveraging language models for commonsense reasoning
Rajani, N. F., McCann, B., Xiong, C., and Socher, R · 2019
Later among the works it cites.
Serrano, S. and Smith, N. A · 2019
Later among the works it cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Wang, A., Pruksachatkun, Y., Nangia, N., Singh, A., Michael, J., Hill, F., Levy, O., and Bowman, S · 2019
Later among the works it cites.