Fetching the paper…
Reading the bibliography…
Neural Sequence-to-Sequence models have proven to be accurate and robust for many sequence prediction tasks, and have become the standard approach for automatic translation of text.
Nonmetric multidimensional scaling: a numerical method
J. B. Kruskal · 1964
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Visualizing data using t-sne
L. v. d. Maaten and G. Hinton · 2008
Earlier work this paper cites.
Visualizing higher-layer features of a deep network
D. Erhan, Y. Bengio, A. Courville, and P. Vincent · 2009
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
Building high-level features using large scale unsupervised learning
Q. V. Le, M. Ranzato, R. Monga, M. Devin, G. Corrado, K. C. 0010, J. Dean, and A. Y. Ng · 2012
Earlier work this paper cites.
Wit3: Web inventory of transcribed and translated talks
C. Mauro, G. Christian, and F. Marcello · 2012
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
K. Simonyan, A. Vedaldi, and A. Zisserman · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Earlier work this paper cites.
Visualizing and Understanding Convolutional Networks
M. D. Zeiler and R. Fergus · 2014
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
S. Bach, A. Binder, G. Montavon, F. Klauschen, K.-R. Müller, and W. Samek · 2015
Earlier work this paper cites.
Visualizing and understanding recurrent networks
A. Karpathy, J. Johnson, and F.-F. Li · 2015
Earlier work this paper cites.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images
A. Nguyen, J. Yosinski, and J. Clune · 2015
Earlier work this paper cites.
A neural attention model for abstractive sentence summarization
A. M. Rush, S. Chopra, and J. Weston · 2015
Earlier work this paper cites.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio · 2015
Earlier work this paper cites.
Understanding neural networks through deep visualization
J. Yosinski, J. Clune, A. Nguyen, T. Fuchs, and H. Lipson · 2015
Earlier work this paper cites.
A Convolutional Encoder Model for Neural Machine Translation
J. Gehring, M. Auli, D. Grangier, and Y. N. Dauphin · 2016
Earlier work this paper cites.
Google’s multilingual neural machine translation system: enabling zero-shot translation
M. Johnson, M. Schuster, Q. V. Le, M. Krikun, Y. Wu, Z. Chen, N. Thorat, F. Viégas, M. Wattenberg, G. Corrado, et al · 2016
Cited alongside, same era.
Interacting with predictions: Visual inspection of black-box machine learning models
J. Krause, A. Perer, and K. Ng · 2016
Cited alongside, same era.
Rationalizing neural predictions
T. Lei, R. Barzilay, and T. S. Jaakkola · 2016
Cited alongside, same era.
Visualizing and Understanding Neural Models in NLP
J. Li, X. Chen, E. Hovy, and D. Jurafsky · 2016
Cited alongside, same era.
Understanding neural networks through representation erasure
J. Li, W. Monroe, and D. Jurafsky · 2016
Cited alongside, same era.
Understanding black-box predictions via influence functions
P. W. Koh and P. Liang · 2017
Later among the works it cites.
Interactive visualization and manipulation of attention-based neural machine translation
J. Lee, J.-H. Shin, and J.-S. Kim · 2017
Later among the works it cites.
Understanding hidden memories of recurrent neural networks
Y. Ming, S. Cao, R. Zhang, Z. Li, Y. Chen, Y. Song, and H. Qu · 2017
Later among the works it cites.
A deep reinforced model for abstractive summarization
R. Paulus, C. Xiong, and R. Socher · 2017
Later among the works it cites.
Right for the right reasons: Training differentiable models by constraining their explanations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Nallapati, B. Zhou, C. Gulcehre, B. Xiang, et al · 2016
Cited alongside, same era.
Attention and augmented recurrent neural networks
C. Olah and S. Carter · 2016
Cited alongside, same era.
Why should i trust you?: Explaining the predictions of any classifier
M. T. Ribeiro, S. Singh, and C. Guestrin · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Y. Wu, M. Schuster, Z. Chen, Q. V. Le, M. Norouzi, W. Macherey, M. Krikun, Y. Cao, Q. Gao, K. Macherey, et al · 2016
Cited alongside, same era.
A causal framework for explaining the predictions of black-box sequence-to-sequence models
D. Alvarez-Melis and T. S. Jaakkola · 2017
Cited alongside, same era.
Explanation and justification in machine learning: A survey
O. Biran and C. Cotton · 2017
Cited alongside, same era.
Rnnbow: Visualizing learning via backpropagation gradients in recurrent neural networks
D. Cashman, G. Patterson, A. Mosca, and R. Chang · 2017
Cited alongside, same era.
A. S. Ross, M. C. Hughes, and F. Doshi-Velez · 2017
Later among the works it cites.
End-to-end non-factoid question answering with an interactive visualization of neural attention weights
A. Rücklé and I. Gurevych · 2017
Later among the works it cites.
Get to the point: Summarization with pointer-generator networks
A. See, P. J. Liu, and C. D. Manning · 2017
Later among the works it cites.
Direct-manipulation visualization of deep networks
D. Smilkov, S. Carter, D. Sculley, F. B. Viégas, and M. Wattenberg · 2017
Later among the works it cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Later among the works it cites.
Visualizing deep neural network decisions: Prediction difference analysis
L. M. Zintgraf, T. S. Cohen, T. Adel, and M. Welling · 2017
Later among the works it cites.
Achieving human parity on automatic chinese to english news translation
H. Hassan Awadalla, A. Aue, C. Chen, V. Chowdhary, J. Clark, C. Federmann, X. Huang, M. Junczys-Dowmunt, W. Lewis, M. Li, S. Liu, T.-Y. Liu, R. Luo, A. Menezes, T. Qin, F. Seide, X. Tan, F. Tian, L. Wu, S. Wu, Y. Xia, D. Zhang, Z. Zhang, and M. Zhou · 2018
Closest in time.
Visual analytics in deep learning: An interrogative survey for the next frontiers
F. Hohman, M. Kahng, R. Pienta, and D. H. Chau · 2018
Closest in time.
Activis: Visual exploration of industry-scale deep neural network models
M. Kahng, P. Y. Andrews, A. Kalro, and D. H. P. Chau · 2018
Closest in time.
Generating wikipedia by summarizing long sequences
P. J. Liu, M. Saleh, E. Pot, B. Goodrich, R. Sepassi, L. Kaiser, and N. Shazeer · 2018
Closest in time.
M. Narayanan, E. Chen, J. He, B. Kim, S. Gershman, and F. Doshi-Velez · 2018
Closest in time.
The building blocks of interpretability
C. Olah, A. Satyanarayan, I. Johnson, S. Carter, L. Schubert, K. Ye, and A. Mordvintsev · 2018
Closest in time.
Conceptvector: text visual analytics via interactive lexicon building using word embedding
D. Park, S. Kim, J. Lee, J. Choo, N. Diakopoulos, and N. Elmqvist · 2018
Closest in time.
Lstmvis: A tool for visual analysis of hidden state dynamics in recurrent neural networks
H. Strobelt, S. Gehrmann, H. Pfister, and A. M. Rush · 2018
Closest in time.