Fetching the paper…
Reading the bibliography…
Although deep convolutional networks have achieved improved performance in many natural language tasks, they have been treated as black boxes because they are difficult to interpret.
Europarl: A Parallel Corpus for Statistical Machine Translation
Philipp Koehn · 2005
Earlier work this paper cites.
Visualizing Higher-layer Features of a Deep Network
Dumitru Erhan, Yoshua Bengio, Aaron Courville, and Pascal Vincent · 2009
Earlier work this paper cites.
Modern Hierarchical, Agglomerative Clustering Algorithms
Daniel Müllner · 2011
Earlier work this paper cites.
Handling the Impact of Low Frequency Events on Co-occurrence based Measures of Word Similarity
François Role and Mohamed Nadif · 2011
Earlier work this paper cites.
Parallel Data, Tools and Interfaces in OPUS
Jörg Tiedemann · 2012
Earlier work this paper cites.
Deep inside Convolutional Networks: Visualising Image Classification Models and Saliency maps
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman · 2013
Earlier work this paper cites.
Morfessor 2.0: Python Implementation and Extensions for Morfessor Baseline
Sami Virpioja, Peter Smit, Stig-Arne Gronroos, and Mikko Kurimo · 2013
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart van Merrienboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Glove: Global Vectors for Word Representation
Jeffrey Pennington, Richard Socher, and Christopher Manning · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Earlier work this paper cites.
TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems, 2015
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dandelion Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng · 2015
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Visualizing and Understanding Recurrent Networks
Andrej Karpathy, Justin Johnson, and Li Fei-Fei · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2015
Cited alongside, same era.
DBpedia–a Large-Scale, Multilingual Knowledge Base Extracted from Wikipedia
Jens Lehmann, Robert Isele, Max Jakob, Anja Jentzsch, Dimitris Kontokostas, Pablo N Mendes, Sebastian Hellmann, Mohamed Morsey, Patrick Van Kleef, Sören Auer, et al · 2015
Cited alongside, same era.
Character-level Convolutional Networks for Text Classification
Xiang Zhang, Junbo Zhao, and Yann LeCun · 2015
Cited alongside, same era.
Object Detectors Emerge in Deep Scene CNNs
Bolei Zhou, Aditya Khosla, Agata Lapedriza, Aude Oliva, and Antonio Torralba · 2015
Cited alongside, same era.
Neural Machine Translation in Linear Time
Nal Kalchbrenner, Lasse Espeholt, Karen Simonyan, Aaron van den Oord, Alex Graves, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Character-Aware Neural Language Models
Feature Visualization
Chris Olah, Alexander Mordvintsev, and Ludwig Schubert · 2017
Later among the works it cites.
Learning to Generate Reviews and Discovering Sentiment
Alec Radford, Rafal Jozefowicz, and Ilya Sutskever · 2017
Later among the works it cites.
ConceptNet 5.5: An Open Multilingual Graph of General Knowledge
Robert Speer, Joshua Chin, and Catherine Havasi · 2017
Later among the works it cites.
Memory Visualization for Gated Recurrent Neural Networks in Speech Recognition
Zhiyuan Tang, Ying Shi, Dong Wang, Yang Feng, and Shiyue Zhang · 2017
Later among the works it cites.
Attention is All You Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Llion Jones, Jakob Uszkoreit, Aidan N Gomez, and Ł ukasz Kaiser · 2017
Later among the works it cites.
What You Can Cram into a Single \ $ & ! # ∗ {\backslash}{\$}{\&}!{\#}* Vector: Probing Sentence Embeddings for Linguistic Properties
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yoon Kim, Yacine Jernite, David Sontag, and Alexander M Rush · 2016
Cited alongside, same era.
How Transferable are Neural Networks in NLP Applications?
Lili Mou, Zhao Meng, Rui Yan, Ge Li, Yan Xu, Lu Zhang, and Zhi Jin · 2016
Cited alongside, same era.
Analyzing Linguistic Knowledge in Sequential Model of Sentence
Peng Qian, Xipeng Qiu, and Xuanjing Huang · 2016
Cited alongside, same era.
Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg · 2017
Cited alongside, same era.
Network Dissection: Quantifying Interpretability of Deep Visual Representations
David Bau, Bolei Zhou, Aditya Khosla, Aude Oliva, and Antonio Torralba · 2017
Cited alongside, same era.
What do Neural Machine Translation Models Learn about Morphology?
Yonatan Belinkov, Nadir Durrani, Fahim Dalvi, Hassan Sajjad, and James Glass · 2017
Cited alongside, same era.
Enriching Word Vectors with Subword Information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov · 2017
Cited alongside, same era.
Alexis Conneau, Germán Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni · 2018
Later among the works it cites.
Net2Vec: Quantifying and Explaining how Concepts are Encoded by Filters in Deep Neural Networks
Ruth Fong and Andrea Vedaldi · 2018
Later among the works it cites.
Constituency Parsing with a Self-Attentive Encoder
Nikita Kitaev and Dan Klein · 2018
Later among the works it cites.
On the Importance of Single Directions for Generalization
Ari S. Morcos, David G.T. Barrett, Neil C. Rabinowitz, and Matthew Botvinick · 2018
Later among the works it cites.
The Building Blocks of Interpretability
Chris Olah, Arvind Satyanarayan, Ian Johnson, Shan Carter, Ludwig Schubert, Katherine Ye, and Alexander Mordvintsev · 2018
Later among the works it cites.
Exploring Semantic Properties of Sentence Embeddings
Xunjie Zhu, Tingfeng Li, and Gerard Melo · 2018
Later among the works it cites.
Visualizing and Understanding Generative Adversarial Networks
David Bau, Jun-Yan Zhu, Hendrik Strobelt, Bolei Zhou, Joshua B. Tenenbaum, William T. Freeman, and Antonio Torralba · 2019
Closest in time.