Fetching the paper…
Reading the bibliography…
Bidirectional Encoder Representations from Transformers (BERT) reach state-of-the-art results in a variety of Natural Language Processing tasks.
LIII. On lines and planes of closest fit to systems of points in space
Karl Pearson F.R.S. 1901 · 1901
Earlier work this paper cites.
Least squares quantization in PCM
Stuart P. Lloyd. 1982 · 1982
Earlier work this paper cites.
Independent component analysis, A new concept?
Pierre Comon. 1994 · 1994
Earlier work this paper cites.
Zadeh, L.A.: Toward a Theory of Fuzzy Information Granulation and Its Centrality in Human Reasoning and Fuzzy Logic. Fuzzy Sets and Systems
Lotfi A Zadeh. 1997 · 1997
Earlier work this paper cites.
Overview of TREC 2001. In Proceedings of TREC 2001
Ellen Voorhees. 2001 · 2001
Earlier work this paper cites.
Learning Question Classifiers. In Proceedings of COLING 2002
Xin Li and Dan Roth. 2002 · 2002
Earlier work this paper cites.
Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition, Chapter 23
Dan Jurafsky and James H. Martin. 2009 · 2009
Earlier work this paper cites.
Learning a Parametric Embedding by Preserving Local Structure. In Proceedings of AISTATS 2009
Laurens van der Maaten. 2009 · 2009
Earlier work this paper cites.
OntoNotes: A Large Training Corpus for Enhanced Processing
Ralph Weischedel, Eduard Hovy, Mitchell Marcus, Martha Palmer, Robert Belvin, Sameer Pradhan, Lance Ramshaw, and Nianwen Xue. 2011 · 2011
Earlier work this paper cites.
Efficient Estimation of Word Representations in Vector Space. In Workshop Track Proceedings of ICLR 2013
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Semi-supervised Sequence Learning. In Proceedings of NIPS 2015
Andrew M. Dai and Quoc V. Le. 2015 · 2015
Earlier work this paper cites.
Exploring How Deep Neural Networks Form Phonemic Categories. In Proceedings of INTERSPEECH 2015
Tasha Nagamine, Michael Seltzer, and Nima Mesgarani. 2015 · 2015
Earlier work this paper cites.
Understanding Neural Networks through Representation Erasure
Jiwei Li, Will Monroe, and Dan Jurafsky. 2016 · 2016
Earlier work this paper cites.
The Mythos of Model Interpretability
Zachary Chase Lipton. 2016 · 2016
Earlier work this paper cites.
SQuAD: 100, 000+ Questions for Machine Comprehension of Text. In Proceedings of EMNLP 2016
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Does String-Based Neural MT Learn Source Syntax?. In Proceedings of EMNLP 2016
Xing Shi, Inkit Padhi, and Kevin Knight. 2016 · 2016
Cited alongside, same era.
Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks. In Proceedings of ICLR 2016
Jason Weston, Antoine Bordes, Sumit Chopra, and Tomas Mikolov. 2016 · 2016
Cited alongside, same era.
What do Neural Machine Translation Models Learn about Morphology?. In Proceedings of ACL 2017
Yonatan Belinkov, Nadir Durrani, Fahim Dalvi, Hassan Sajjad, and James R. Glass. 2017 · 2017
Cited alongside, same era.
Bidirectional Attention Flow for Machine Comprehension. In Proceedings of ICLR 2017
Min Joon Seo, Aniruddha Kembhavi, Ali Farhadi, and Hannaneh Hajishirzi. [n. d.] · 2017
Cited alongside, same era.
The Natural Language Decathlon: Multitask Learning as Question Answering
Bryan McCann, Nitish Shirish Keskar, Caiming Xiong, and Richard Socher. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations. In Proceedings of NAACL-HLT 2018
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Improving Language Understanding by Generative Pre-Training
Alec Radford. 2018 · 2018
Later among the works it cites.
Know What You Don’t Know: Unanswerable Questions for SQuAD. In Proceedings of ACL 2018
Pranav Rajpurkar, Robin Jia, and Percy Liang. 2018 · 2018
Later among the works it cites.
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering. In Proceedings of EMNLP 2018
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W. Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Attention Is All You Need. In Proceedings of NIPS 2017
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
SentEval: An Evaluation Toolkit for Universal Sentence Representations. In Proceedings of LREC 2018
Alexis Conneau and Douwe Kiela. 2018 · 2018
Cited alongside, same era.
Universal Transformers. In Proceedings of SMACD 2018
Mostafa Dehghani, Stephan Gouws, Oriol Vinyals, Jakob Uszkoreit, and Lukasz Kaiser. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Explainable artificial intelligence: A survey. In MIPRO 2018
F. K. Došilović, M. Brčić, and N. Hlupić. 2018 · 2018
Cited alongside, same era.
A Survey Of Methods For Explaining Black Box Models
Riccardo Guidotti, Anna Monreale, Franco Turini, Dino Pedreschi, and Fosca Giannotti. 2018 · 2018
Cited alongside, same era.
Fine-tuned Language Models for Text Classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.
Visual interpretability for deep learning: a survey
Quan-shi Zhang and Song-chun Zhu. 2018 · 2018
Later among the works it cites.
Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime G. Carbonell, Quoc V. Le, and Ruslan Salakhutdinov. 2019 · 2019
Closest in time.
Assessing BERT’s Syntactic Abilities
Yoav Goldberg. 2019 · 2019
Closest in time.
Attention is not Explanation. In Proceedings of NAACL 2019
Sarthak Jain and Byron C. Wallace. 2019 · 2019
Closest in time.
Linguistic Knowledge and Transferability of Contextual Representations. In Proceedings of NAACL 2019
Nelson F. Liu, Matt Gardner, Yonatan Belinkov, Matthew Peters, and Noah A. Smith. 2019 · 2019
Closest in time.
Understanding the Behaviors of BERT in Ranking
Yifan Qiao, Chenyan Xiong, Zheng-Hao Liu, and Zhiyuan Liu. 2019 · 2019
Closest in time.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Closest in time.
What do you learn from context? Probing for sentence structure in contextualized word representations. In Proceedings of ICLR 2019
Ian Tenney, Patrick Xia, Berlin Chen, Alex Wang, Adam Poliak, R Thomas McCoy, Najoung Kim, Benjamin Van Durme, Sam Bowman, Dipanjan Das, and Ellie Pavlick. 2019 · 2019
Closest in time.