Fetching the paper…
Reading the bibliography…
Language models trained on billions of tokens have recently led to unprecedented results on many NLP tasks.
Principia Mathematica
Alfred North Whitehead and Bertrand Russell. 1925–1927 · 1927
Earlier work this paper cites.
On computable numbers, with an application to the Entscheidungsproblem
Alan M. Turing. 1936 · 1936
Earlier work this paper cites.
Computing machinery and intelligence
Alan M. Turing. 1950 · 1950
Earlier work this paper cites.
Distributional structure
Zellig S. Harris. 1954 · 1954
Earlier work this paper cites.
Aspects of the Theory of Syntax , volume 11
Noam Chomsky. 1965 · 1965
Earlier work this paper cites.
Logic and conversation
Herbert P Grice. 1975 · 1975
Earlier work this paper cites.
Learning regular sets from queries and counterexamples
Dana Angluin. 1987 · 1987
Earlier work this paper cites.
Semantics in Generative Grammar
Irene Heim and Angelika Kratzer. 1998 · 1998
Earlier work this paper cites.
More fragments of language
Ian Pratt-Hartmann and Allan Third. 2006 · 2006
Earlier work this paper cites.
Three learnable models for the description of language
Alexander Clark. 2010 · 2010
Earlier work this paper cites.
Intensional semantics
Kai Von Fintel and Irene Heim. 2011 · 2011
Cited alongside, same era.
On our best behaviour
Hector Levesque. 2014 · 2014
Cited alongside, same era.
Elements of Formal Semantics: An Introduction to the Mathematical Theory of Meaning in Natural Language
Yoad Winter. 2016 · 2016
Cited alongside, same era.
Fine-grained analysis of sentence embeddings using auxiliary prediction tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2017 · 2017
Cited alongside, same era.
What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, German Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Visualisation and ’diagnostic classifiers’ reveal how recurrent and recursive neural networks process hierarchical structure (extended abstract)
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2019 · 2019
Later among the works it cites.
BERT rediscovers the classical NLP pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019 · 2019
Later among the works it cites.
Learning and evaluating general linguistic intelligence
Dani Yogatama, Cyprien de Masson d’Autume, Jerome Connor, Tomas Kocisky, Mike Chrzanowski, Lingpeng Kong, Angeliki Lazaridou, Wang Ling, Lei Yu, Chris Dyer, and Phil Blunsom. 2019 · 2019
Later among the works it cites.
Climbing towards NLU: On meaning, form, and understanding in the age of data
Emily M. Bender and Alexander Koller. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dieuwke Hupkes and Willem Zuidema. 2018 · 2018
Cited alongside, same era.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
Analysis methods in neural language processing: A survey
Yonatan Belinkov and James Glass. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Designing and interpreting probes with control tasks
John Hewitt and Percy Liang. 2019 · 2019
Cited alongside, same era.
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Later among the works it cites.
Negation
Laurence R. Horn and Heinrich Wansing. 2020 · 2020
Later among the works it cites.
To dissect an octopus: Making sense of the form/meaning debate
Julian Michael. 2020 · 2020
Later among the works it cites.
Is it possible for language models to achieve understanding?
Christopher Potts. 2020 · 2020
Later among the works it cites.
A primer in BERTology: What we know about how bert works
Anna Rogers, Olga Kovaleva, and Anna Rumshisky. 2020 · 2020
Later among the works it cites.
WinoWhy: A deep diagnosis of essential commonsense knowledge for answering Winograd schema challenge
Hongming Zhang, Xinran Zhao, and Yangqiu Song. 2020 · 2020
Later among the works it cites.