Fetching the paper…
Reading the bibliography…
Although pretrained language models (PTLMs) contain significant amounts of world knowledge, they can still produce inconsistent answers to questions when probed, even after specialized training.
Programs with common sense
John W. McCarthy. 1959 · 1959
Earlier work this paper cites.
Mental Models
Dedre Gentner and Albert L. Stevens. 1983 · 1983
Earlier work this paper cites.
Mental Models : Towards a Cognitive Science of Language
P. Johnson-Laird. 1983 · 1983
Earlier work this paper cites.
Semantical considerations on nonmonotonic logic
R. Moore. 1983 · 1983
Earlier work this paper cites.
An assumption-based tms
Johan De Kleer. 1986 · 1986
Earlier work this paper cites.
Fusion, propagation, and structuring in belief networks
J. Pearl. 1986 · 1986
Earlier work this paper cites.
Logical foundations of artificial intelligence
M. Genesereth and N. Nilsson. 1987 · 1987
Earlier work this paper cites.
Belief maintenance in dynamic constraint networks
Rina Dechter and Avi Dechter. 1988 · 1988
Earlier work this paper cites.
Mental models and causal explanation: Judgements of probable cause and explanatory relevance
D. Hilton. 1996 · 1996
Earlier work this paper cites.
Realm: Retrieval-augmented language model pre-training
Kelvin Guu, Kenton Lee, Z. Tung, Panupong Pasupat, and Ming-Wei Chang. 2020 · 2002
Earlier work this paper cites.
Logic-guided data augmentation and regularization for consistent question answering
Akari Asai and Hannaneh Hajishirzi. 2020 · 2004
Earlier work this paper cites.
Wordnet and wordnets
Christiane Fellbaum. 2005 · 2005
Earlier work this paper cites.
Z3: An efficient SMT solver
Leonardo De Moura and Nikolaj Bjørner. 2008 · 2008
Earlier work this paper cites.
Probabilistic similarity logic
Matthias Broecheler, Lilyana Mihalkova, and L. Getoor. 2010 · 2010
Earlier work this paper cites.
Toward an architecture for never-ending language learning
Andrew Carlson, J. Betteridge, Bryan Kisiel, Burr Settles, Estevam R. Hruschka, and Tom Michael Mitchell. 2010 · 2010
Earlier work this paper cites.
Thinking, fast and slow
D. Kahneman. 2011 · 2011
Cited alongside, same era.
Knowledge graph identification
J. Pujara, H. Miao, L. Getoor, and William W. Cohen. 2013 · 2013
Cited alongside, same era.
End-to-end memory networks
Sainbayar Sukhbaatar, Arthur D. Szlam, J. Weston, and R. Fergus. 2015 · 2015
Cited alongside, same era.
Machine teaching: An inverse problem to machine learning and an approach toward optimal education
Xiaojin Zhu. 2015 · 2015
Cited alongside, same era.
Hybrid computing using a neural network with dynamic external memory
A. Graves, Greg Wayne, M. Reynolds, Tim Harley, Ivo Danihelka, Agnieszka Grabska-Barwinska, Sergio Gomez Colmenarejo, Edward Grefenstette, Tiago Ramalho, J. Agapiou, Adrià Puigdomènech Badia, K. Hermann, Yori Zwols, Georg Ostrovski, Adam Cain, Helen King, C. Summerfield, P. Blunsom, K. Kavukcuoglu, and D. Hassabis. 2016 · 2016
Cited alongside, same era.
Tracking the world state with recurrent entity networks
Mikael Henaff, J. Weston, Arthur D. Szlam, Antoine Bordes, and Y. LeCun. 2016 · 2016
What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models
Allyson Ettinger. 2020 · 2020
Later among the works it cites.
Negated and misprimed probes for pretrained language models: Birds can talk, but cannot fly
Nora Kassner and H. Schütze. 2020 · 2020
Later among the works it cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandara Piktus, Fabio Petroni, V. Karpukhin, Naman Goyal, Heinrich Kuttler, M. Lewis, Wen tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. 2020 · 2020
Later among the works it cites.
On the systematicity of probing contextualized word representations: The case of hypernymy in BERT
Abhilasha Ravichander, Eduard Hovy, Kaheer Suleman, Adam Trischler, and Jackie Chi Kit Cheung. 2020 · 2020
Later among the works it cites.
How much knowledge can you pack into the parameters of a language model?
Adam Roberts, Colin Raffel, and Noam Shazeer. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Conceptnet 5.5: An open multilingual graph of general knowledge
Robyn Speer, Joshua Chin, and Catherine Havasi. 2017 · 2017
Cited alongside, same era.
Adversarially regularising neural NLI models to integrate logical background knowledge
Pasquale Minervini and S. Riedel. 2018 · 2018
Cited alongside, same era.
Commonsense knowledge mining from pretrained models
Joe Davison, Joshua Feldman, and Alexander Rush. 2019 · 2019
Cited alongside, same era.
A logic-driven framework for consistency of neural models
Tao Li, Vivek Gupta, Maitrey Mehta, and Vivek Srikumar. 2019 · 2019
Cited alongside, same era.
Continual lifelong learning with neural networks: A review
G. I. Parisi, Ronald Kemker, Jose L. Part, Christopher Kanan, and S. Wermter. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
F. Petroni, Tim Rocktäschel, Patrick Lewis, A. Bakhtin, Yuxiang Wu, Alexander H. Miller, and S. Riedel. 2019 · 2019
Cited alongside, same era.
Unsupervised commonsense question answering with self-talk
Vered Shwartz, Peter West, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2020 · 2020
Later among the works it cites.
Obtaining faithful interpretations from compositional neural networks
Sanjay Subramanian, Ben Bogin, Nitish Gupta, Tomer Wolfson, Sameer Singh, Jonathan Berant, and Matt Gardner. 2020 · 2020
Later among the works it cites.
Leap-of-thought: Teaching pre-trained models to systematically reason over implicit knowledge
Alon Talmor, Oyvind Tafjord, Peter Clark, Yoav Goldberg, and Jonathan Berant. 2020 · 2020
Later among the works it cites.
Remembering for the right reasons: Explanations reduce catastrophic forgetting
Sayna Ebrahimi, Suzanne Petryk, Akash Gokul, William Gan, J. Gonzalez, Marcus Rohrbach, and Trevor Darrell. 2021 · 2021
Closest in time.
Measuring and improving consistency in pretrained language models
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, E. Hovy, Hinrich Schütze, and Yoav Goldberg. 2021 · 2021
Closest in time.
Formal Representations of Belief
Konstantin Genin and Franz Huber. 2021 · 2021
Closest in time.
What makes good in-context examples for GPT-3?
Jiachang Liu, Dinghan Shen, Yizhe Zhang, Bill Dolan, L. Carin, and Weizhu Chen. 2021 · 2021
Closest in time.
Conditionally adaptive multi-task learning: Improving transfer learning in nlp using fewer parameters & less data
Jonathan Pilault, Amine Elhattami, and C. Pal. 2021 · 2021
Closest in time.
General-purpose question-answering with Macaw
Oyvind Tafjord and Peter Clark. 2021 · 2021
Closest in time.