Fetching the paper…
Reading the bibliography…
Knowledge claims are abundant in the literature on large language models (LLMs); but can we say that GPT-4 truly "knows" the Earth is round? To address this question, we review standard definitions of knowledge in epistemology and we formalize interpretations applicable to LLMs.
The Concept of Mind: 60Th Anniversary Edition
Gilbert Ryle. 1949 · 1949
Earlier work this paper cites.
Is Justified True Belief Knowledge?
Edmund L. Gettier. 1963 · 1963
Earlier work this paper cites.
Semantics for relevant logics
Alasdair Urquhart. 1972 · 1972
Earlier work this paper cites.
The influence curve and its role in robust estimation
Frank R Hampel. 1974 · 1974
Earlier work this paper cites.
The raft and the pyramid: Coherence versus foundations in the theory of knowledge
Ernest Sosa. 1980 · 1980
Earlier work this paper cites.
Fodor?s guide to mental representation: The intelligent auntie?s vade-mecum
Jerry A. Fodor. 1985 · 1985
Earlier work this paper cites.
Theory of knowledge , volume 3
Roderick M Chisholm, Roderick Milton Chisholm, Roderick Milton Chisholm, and Roderick Milton Chisholm. 1989 · 1989
Earlier work this paper cites.
The role of trust in knowledge
John Hardwig. 1991 · 1991
Earlier work this paper cites.
Why knowledge is merely true belief
Crispin Sartwell. 1992 · 1992
Earlier work this paper cites.
Virtues and vices of virtue epistemology
John Greco. 1993 · 1993
Earlier work this paper cites.
Warrant and proper function
Alvin Plantinga. 1993 · 1993
Earlier work this paper cites.
"what is knowledge?"
Linda Zagzebski. 1999 · 1999
Earlier work this paper cites.
Other minds
J. L. Austin. 2000 · 2000
Earlier work this paper cites.
. knowledge and scepticism
Robert Nozick. 2000 · 2000
Earlier work this paper cites.
Using weighted max-sat engines to solve mpe
James D. Park. 2002 · 2002
Earlier work this paper cites.
Knowledge, context, and the agent’s point of view
Timothy Williamson. 2005 · 2005
Earlier work this paper cites.
The logic of justification
Sergei Artemov. 2008 · 2008
Earlier work this paper cites.
Understanding black-box predictions via influence functions
Pang Wei Koh and Percy Liang. 2017 · 2017
Earlier work this paper cites.
A regularization approach for incorporating event knowledge and coreference relations into neural discourse parsing
Zeyu Dai and Ruihong Huang. 2019 · 2019
Earlier work this paper cites.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Earlier work this paper cites.
Theaetetus
Plato Plato. 2019 · 2019
Earlier work this paper cites.
How can we know what language models know?
Zhengbao Jiang, Frank F. Xu, Jun Araki, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
Are pretrained language models symbolic reasoners over knowledge?
Nora Kassner, Benno Krojer, and Hinrich Schütze. 2020 · 2020
Cited alongside, same era.
Negated and misprimed probes for pretrained language models: Birds can talk, but cannot fly
Nora Kassner and Hinrich Schütze. 2020 · 2020
Cited alongside, same era.
Estimating training data influence by tracing gradient descent
Garima Pruthi, Frederick Liu, Satyen Kale, and Mukund Sundararajan. 2020 · 2020
Cited alongside, same era.
How much knowledge can you pack into the parameters of a language model?
Adam Roberts, Colin Raffel, and Noam Shazeer. 2020 · 2020
Cited alongside, same era.
BERTnesia: Investigating the capture and forgetting of knowledge in BERT
Jonas Wallat, Jaspreet Singh, and Avishek Anand. 2020 · 2020
RARR: Researching and revising what language models say, using language models
Luyu Gao, Zhuyun Dai, Panupong Pasupat, Anthony Chen, Arun Tejasvi Chaganty, Yicheng Fan, Vincent Zhao, Ni Lao, Hongrae Lee, Da-Cheng Juan, and Kelvin Guu. 2023 · 2023
Later among the works it cites.
Dissecting recall of factual associations in auto-regressive language models
Mor Geva, Jasmijn Bastings, Katja Filippova, and Amir Globerson. 2023 · 2023
Later among the works it cites.
ROSCOE: A suite of metrics for scoring step-by-step reasoning
Olga Golovneva, Moya Peng Chen, Spencer Poff, Martin Corredor, Luke Zettlemoyer, Maryam Fazel-Zarandi, and Asli Celikyilmaz. 2023 · 2023
Later among the works it cites.
Are machine rationales (not) useful to humans? measuring and improving human utility of free-text rationales
Brihi Joshi, Ziyi Liu, Sahana Ramnath, Aaron Chan, Zhewei Tong, Shaoliang Nie, Qifan Wang, Yejin Choi, and Xiang Ren. 2023 · 2023
Later among the works it cites.
Language models with rationality
Nora Kassner, Oyvind Tafjord, Ashish Sabharwal, Kyle Richardson, Hinrich Schuetze, and Peter Clark. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Making Ai Intelligible: Philosophical Foundations
Herman Cappelen and Josh Dever. 2021 · 2021
Cited alongside, same era.
Editing factual knowledge in language models
Nicola De Cao, Wilker Aziz, and Ivan Titov. 2021 · 2021
Cited alongside, same era.
Measuring and improving consistency in pretrained language models
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, Eduard Hovy, Hinrich Schütze, and Yoav Goldberg. 2021 · 2021
Cited alongside, same era.
How can we know when language models know? on the calibration of language models for question answering
Zhengbao Jiang, Jun Araki, Haibo Ding, and Graham Neubig. 2021 · 2021
Cited alongside, same era.
BeliefBank: Adding memory to a pre-trained language model for a systematic notion of belief
Nora Kassner, Oyvind Tafjord, Hinrich Schütze, and Peter Clark. 2021b · 2021
Cited alongside, same era.
The World of an Octopus: How Reporting Bias Influences a Language Model’s Perception of Color
Cory Paik, Stéphane Aroca-Ouellette, Alessandro Roncone, and Katharina Kann. 2021 · 2021
Cited alongside, same era.
DLAMA: A framework for curating culturally diverse facts for probing the knowledge of pretrained language models
Amr Keleg and Walid Magdy. 2023 · 2023
Later among the works it cites.
Mass-editing memory in a transformer
Kevin Meng, Arnab Sen Sharma, Alex J Andonian, Yonatan Belinkov, and David Bau. 2023 · 2023
Later among the works it cites.
DisentQA: Disentangling parametric and contextual knowledge with counterfactual question answering
Ella Neeman, Roee Aharoni, Or Honovich, Leshem Choshen, Idan Szpektor, and Omri Abend. 2023 · 2023
Later among the works it cites.
Cross-lingual consistency of factual knowledge in multilingual language models
Jirui Qi, Raquel Fernández, and Arianna Bisazza. 2023 · 2023
Later among the works it cites.
Characterizing mechanisms for factual recall in language models
Qinan Yu, Jack Merullo, and Ellie Pavlick. 2023 · 2023
Later among the works it cites.
MQuAKE: Assessing knowledge editing in language models via multi-hop questions
Zexuan Zhong, Zhengxuan Wu, Christopher Manning, Christopher Potts, and Danqi Chen. 2023a · 2023
Later among the works it cites.
Hopping too late: Exploring the limitations of large language models on multi-hop queries
Eden Biran, Daniela Gottesman, Sohee Yang, Mor Geva, and Amir Globerson. 2024 · 2024
Closest in time.
Evaluating the ripple effects of knowledge editing in language models
Roi Cohen, Eden Biran, Ori Yoran, Amir Globerson, and Mor Geva. 2024 · 2024
Closest in time.
A chain-of-thought is as strong as its weakest link: A benchmark for verifiers of reasoning chains
Alon Jacovi, Yonatan Bitton, Bernd Bohnet, Jonathan Herzig, Or Honovich, Michael Tseng, Michael Collins, Roee Aharoni, and Mor Geva. 2024 · 2024
Closest in time.
Ra-isf: Learning to answer and understand from retrieval augmentation via iterative self-feedback
Yanming Liu, Xinyue Peng, Xuhong Zhang, Weihao Liu, Jianwei Yin, Jiannan Cao, and Tianyu Du. 2024 · 2024
Closest in time.
An analysis of bias and distrust in social hinge epistemology
Anna Pederneschi. 2024 · 2024
Closest in time.
Locating and editing factual associations in mamba
Arnab Sen Sharma, David Atkinson, and David Bau. 2024 · 2024
Closest in time.
Localizing paragraph memorization in language models
Niklas Stoehr, Mitchell Gordon, Chiyuan Zhang, and Owen Lewis. 2024 · 2024
Closest in time.
Cross-lingual knowledge editing in large language models
Jiaan Wang, Yunlong Liang, Zengkui Sun, Yuxuan Cao, Jiarong Xu, and Fandong Meng. 2024 · 2024
Closest in time.
Retrieval head mechanistically explains long-context factuality
Wenhao Wu, Yizhong Wang, Guangxuan Xiao, Hao Peng, and Yao Fu. 2024 · 2024
Closest in time.
To believe or not to believe your llm
Yasin Abbasi Yadkori, Ilja Kuzborskij, András György, and Csaba Szepesvári. 2024 · 2024
Closest in time.