Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have majorly advanced NLP and AI, and next to their ability to perform a wide range of procedural tasks, a major success factor is their internalized factual knowledge.
As we may think
Vannevar Bush. 1945 · 1945
Earlier work this paper cites.
Availability: A heuristic for judging frequency and probability
Amos Tversky and Daniel Kahneman. 1973 · 1973
Earlier work this paper cites.
Extracting patterns and relations from the world wide web
Sergey Brin. 1998 · 1998
Earlier work this paper cites.
Snowball : extracting relations from large plain-text collections
Eugene Agichtein and Luis Gravano. 2000 · 2000
Earlier work this paper cites.
Web-scale information extraction in knowitall: (preliminary results)
Oren Etzioni, Michael J. Cafarella, Doug Downey, Stanley Kok, Ana-Maria Popescu, Tal Shaked, Stephen Soderland, Daniel S. Weld, and Alexander Yates. 2004 · 2004
Earlier work this paper cites.
DBpedia: A nucleus for a web of open data
Sören Auer, Christian Bizer, Georgi Kobilarov, Jens Lehmann, Richard Cyganiak, and Zachary G. Ives. 2007 · 2007
Earlier work this paper cites.
Yago: a core of semantic knowledge
Fabian M. Suchanek, Gjergji Kasneci, and Gerhard Weikum. 2007 · 2007
Earlier work this paper cites.
Frameworks for entity matching: A comparison
Hanna Köpcke and Erhard Rahm. 2010 · 2010
Earlier work this paper cites.
Identifying relations for open information extraction
Anthony Fader, Stephen Soderland, and Oren Etzioni. 2011 · 2011
Earlier work this paper cites.
Entity linking at web scale
Thomas Lin, Mausam, and Oren Etzioni. 2012 · 2012
Earlier work this paper cites.
Wikidata: a free collaborative knowledgebase
Denny Vrandecic and Markus Krötzsch. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
DBpedia–a large-scale, multilingual knowledge base extracted from wikipedia
Jens Lehmann, Robert Isele, Max Jakob, Anja Jentzsch, Dimitris Kontokostas, Pablo N Mendes, Sebastian Hellmann, Mohamed Morsey, Patrick Van Kleef, Sören Auer, et al. 2015 · 2015
Earlier work this paper cites.
Semeval-2016 task 13: Taxonomy extraction evaluation (texeval-2)
Georgeta Bordea, Els Lefever, and Paul Buitelaar. 2016 · 2016
Earlier work this paper cites.
The knowledge awakens: Keeping knowledge bases fresh with emerging entities
Johannes Hoffart, Dragan Milchevski, Gerhard Weikum, Avishek Anand, and Jaspreet Singh. 2016 · 2016
Earlier work this paper cites.
Linked data quality of DBpedia, Freebase, OpenCyc, Wikidata, and Yago
Michael Färber, Frederic Bartscherer, Carsten Menne, and Achim Rettinger. 2018 · 2018
Earlier work this paper cites.
Never-ending learning
Tom Mitchell, William Cohen, Estevam Hruschka, Partha Talukdar, Bishan Yang, Justin Betteridge, Andrew Carlson, Bhavana Dalvi, Matt Gardner, Bryan Kisiel, et al. 2018 · 2018
Earlier work this paper cites.
How much is a triple? estimating the cost of knowledge graph creation
Heiko Paulheim. 2018 · 2018
Earlier work this paper cites.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Autoknow: Self-driving knowledge collection for products of thousands of types
Xin Luna Dong, Xiang He, Andrey Kan, Xian Li, Yan Liang, Jun Ma, Yifan Ethan Xu, Chenwei Zhang, Tong Zhao, Gabriel Blanco Saldana, et al. 2020 · 2020
Cited alongside, same era.
How can we know what language models know?
Zhengbao Jiang, Frank F. Xu, Jun Araki, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
How much knowledge can you pack into the parameters of a language model?
Adam Roberts, Colin Raffel, and Noam Shazeer. 2020 · 2020
Cited alongside, same era.
Semantics-aware BERT for language understanding
Zhuosheng Zhang, Yuwei Wu, Hai Zhao, Zuchao Li, Shuailiang Zhang, Xi Zhou, and Xiang Zhou. 2020 · 2020
Cited alongside, same era.
BeliefBank: Adding memory to a pre-trained language model for a systematic notion of belief
Physics of language models: Part 3.3, knowledge capacity scaling laws
Zeyuan Allen-Zhu and Yuanzhi Li. 2024 · 2024
Closest in time.
The reversal curse: LLMs trained on “a is b” fail to learn “b is a”
Lukas Berglund, Meg Tong, Maximilian Kaufmann, Mikita Balesni, Asa Cooper Stickland, Tomasz Korbak, and Owain Evans. 2024 · 2024
Closest in time.
Entgpt: Linking generative large language models with knowledge bases
Yifan Ding, Amrit Poudel, Qingkai Zeng, Tim Weninger, Balaji Veeramani, and Sanmitra Bhattacharya. 2024 · 2024
Closest in time.
Abhimanyu Dubey, Abhinav Jauhri, and Abhinav Pandey et al. 2024 · 2024
Closest in time.
Defining knowledge: Bridging epistemology and large language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nora Kassner, Oyvind Tafjord, Hinrich Schütze, and Peter Clark. 2021 · 2021
Cited alongside, same era.
Analyzing race and citizenship bias in wikidata
Zaina Shaik, Filip Ilievski, and Fred Morstatter. 2021 · 2021
Cited alongside, same era.
KEPLER: A unified model for knowledge embedding and pre-trained language representation
Xiaozhi Wang, Tianyu Gao, Zhaocheng Zhu, Zhengyan Zhang, Zhiyuan Liu, Juanzi Li, and Jian Tang. 2021 · 2021
Cited alongside, same era.
Machine knowledge: Creation and curation of comprehensive knowledge bases
Gerhard Weikum, Xin Luna Dong, Simon Razniewski, Fabian Suchanek, et al. 2021 · 2021
Cited alongside, same era.
Bertnet: Harvesting knowledge graphs from pretrained language models
Shibo Hao, Bowen Tan, and Kaiwen Tang. 2022 · 2022
Cited alongside, same era.
Knowledge graphs
Aidan Hogan, Eva Blomqvist, Michael Cochez, Claudia d’Amato, Gerard de Melo, Claudio Gutierrez, Sabrina Kirrane, José Emilio Labra Gayo, Roberto Navigli, Sebastian Neumaier, Axel-Cyrille Ngonga Ngomo, Axel Polleres, Sabbir M. Rashid, Anisa Rula, Lukas Schmelzeisen, Juan F. Sequeda, Steffen Staab, and Antoine Zimmermann. 2022 · 2022
Cited alongside, same era.
LM-KBC: Knowledge base construction from pre-trained language models
Sneha Singhania, Tuan-Phong Nguyen, and Simon Razniewski. 2022 · 2022
Cited alongside, same era.
Constanza Fierro, Ruchira Dhar, Filippos Stamatiou, Nicolas Garneau, and Anders Søgaard. 2024 · 2024
Closest in time.
Number of parameters in gpt-4 (latest data)
Josh Howarth. 2024 · 2024
Closest in time.
Cultural commonsense knowledge for intercultural dialogues
Tuan-Phong Nguyen, Simon Razniewski, and Gerhard Weikum. 2024 · 2024
Closest in time.
Sharing & publication policy
OpenAI. 2022 · 2024
Closest in time.
OpenAI. 2024 · 2024
Closest in time.
Refining Wikidata taxonomy using large language models
Yiwen Peng, Thomas Bonald, and Mehwish Alam. 2024 · 2024
Closest in time.
Completeness, recall, and negation in open-world knowledge bases: A survey
Simon Razniewski, Hiba Arnaout, Shrestha Ghosh, and Fabian Suchanek. 2024 · 2024
Closest in time.
Introducing the knowledge graph: Things, not strings
Amit Singhal. 2012 · 2024
Closest in time.
YAGO 4.5: A large and clean knowledge base with a rich taxonomy
Fabian M. Suchanek, Mehwish Alam, Thomas Bonald, Lihu Chen, Pierre-Henri Paris, and Jules Soria. 2024 · 2024
Closest in time.
Head-to-tail: How knowledgeable are large language models (LLMs)? A.K.A. will LLMs replace knowledge graphs?
Kai Sun, Yifan Xu, Hanwen Zha, Yue Liu, and Xin Luna Dong. 2024 · 2024
Closest in time.
Can llms express their uncertainty? an empirical evaluation of confidence elicitation in llms
Miao Xiong, Zhiyuan Hu, Xinyang Lu, Yifei Li, Jie Fu, Junxian He, and Bryan Hooi. 2024 · 2024
Closest in time.
Openai unveils gpt-4o mini, a smaller and cheaper ai model
Maxwell Zeff. 2024 · 2024
Closest in time.
Automated mining of structured knowledge from text in the era of large language models
Yunyi Zhang, Ming Zhong, Siru Ouyang, Yizhu Jiao, Sizhe Zhou, Linyi Ding, and Jiawei Han. 2024 · 2024
Closest in time.
Towards reliable latent knowledge estimation in LLMs: In-context learning vs. prompting based factual knowledge extraction
Qinyuan Wu, Mohammad Aflah Khan, Soumi Das, Vedant Nanda, Bishwamittra Ghosh, Camila Kolling, Till Speicher, Laurent Bindschaedler, Krishna P Gummadi, and Evimaria Terzi. 2025 · 2025
Closest in time.