Fetching the paper…
Reading the bibliography…
Model editing techniques modify a minor proportion of knowledge in Large Language Models (LLMs) at a relatively low cost, which have demonstrated notable success.
Knowledge Consistency between Neural Networks and Beyond
Liang, R.; Li, T.; Li, L.; Wang, J.; and Zhang, Q. 2020 · 1908
Earlier work this paper cites.
The Finley affair: A signal event in the history of forecast verification
Murphy, A. H. 1996 · 1996
Earlier work this paper cites.
Sinitsin, A.; Plokhotnyuk, V.; Pyrkin, D.; Popov, S.; and Babenko, A. 2020 · 2004
Earlier work this paper cites.
Modifying Memories in Transformer Models
Zhu, C.; Rawat, A. S.; Zaheer, M.; Bhojanapalli, S.; Li, D.; Yu, F.; and Kumar, S. 2020 · 2012
Earlier work this paper cites.
Zero-Shot Relation Extraction via Reading Comprehension
Levy, O.; Seo, M.; Choi, E.; and Zettlemoyer, L. 2017 · 2017
Earlier work this paper cites.
Revealing the Dark Secrets of BERT
Kovaleva, O.; Romanov, A.; Rogers, A.; and Rumshisky, A. 2019 · 2019
Earlier work this paper cites.
Language Models as Knowledge Bases?
Petroni, F.; Rocktäschel, T.; Riedel, S.; Lewis, P.; Bakhtin, A.; Wu, Y.; and Miller, A. 2019 · 2019
Earlier work this paper cites.
Knowledgeable or Educated Guess? Revisiting Language Models as Knowledge Bases
Cao, B.; Lin, H.; Han, X.; Sun, L.; Yan, L.; Liao, M.; Xue, T.; and Xu, J. 2021 · 2021
Earlier work this paper cites.
Editing Factual Knowledge in Language Models
De Cao, N.; Aziz, W.; and Titov, I. 2021 · 2021
Earlier work this paper cites.
Transformer Feed-Forward Layers Are Key-Value Memories
Geva, M.; Schuster, R.; Berant, J.; and Levy, O. 2021 · 2021
Cited alongside, same era.
Self-Attention Attribution: Interpreting Information Interactions Inside Transformer
Hao, Y.; Dong, L.; Wei, F.; and Xu, K. 2021 · 2021
Cited alongside, same era.
Language Models as Knowledge Bases: On Entity Representations, Storage Capacity, and Paraphrased Queries
Heinzerling, B.; and Inui, K. 2021 · 2021
Cited alongside, same era.
GPT-NeoX-20B: An Open-Source Autoregressive Language Model
Black, S.; Biderman, S.; Hallahan, E.; Anthony, Q.; Gao, L.; Golding, L.; He, H.; Leahy, C.; McDonell, K.; Phang, J.; Pieler, M.; Prashanth, U. S.; Purohit, S.; Reynolds, L.; Tow, J.; Wang, B.; and Weinbach, S. 2022 · 2022
Cited alongside, same era.
Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space
Geva, M.; Caciularu, A.; Wang, K.; and Goldberg, Y. 2022 · 2022
Cited alongside, same era.
Inspecting and Editing Knowledge Representations in Language Models
Hernandez, E.; Li, B. Z.; and Andreas, J. 2023 · 2023
Closest in time.
Survey of Hallucination in Natural Language Generation
Ji, Z.; Lee, N.; Frieske, R.; Yu, T.; Su, D.; Xu, Y.; Ishii, E.; Bang, Y. J.; Madotto, A.; and Fung, P. 2023 · 2023
Closest in time.
Feed-Forward Blocks Control Contextualization in Masked Language Models
Kobayashi, G.; Kuribayashi, T.; Yokoi, S.; and Inui, K. 2023 · 2023
Closest in time.
PMET: Precise Model Editing in a Transformer
Li, X.; Li, S.; Song, S.; Yang, J.; Ma, J.; and Yu, J. 2023 · 2023
Closest in time.
GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model
Wang, B.; and Komatsuzaki, A. 2021 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hassid, M.; Peng, H.; Rotem, D.; Kasai, J.; Montero, I.; Smith, N. A.; and Schwartz, R. 2022 · 2022
Cited alongside, same era.
Interpretability in the wild: a circuit for indirect object identification in gpt-2 small
Wang, K.; Variengien, A.; Conmy, A.; Shlegeris, B.; and Steinhardt, J. 2022 · 2022
Cited alongside, same era.
Evaluating the Ripple Effects of Knowledge Editing in Language Models
Cohen, R.; Biran, E.; Yoran, O.; Globerson, A.; and Geva, M. 2023 · 2023
Cited alongside, same era.
Dissecting Recall of Factual Associations in Auto-Regressive Language Models
Geva, M.; Bastings, J.; Filippova, K.; and Globerson, A. 2023 · 2023
Cited alongside, same era.
Knowledge Graph Contrastive Learning Based on Relation-Symmetrical Structure
Liang, K.; Liu, Y.; Zhou, S.; Tu, W.; Wen, Y.; Yang, X.; Dong, X.; and Liu, X. 2023a
Cited in the paper.
Learn from relational correlations and periodic events for temporal knowledge graph reasoning
Liang, K.; Meng, L.; Liu, M.; Liu, Y.; Tu, W.; Wang, S.; Zhou, S.; and Liu, X. 2023b
Cited in the paper.
Locating and Editing Factual Associations in GPT
Meng, K.; Bau, D.; Andonian, A.; and Belinkov, Y. 2022a
Cited in the paper.
Editing Large Language Models: Problems, Methods, and Opportunities
Yao, Y.; Wang, P.; Tian, B.; Cheng, S.; Li, Z.; Deng, S.; Chen, H.; and Zhang, N. 2023 · 2023
Closest in time.
A Survey of Large Language Models
Zhao, W. X.; Zhou, K.; Li, J.; Tang, T.; Wang, X.; Hou, Y.; Min, Y.; Zhang, B.; Zhang, J.; Dong, Z.; Du, Y.; Yang, C.; Chen, Y.; Chen, Z.; Jiang, J.; Ren, R.; Li, Y.; Tang, X.; Liu, Z.; Liu, P.; Nie, J.-Y.; and Wen, J.-R. 2023 · 2023
Closest in time.
Can We Edit Factual Knowledge by In-Context Learning?
Zheng, C.; Li, L.; Dong, Q.; Fan, Y.; Wu, Z.; Xu, J.; and Chang, B. 2023 · 2023
Closest in time.