Fetching the paper…
Reading the bibliography…
Large language models (LLMs) acquire information from pre-training corpora, but their stored knowledge can become inaccurate or outdated over time.
A markovian decision process
Bellman, R · 1957
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Puterman, M. L · 1994
Earlier work this paper cites.
Reinforcement learning: An introduction
Sutton, R. and Barto, A · 1998
Earlier work this paper cites.
Automatically constructing a corpus of sentential paraphrases
Dolan, W. B. and Brockett, C · 2005
Earlier work this paper cites.
The fifth PASCAL recognizing textual entailment challenge
Bentivogli, L., Magnini, B., Dagan, I., Dang, H. T., and Giampiccolo, D · 2009
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Socher, R., Perelygin, A., Wu, J., Chuang, J., Manning, C. D., Ng, A., and Potts, C · 2013
Earlier work this paper cites.
Deep reinforcement learning: A brief survey
Arulkumaran, K., Deisenroth, M. P., Brundage, M., and Bharath, A. A · 2017
Earlier work this paper cites.
Zero-shot relation extraction via reading comprehension
Levy, O., Seo, M., Choi, E., and Zettlemoyer, L · 2017
Earlier work this paper cites.
FEVER: a large-scale dataset for fact extraction and VERification
Thorne, J., Vlachos, A., Christodoulopoulos, C., and Mittal, A · 2018
Earlier work this paper cites.
A broad-coverage challenge corpus for sentence understanding through inference
Williams, A., Nangia, N., and Bowman, S · 2018
Earlier work this paper cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Wang, A., Singh, A., Michael, J., Hill, F., Levy, O., and Bowman, S. R · 2019
Earlier work this paper cites.
Neural network acceptability judgments
Warstadt, A., Singh, A., and Bowman, S. R · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., and Amodei, D · 2020
Earlier work this paper cites.
Editing factual knowledge in language models
De Cao, N., Aziz, W., and Titov, I · 2021
Earlier work this paper cites.
Measuring massive multitask language understanding
Hendrycks, D., Burns, C., Basart, S., Zou, A., Mazeika, M., Song, D., and Steinhardt, J · 2021
Earlier work this paper cites.
Mind the gap: Assessing temporal generalization in neural language models
Lazaridou, A., Kuncoro, A., Gribovskaya, E., Agrawal, D., Liska, A., Terzi, T., Gimenez, M., de Masson d'Autume, C., Kocisky, T., Ruder, S., Yogatama, D., Cao, K., Young, S., and Blunsom, P · 2021
Cited alongside, same era.
Modifying memories in transformer models, 2021
Zhu, C., Rawat, A. S., Zaheer, M., Bhojanapalli, S., Li, D., Yu, F., and Kumar, S · 2021
Cited alongside, same era.
Knowledge neurons in pretrained transformers
Dai, D., Dong, L., Hao, Y., Sui, Z., Chang, B., and Wei, F · 2022
Cited alongside, same era.
Locating and editing factual associations in gpt
Meng, K., Bau, D., Andonian, A., and Belinkov, Y · 2022
Cited alongside, same era.
Differentiable neuro-symbolic reasoning on large-scale knowledge graphs
Chen, S., Cai, Y., Fang, H., Huang, X., and Sun, M · 2023
Cited alongside, same era.
Aging with grace: Lifelong model editing with discrete key-value adaptors
Pmet: Precise model editing in a transformer
Li, X., Li, S., Song, S., Yang, J., Ma, J., and Yu, J · 2024
Later among the works it cites.
Flipattack: Jailbreak llms via flipping
Liu, Y., He, X., Xiong, M., Fu, J., Deng, S., and Hooi, B · 2024
Later among the works it cites.
Perturbation-restrained sequential model editing, 2024
Ma, J.-Y., Wang, H., Xu, H.-X., Ling, Z.-H., and Gu, J.-C · 2024
Later among the works it cites.
Massive editing for large language models via meta learning
Tan, C., Zhang, G., and Fu, J · 2024
Later among the works it cites.
Gemma 2: Improving open language models at a practical size, 2024
Team, G., Riviere, M., Pathak, S., Sessa, P. G., and et al., C. H · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hartvigsen, T., Sankaranarayanan, S., Palangi, H., Kim, Y., and Ghassemi, M · 2023
Cited alongside, same era.
Transformer-patcher: One mistake worth one neuron
Huang, Z., Shen, Y., Zhang, X., Zhou, J., Rong, W., and Xiong, Z · 2023
Cited alongside, same era.
Jiang, A. Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D. S., de las Casas, D., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., Lavaud, L. R., Lachaux, M.-A., Stock, P., Scao, T. L., Lavril, T., Wang, T., Lacroix, T., and Sayed, W. E · 2023
Cited alongside, same era.
Mass-editing memory in a transformer
Meng, K., Sharma, A. S., Andonian, A. J., Belinkov, Y., and Bau, D · 2023
Cited alongside, same era.
Can we edit factual knowledge by in-context learning?
Zheng, C., Li, L., Dong, Q., Fan, Y., Wu, Z., Xu, J., and Chang, B · 2023
Cited alongside, same era.
O-edit: Orthogonal subspace editing for language model sequential editing, 2024
Cai, Y. and Cao, D · 2024
Cited alongside, same era.
Alphaedit: Null-space constrained knowledge editing for language models
Fang, J., Jiang, H., Wang, K., Ma, Y., Wang, X., He, X., and Chua, T · 2024
Cited alongside, same era.
Wise: Rethinking the knowledge memory for lifelong model editing of large language models, 2024
Wang, P., Li, Z., Zhang, N., Xu, Z., Yao, Y., Jiang, Y., Xie, P., Huang, F., and Chen, H · 2024
Later among the works it cites.
The butterfly effect of model editing: Few edits can trigger large language models collapse
Yang, W., Sun, F., Ma, X., Liu, X., Yin, D., and Cheng, X · 2024
Later among the works it cites.
The fall of ROME: Understanding the collapse of LLMs in model editing
Yang, W., Sun, F., Tan, J., Ma, X., Su, D., Yin, D., and Shen, H · 2024
Later among the works it cites.
DAFNet: Dynamic auxiliary fusion for sequential model editing in large language models
Zhang, T., Chen, Q., Li, D., Wang, C., He, X., Huang, L., Xue’, H., and Huang, J · 2024
Later among the works it cites.
A survey of large language models, 2024
Zhao, W. X., Zhou, K., Li, J., Tang, T., Wang, X., Hou, Y., Min, Y., Zhang, B., Zhang, J., Dong, Z., Du, Y., Yang, C., Chen, Y., Chen, Z., Jiang, J., Ren, R., Li, Y., Tang, X., Liu, Z., Liu, P., Nie, J.-Y., and Wen, J.-R · 2024
Later among the works it cites.
On the role of attention heads in large language model safety
Zhou, Z., Yu, H., Zhang, X., Xu, R., Huang, F., Wang, K., Liu, Y., Fang, J., and Li, Y · 2024
Later among the works it cites.
Can knowledge editing really correct hallucinations?, 2025
Huang, B., Chen, C., Xu, X., Payani, A., and Shu, K · 2025
Closest in time.
Anyedit: Edit any knowledge encoded in language models
Jiang, H., Fang, J., Zhang, N., Ma, G., Wan, M., Wang, X., He, X., and Chua, T.-s · 2025
Closest in time.
Explainable and efficient editing for large language models
Zhang, T., Fang, J., Jiang, H., Bi, B., Wang, X., and He, X · 2025
Closest in time.