Fetching the paper…
Reading the bibliography…
Despite their exceptional capabilities, large language models (LLMs) are prone to generating unintended text due to false or outdated knowledge.
Linear hinge loss and average margin
Gentile, C. and Warmuth, M. K · 1998
Earlier work this paper cites.
Direct and indirect effects
Pearl, J · 2001
Earlier work this paper cites.
Robust truncated hinge loss support vector machines
Wu, Y. and Liu, Y · 2007
Earlier work this paper cites.
Modifying memories in transformer models
Zhu, C., Rawat, A. S., Zaheer, M., Bhojanapalli, S., Li, D., Yu, F. X., and Kumar, S · 2012
Earlier work this paper cites.
Wikidata: a free collaborative knowledgebase
Vrandecic, D. and Krötzsch, M · 2014
Earlier work this paper cites.
YAGO3: A knowledge base from multilingual wikipedias
Mahdisoltani, F., Biega, J., and Suchanek, F. M · 2015
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al · 2019
Earlier work this paper cites.
Rewriting a deep generative model
Bau, D., Liu, S., Wang, T., Zhu, J., and Torralba, A · 2020
Earlier work this paper cites.
Editable neural networks
Sinitsin, A., Plokhotnyuk, V., Pyrkin, D. V., Popov, S., and Babenko, A · 2020
Earlier work this paper cites.
Investigating gender bias in language models using causal mediation analysis
Vig, J., Gehrmann, S., Belinkov, Y., Qian, S., Nevo, D., Singer, Y., and Shieber, S. M · 2020
Earlier work this paper cites.
Editing factual knowledge in language models
Cao, N. D., Aziz, W., and Titov, I · 2021
Earlier work this paper cites.
Transformer feed-forward layers are key-value memories
Geva, M., Schuster, R., Berant, J., and Levy, O · 2021
Earlier work this paper cites.
Gpt-j-6b: A 6 billion parameter autoregressive language model, 2021
Wang, B. and Komatsuzaki, A · 2021
Earlier work this paper cites.
Knowledge neurons in pretrained transformers
Dai, D., Dong, L., Hao, Y., Sui, Z., Chang, B., and Wei, F · 2022
Earlier work this paper cites.
Transformer feed-forward layers build predictions by promoting concepts in the vocabulary space
Geva, M., Caciularu, A., Wang, K. R., and Goldberg, Y · 2022
Cited alongside, same era.
Wider & closer: Mixture of short-channel distillers for zero-shot cross-lingual named entity recognition
Ma, J., Chen, B., Gu, J., Ling, Z., Guo, W., Liu, Q., Chen, Z., and Liu, C · 2022
Cited alongside, same era.
Locating and editing factual associations in GPT
Meng, K., Bau, D., Andonian, A., and Belinkov, Y · 2022
Cited alongside, same era.
Fast model editing at scale
Mitchell, E., Lin, C., Bosselut, A., Finn, C., and Manning, C. D · 2022
Cited alongside, same era.
Memory-based model editing at scale
Mitchell, E., Lin, C., Bosselut, A., Manning, C. D., and Finn, C · 2022
Cited alongside, same era.
Mass-editing memory in a transformer
Meng, K., Sharma, A. S., Andonian, A. J., Belinkov, Y., and Bau, D · 2023
Later among the works it cites.
OpenAI · 2023
Later among the works it cites.
Peng, B., Galley, M., He, P., Cheng, H., Xie, Y., Hu, Y., Huang, Q., Liden, L., Yu, Z., Chen, W., and Gao, J · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Touvron, H., Martin, L., Stone, K., Albert, P., Almahairi, A., Babaei, Y., Bashlykov, N., Batra, S., Bhargava, P., Bhosale, S., Bikel, D., Blecher, L., Canton-Ferrer, C., Chen, M., Cucurull, G., Esiobu, D., Fernandes, J., Fu, J., Fu, W., Fuller, B., Gao, C., Goswami, V., Goyal, N., Hartshorn, A., Hosseini, S., Hou, R., Inan, H., Kardas, M., Kerkez, V., Khabsa, M., Kloumann, I., Korenev, A., Koura, P. S., Lachaux, M., Lavril, T., Lee, J., Liskovich, D., Lu, Y., et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C. L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P. F., Leike, J., and Lowe, R · 2022
Cited alongside, same era.
Can we edit multimodal large language models?
Cheng, S., Tian, B., Liu, Q., Chen, X., Wang, Y., Chen, H., and Zhang, N · 2023
Cited alongside, same era.
Evaluating the ripple effects of knowledge editing in language models
Cohen, R., Biran, E., Yoran, O., Globerson, A., and Geva, M · 2023
Cited alongside, same era.
Erasing concepts from diffusion models
Gandikota, R., Materzynska, J., Fiotto-Kaufman, J., and Bau, D · 2023
Cited alongside, same era.
Huang, L., Yu, W., Ma, W., Zhong, W., Feng, Z., Wang, H., Chen, Q., Peng, W., Feng, X., Qin, B., and Liu, T · 2023
Cited alongside, same era.
Survey of hallucination in natural language generation
Ji, Z., Lee, N., Frieske, R., Yu, T., Su, D., Xu, Y., Ishii, E., Bang, Y., Madotto, A., and Fung, P · 2023
Cited alongside, same era.
Unveiling the pitfalls of knowledge editing for large language models
Li, Z., Zhang, N., Yao, Y., Wang, M., Chen, X., and Chen, H · 2023
Cited alongside, same era.
Wang, P., Zhang, N., Xie, X., Yao, Y., Tian, B., Wang, M., Xi, Z., Cheng, S., Liu, K., Zheng, G., and Chen, H · 2023
Later among the works it cites.
DEPN: detecting and editing privacy neurons in pretrained language models
Wu, X., Li, J., Xu, M., Dong, W., Wu, S., Bian, C., and Xiong, D · 2023
Later among the works it cites.
Editing large language models: Problems, methods, and opportunities
Yao, Y., Wang, P., Tian, B., Cheng, S., Li, Z., Deng, S., Chen, H., and Zhang, N · 2023
Later among the works it cites.
History matters: Temporal knowledge editing in large language model
Yin, X., Jiang, J., Yang, L., and Wan, X · 2023
Later among the works it cites.
Prompting large language model for machine translation: A case study
Zhang, B., Haddow, B., and Birch, A · 2023
Later among the works it cites.
Mquake: Assessing knowledge editing in language models via multi-hop questions
Zhong, Z., Wu, Z., Manning, C. D., Potts, C., and Chen, D · 2023
Later among the works it cites.
Model editing can hurt general abilities of large language models
Gu, J., Xu, H., Ma, J., Lu, P., Ling, Z., Chang, K., and Peng, N · 2024
Closest in time.
Deepedit: Knowledge editing as decoding with constraints
Wang, Y., Chen, M., Peng, N., and Chang, K.-W · 2024
Closest in time.
A comprehensive study of knowledge editing for large language models
Zhang, N., Yao, Y., Tian, B., Wang, P., Deng, S., Wang, M., Xi, Z., Mao, S., Zhang, J., Ni, Y., et al · 2024
Closest in time.