Fetching the paper…
Reading the bibliography…
Our ability to continuously acquire, organize, and leverage knowledge is a key feature of human intelligence that AI systems must approximate to unlock their full potential.
Huggingface’s transformers: State-of-the-art natural language processing
Wolf, T., Debut, L., Sanh, V., Chaumond, J., Delangue, C., Moi, A., Cistac, P., Rault, T., Louf, R., Funtowicz, M., and Brew, J · 1910
Earlier work this paper cites.
Topic-sensitive pagerank
Haveliwala, T. H · 2002
Earlier work this paper cites.
Associative learning and the hippocampus
Suzuki, W. A · 2005
Earlier work this paper cites.
Making sense of sensemaking 1: Alternative perspectives
Klein, G., Moon, B., and Hoffman, R. R · 2006
Earlier work this paper cites.
Neural correlates of sparse coding and dimensionality reduction
Beyeler, M., Rounds, E. L., Carlson, K. D., Dutt, N., and Krichmar, J. L · 2019
Earlier work this paper cites.
PyTorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., Desmaison, A., Köpf, A., Yang, E. Z., DeVito, Z., Raison, M., Tejani, A., Chilamkurthy, S., Steiner, B., Fang, L., Bai, J., and Chintala, S · 2019
Earlier work this paper cites.
BEIR: A heterogeneous benchmark for zero-shot evaluation of information retrieval models
Thakur, N., Reimers, N., Rücklé, A., Srivastava, A., and Gurevych, I · 2021
Earlier work this paper cites.
Unsupervised dense information retrieval with contrastive learning
Izacard, G., Caron, M., Hosseini, L., Riedel, S., Bojanowski, P., Joulin, A., and Grave, E · 2022
Earlier work this paper cites.
Lifelong pretraining: Continually adapting language models to emerging corpora
Jin, X., Zhang, D., Zhu, H., Xiao, W., Li, S.-W., Wei, X., Arnold, A., and Ren, X · 2022
Earlier work this paper cites.
StreamingQA: A benchmark for adaptation to new knowledge over time in question answering models
Liska, A., Kocisky, T., Gribovskaya, E., Terzi, T., Sezener, E., Agrawal, D., De Masson D’Autume, C., Scholtes, T., Zaheer, M., Young, S., Gilsenan-Mcmahon, E., Austin, S., Blunsom, P., and Lazaridou, A · 2022
Earlier work this paper cites.
Large dual encoders are generalizable retrievers
Ni, J., Qu, C., Lu, J., Dai, Z., Hernandez Abrego, G., Ma, J., Zhao, V., Luan, Y., Hall, K., Chang, M.-W., and Yang, Y · 2022
Earlier work this paper cites.
MuSiQue: Multihop questions via single-hop question composition
Trivedi, H., Balasubramanian, N., Khot, T., and Sabharwal, A · 2022
Earlier work this paper cites.
Do recall and recognition lead to different retrieval experiences?
Uner, O. and Roediger III, H. L · 2022
Earlier work this paper cites.
Walking down the memory maze: Beyond context limit through interactive reading, 2023
Chen, H., Pasunuru, R., Weston, J., and Celikyilmaz, A · 2023
Earlier work this paper cites.
Detecting edit failures in large language models: An improved specificity benchmark
Hoelscher-Obermaier, J., Persson, J., Kran, E., Konstas, I., and Barez, F · 2023
Earlier work this paper cites.
Efficient memory management for large language model serving with pagedattention
Kwon, W., Li, Z., Zhuang, S., Sheng, Y., Zheng, L., Yu, C. H., Gonzalez, J. E., Zhang, H., and Stoica, I · 2023
Cited alongside, same era.
Towards general text embeddings with multi-stage contrastive learning
Li, Z., Zhang, X., Zhang, Y., Long, D., Xie, P., and Zhang, M · 2023
Cited alongside, same era.
When not to trust language models: Investigating effectiveness of parametric and non-parametric memories
Mallen, A., Asai, A., Zhong, V., Das, R., Khashabi, D., and Hajishirzi, H · 2023
Cited alongside, same era.
Editing large language models: Problems, methods, and opportunities
Yao, Y., Wang, P., Tian, B., Cheng, S., Li, Z., Deng, S., Chen, H., and Zhang, N · 2023
Cited alongside, same era.
CITB: A benchmark for continual instruction tuning
Zhang, Z., Fang, M., Chen, L., and Namazi-Rad, M.-R · 2023
Cited alongside, same era.
Carpe diem: On the evaluation of world knowledge in lifelong language models
Kim, Y., Yoon, J., Ye, S., Bae, S., Ho, N., Hwang, S. J., and Yun, S.-Y · 2024
Later among the works it cites.
Sensemaking of socially-mediated crisis information
Koli, V., Yuan, J., and Dasgupta, A · 2024
Later among the works it cites.
Tic-LM: A multi-year benchmark for continual pretraining of language models
Li, J., Armandpour, M., Mirzadeh, S. I., Mehta, S., Shankar, V., Vemulapalli, R., Tuzel, O., Farajtabar, M., Pouransari, H., and Faghri, F · 2024
Later among the works it cites.
BM25S: Orders of magnitude faster lexical search via eager sparse scoring, 2024
Lù, X. H · 2024
Later among the works it cites.
A practitioner’s guide to continual multimodal pretraining
Roth, K., Udandarao, V., Dziadzio, S., Prabhu, A., Cherti, M., Vinyals, O., Henaff, O. J., Albanie, S., Bethge, M., and Akata, Z · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
MQuAKE: Assessing knowledge editing in language models via multi-hop questions
Zhong, Z., Wu, Z., Manning, C., Potts, C., and Chen, D · 2023
Cited alongside, same era.
Llama 3 model card
AI@Meta · 2024
Cited alongside, same era.
Evaluating the ripple effects of knowledge editing in language models
Cohen, R., Biran, E., Yoran, O., Globerson, A., and Geva, M · 2024
Cited alongside, same era.
From local to global: A graph rag approach to query-focused summarization, 2024
Edge, D., Trinh, H., Cheng, N., Bradley, J., Chao, A., Mody, A., Truitt, S., and Larson, J · 2024
Cited alongside, same era.
Model editing harms general abilities of large language models: Regularization to the rescue
Gu, J.-C., Xu, H.-X., Ma, J.-Y., Lu, P., Ling, Z.-H., Chang, K.-W., and Peng, N · 2024
Cited alongside, same era.
LightRAG: Simple and fast retrieval-augmented generation, 2024
Guo, Z., Xia, L., Yu, Y., Ao, T., and Huang, C · 2024
Cited alongside, same era.
Hipporag: Neurobiologically inspired long-term memory for large language models
Gutiérrez, B. J., Shu, Y., Gu, Y., Yasunaga, M., and Su, Y · 2024
Cited alongside, same era.
RAPTOR: recursive abstractive processing for tree-organized retrieval
Sarthi, P., Abdullah, S., Tuli, A., Khanna, S., Goldie, A., and Manning, C. D · 2024
Later among the works it cites.
Continual learning of large language models: A comprehensive survey
Shi, H., Xu, Z., Wang, H., Qin, W., Wang, W., Wang, Y., Wang, Z., Ebrahimi, S., and Wang, H · 2024
Later among the works it cites.
REAR: A relevance-aware retrieval-augmented framework for open-domain question answering
Wang, Y., Ren, R., Li, J., Zhao, X., Liu, J., and Wen, J · 2024
Later among the works it cites.
Adaptive chameleon or stubborn sloth: Revealing the behavior of large language models in knowledge conflicts
Xie, J., Zhang, K., Chen, J., Lou, R., and Su, Y · 2024
Later among the works it cites.
LV-Eval: A balanced long-context benchmark with 5 length levels up to 256k, 2024
Yuan, T., Ning, X., Zhou, D., Yang, Z., Li, S., Zhuang, M., Tan, Z., Yao, Z., Lin, D., Li, B., Dai, G., Yan, S., and Wang, Y · 2024
Later among the works it cites.
Copr: Continual learning human preference through optimal policy regularization, 2024
Zhang, H., Gui, L., Zhai, Y., Wang, H., Lei, Y., and Xu, R · 2024
Later among the works it cites.
NV-embed: Improved techniques for training LLMs as generalist embedding models
Lee, C., Roy, R., Xu, M., Raiman, J., Shoeybi, M., Catanzaro, B., and Ping, W · 2025
Closest in time.
Generative representational instruction tuning
Muennighoff, N., SU, H., Wang, L., Yang, N., Wei, F., Yu, T., Singh, A., and Kiela, D · 2025
Closest in time.
Some simple effective approximations to the 2-poisson model for probabilistic weighted retrieval
Robertson, S. E. and Walker, S · 2099
Closest in time.