Fetching the paper…
Reading the bibliography…
Large language model (LLM) agents are constrained by limited context windows, necessitating external memory systems for long-term information understanding.
Benchmarking natural language understanding services for building conversational agents, 2019
Xingkun Liu, Arash Eshghi, Pawel Swietojanski, and Verena Rieser · 1903
Earlier work this paper cites.
Learning question classifiers
Xin Li and Dan Roth · 2002
Earlier work this paper cites.
Mikhail S. Burtsev and Grigory V. Sapunov · 2006
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang · 2016
Earlier work this paper cites.
Pubmed 200k rct: a dataset for sequential sentence classification in medical abstracts
Franck Dernoncourt and Ji Young Lee · 2017
Earlier work this paper cites.
Hotpotqa: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W Cohen, Ruslan Salakhutdinov, and Christopher D Manning · 2018
Earlier work this paper cites.
An evaluation dataset for intent classification and out-of-scope prediction
Stefan Larson, Anish Mahendran, Joseph J. Peper, Christopher Clarke, Andrew Lee, Parker Hill, Jonathan K. Kummerfeld, Kevin Leach, Michael A. Laurenzano, Lingjia Tang, and Jason Mars · 2019
Earlier work this paper cites.
Efficient intent detection with dual sentence encoders
Iñigo Casanueva, Tadas Temčinas, Daniela Gerz, Matthew Henderson, and Ivan Vulić · 2020
Earlier work this paper cites.
Booksum: A collection of datasets for long-form narrative summarization
Wojciech Kryściński, Nazneen Rajani, Divyansh Agarwal, Caiming Xiong, and Dragomir Radev · 2021
Earlier work this paper cites.
Recurrent memory transformer
Aydar Bulatov, Yuri Kuratov, and Mikhail S. Burtsev · 2022
Earlier work this paper cites.
In-context autoencoder for context compression in a large language model
Tao Ge, Jing Hu, Lei Wang, Xun Wang, Si-Qing Chen, and Furu Wei · 2023
Earlier work this paper cites.
Memochat: Tuning llms to use memos for consistent long-range open-domain conversation
Junru Lu, Siyu An, Mingbao Lin, Gabriele Pergola, Yulan He, Di Yin, Xing Sun, and Yunsheng Wu · 2023
Earlier work this paper cites.
Memgpt: Towards llms as operating systems
Charles Packer, Vivian Fang, Shishir_G Patil, Kevin Lin, Sarah Wooders, and Joseph_E Gonzalez · 2023
Earlier work this paper cites.
Enhancing large language model with self-controlled memory framework
Bing Wang, Xinnian Liang, Jian Yang, Hui Huang, Shuangzhi Wu, Peihao Wu, Lu Lu, Zejun Ma, and Zhoujun Li · 2023
Earlier work this paper cites.
Memorybank: Enhancing large language models with long-term memory
Wanjun Zhong, Lianghong Guo, Qiqi Gao, and Yanlin Wang · 2023
Earlier work this paper cites.
Arigraph: Learning knowledge graph world models with episodic memory for llm agents
Petr Anokhin, Nikita Semenov, Artyom Sorokin, Dmitry Evseev, Andrey Kravchenko, Mikhail Burtsev, and Evgeny Burnaev · 2024
Earlier work this paper cites.
Titans: Learning to memorize at test time
Ali Behrouz, Peilin Zhong, and Vahab Mirrokni · 2024
Earlier work this paper cites.
Vincent-Pierre Berges, Barlas Oğuz, Daniel Haziza, Wen-tau Yih, Luke Zettlemoyer, and Gargi Ghosh · 2024
Cited alongside, same era.
Larimar: Large language models with episodic memory control
Payel Das, Subhajit Chaudhury, Elliot Nelson, Igor Melnyk, Sarathkrishna Swaminathan, Sihui Dai, Aurélie C. Lozano, Georgios Kollias, Vijil Chenthamarakshan, Jirí Navrátil, Soham Dan, and Pin-Yu Chen · 2024
Cited alongside, same era.
Perltqa: A personal long-term memory dataset for memory classification, retrieval, and fusion in question answering
Yiming Du, Hongru Wang, Zhengyi Zhao, Bin Liang, Baojun Wang, Wanjun Zhong, Zezhong Wang, and Kam-Fai Wong · 2024
Cited alongside, same era.
Human-like episodic memory for infinite context llms
Zafeirios Fountas, Martin A Benfeghoul, Adnan Oomerjee, Fenia Christopoulou, Gerasimos Lampouras, Haitham Bou-Ammar, and Jun Wang · 2024
Cited alongside, same era.
Camelot: Towards large language models with training-free consolidated associative memory
Jinyuan Fang, Yanwen Peng, Xi Zhang, Yingxu Wang, Xinhao Yi, Guibin Zhang, Yi Xu, Bin Wu, Siwei Liu, Zihao Li, et al · 2025
Closest in time.
Evaluating memory in llm agents via incremental multi-turn interactions
Yuanzhe Hu, Yu Wang, and Julian McAuley · 2025
Closest in time.
Sleep-time compute: Beyond inference scaling at test-time
Kevin Lin, Charlie Snell, Yu Wang, Charles Packer, Sarah Wooders, Ion Stoica, and Joseph E Gonzalez · 2025
Closest in time.
Echo: A large language model with temporal episodic memory
WenTao Liu, Ruohua Zhang, Aimin Zhou, Feng Gao, and JiaLi Liu · 2025
Closest in time.
Nemori: Self-organizing agent memory inspired by cognitive science
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zexue He, Leonid Karlinsky, Donghyun Kim, Julian McAuley, Dmitry Krotov, and Rogerio Feris · 2024
Cited alongside, same era.
RULER: What’s the Real Context Size of Your Long-Context Language Models?, August 2024
Cheng-Ping Hsieh, Simeng Sun, Samuel Kriman, Shantanu Acharya, Dima Rekesh, Fei Jia, Yang Zhang, and Boris Ginsburg · 2024
Cited alongside, same era.
Memory, consciousness and large language model
Jitang Li and Jinzheng Li · 2024
Cited alongside, same era.
Snapkv: LLM knows what you are looking for before generation
Yuhong Li, Yingbing Huang, Bowen Yang, Bharat Venkitesh, Acyr Locatelli, Hanchen Ye, Tianle Cai, Patrick Lewis, and Deming Chen · 2024
Cited alongside, same era.
Evaluating very long-term conversational memory of llm agents
Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov, Mohit Bansal, Francesco Barbieri, and Yuwei Fang · 2024
Cited alongside, same era.
From isolated conversations to hierarchical schemas: Dynamic tree memory representation for llms
Alireza Rezazadeh, Zichao Li, Wei Wei, and Yujia Bao · 2024
Cited alongside, same era.
Deepseekmath: Pushing the limits of mathematical reasoning in open language models
Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu, Junxiao Song, Xiao Bi, Haowei Zhang, Mingchuan Zhang, YK Li, Yang Wu, et al · 2024
Cited alongside, same era.
Memoryllm: Towards self-updatable large language models
Yu Wang, Yifan Gao, Xiusi Chen, Haoming Jiang, Shiyang Li, Jingfeng Yang, Qingyu Yin, Zheng Li, Xian Li, Bing Yin, et al · 2024
Cited alongside, same era.
Jiayan Nan, Wenquan Ma, Wenlong Wu, and Yize Chen · 2025
Closest in time.
Position: Episodic memory is the missing piece for long-term llm agents
Mathis Pink, Qinyuan Wu, Vy Ai Vo, Javier Turek, Jianing Mu, Alexander Huth, and Mariya Toneva · 2025
Closest in time.
Memorag: Boosting long context processing with global memory-enhanced retrieval augmentation
Hongjin Qian, Zheng Liu, Peitian Zhang, Kelong Mao, Defu Lian, Zhicheng Dou, and Tiejun Huang · 2025
Closest in time.
Zep: A temporal knowledge graph architecture for agent memory
Preston Rasmussen, Pavlo Paliychuk, Travis Beauvais, Jack Ryan, and Daniel Chalef · 2025
Closest in time.
Mirix: Multi-agent memory system for llm-based agents
Yu Wang and Xi Chen · 2025
Closest in time.
Towards lifespan cognitive systems
Yu Wang, Chi Han, Tongtong Wu, Xiaoxin He, Wangchunshu Zhou, Nafis Sadeq, Xiusi Chen, Zexue He, Wei Wang, Gholamreza Haffari, Heng Ji, and Julian J. McAuley · 2025
Closest in time.
Ai-native memory 2.0: Second me
Jiale Wei, Xiang Ying, Tao Gao, Fangyi Bao, Felix Tao, and Jingbo Shang · 2025
Closest in time.
Sikuan Yan, Xiufeng Yang, Zuchao Huang, Ercong Nie, Zifeng Ding, Zonggen Li, Xiaowen Ma, Hinrich Schütze, Volker Tresp, and Yunpu Ma · 2025
Closest in time.
Egomem: Lifelong memory agent for full-duplex omnimodal models
Yiqun Yao, Naitong Yu, Xiang Li, Xin Jiang, Xuezhi Fang, Wenjia Ma, Xuying Meng, Jing Li, Aixin Sun, and Yequan Wang · 2025
Closest in time.
Memagent: Reshaping long-context llm with multi-conv rl-based memory agent
Hongli Yu, Tinghong Chen, Jiangtao Feng, Jiangjie Chen, Weinan Dai, Qiying Yu, Ya-Qin Zhang, Wei-Ying Ma, Jingjing Liu, Mingxuan Wang, et al · 2025
Closest in time.
Memengine: A unified and modular library for developing advanced memory of llm-based agents
Zeyu Zhang, Quanyu Dai, Xu Chen, Rui Li, Zhongyang Li, and Zhenhua Dong · 2025
Closest in time.
Mem1: Learning to synergize memory and reasoning for efficient long-horizon agents
Zijian Zhou, Ao Qu, Zhaoxuan Wu, Sunghwan Kim, Alok Prakash, Daniela Rus, Jinhua Zhao, Bryan Kian Hsiang Low, and Paul Pu Liang · 2025
Closest in time.