Fetching the paper…
Reading the bibliography…
Memory is fundamental to large language model (LLM)-based agents, but existing surveys emphasize application-level use (e.g., personalized dialogue), while overlooking the atomic operations governing memory dynamics.
Episodic and Semantic Memory
Endel Tulving et al · 1972
Earlier work this paper cites.
K-Lines: A theory of memory
Marvin Minsky. 1980 · 1980
Earlier work this paper cites.
Cognitive Psychology and Human Memory
Alan Baddeley. 1988 · 1988
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: Insights from the successes and failures of connectionist models of learning and memory
James L McClelland, Bruce L McNaughton, and Randall C O’Reilly. 1995 · 1995
Earlier work this paper cites.
Okapi at TREC-3
Stephen E Robertson, Steve Walker, Susan Jones, Micheline M Hancock-Beaulieu, Mike Gatford, et al · 1995
Earlier work this paper cites.
The magical number 4 in short-term memory: A reconsideration of mental storage capacity
Nelson Cowan. 2001 · 2001
Earlier work this paper cites.
Amazon.com Recommendations: Item-to-Item Collaborative Filtering
Greg Linden, Brent Smith, and Jeremy York. 2003 · 2003
Earlier work this paper cites.
Concept and Role Forgetting in ALC Ontologies. In Proceedings of the 8th International Semantic Web Conference (ISWC)
Kewen Wang, Zhe Wang, Rodney Topor, et al · 2009
Earlier work this paper cites.
Neo4j - The World’s Leading Graph Database
Neo4j. 2012 · 2012
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Peter Young, Alice Lai, Micah Hodosh, and Julia Hockenmaier. 2014 · 2014
Earlier work this paper cites.
The consolidation and transformation of memory
Yadin Dudai, Avi Karni, and Jan Born. 2015 · 2015
Earlier work this paper cites.
Memory consolidation
Larry R Squire, Lisa Genzel, John T Wixted, and Richard G Morris. 2015 · 2015
Earlier work this paper cites.
Relative Citation Ratio (RCR): A New Metric That Uses Citation Rates to Measure Influence at the Article Level
B. Ian Hutchins, Xin Yuan, and James M. et al. Anderson. 2016 · 2016
Earlier work this paper cites.
What learning systems do intelligent agents need? Complementary learning systems theory updated
Dharshan Kumaran, Demis Hassabis, and James L McClelland. 2016 · 2016
Earlier work this paper cites.
Abstractive Text Summarization using Sequence-to-sequence RNNs and Beyond. In Proceedings of the 20th SIGNLL Conference on Computational Natural Language Learning
Ramesh Nallapati, Bowen Zhou, Cicero dos Santos, Çağlar Gu ˙ \dot{} lçehre, and Bing Xiang. 2016 · 2016
Earlier work this paper cites.
The Value of Semantic Parse Labeling for Knowledge Base Question Answering. In Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Wen-tau Yih, Ming-Wei Chang, Xiaodong He, and Jianfeng Gao. 2016 · 2016
Earlier work this paper cites.
The biology of forgetting—a perspective
Ronald L Davis and Yi Zhong. 2017 · 2017
Earlier work this paper cites.
TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Mandar Joshi, Eunsol Choi, Daniel Weld, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Overcoming Catastrophic Forgetting in Neural Networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, et al · 2017
Earlier work this paper cites.
Zero-shot Relation Extraction via Reading Comprehension. In Proceedings of CoNLL
Omer Levy, Minjoon Seo, Eunsol Choi, et al · 2017
Earlier work this paper cites.
Pointer Sentinel Mixture Models. In International Conference on Learning Representations
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher. 2017 · 2017
Earlier work this paper cites.
The NarrativeQA Reading Comprehension Challenge
Tom’aš Kočisk’y, Jonathan Schwarz, Phil Blunsom, Chris Dyer, Karl Moritz Hermann, G’abor Melis, and Edward Grefenstette. 2018 · 2018
Earlier work this paper cites.
Meta-learning through Hebbian plasticity in random networks. In Advances in Neural Information Processing Systems (NeurIPS)
Steven Ritter, Jane X Wang, Zeb Kurth-Nelson, Siddhant Jayakumar, Charles Blundell, and Timothy Lillicrap. 2018 · 2018
Earlier work this paper cites.
The Web as a Knowledge-base for Answering Complex Questions. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies
Alon Talmor and Jonathan Berant. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W Cohen, Ruslan Salakhutdinov, and Christopher D Manning. 2018 · 2018
Earlier work this paper cites.
Episodic Memory in Lifelong Language Learning
Cyprien de Masson D’Autume, Sebastian Ruder, Lingpeng Kong, et al · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, et al · 2019
Earlier work this paper cites.
CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers)
Alon Talmor, Jonathan Herzig, Nicholas Lourie, and Jonathan Berant. 2019 · 2019
Earlier work this paper cites.
Structured event memory: A neuro-symbolic model of event cognition
Nicholas T Franklin, Kenneth A Norman, Charan Ranganath, Jeffrey M Zacks, and Samuel J Gershman. 2020 · 2020
Earlier work this paper cites.
INSPIRED: Toward Sociable Recommendation Dialog Systems. In Proceedings of the EMNLP
Shirley Anugrah Hayati, Dongyeop Kang, Qingxiaoyang Zhu, Weiyan Shi, and Zhou Yu. 2020 · 2020
Earlier work this paper cites.
Constructing a multi-hop qa dataset for comprehensive evaluation of reasoning steps
Xanh Ho, Anh-Khoa Nguyen Duong, Saku Sugawara, and Akiko Aizawa. 2020 · 2020
Earlier work this paper cites.
Compressive Transformers for Long-Range Sequence Modelling. In International Conference on Learning Representations
Jack W. Rae, Anna Potapenko, Siddhant M. Jayakumar, Chloe Hillier, and Timothy P. Lillicrap. 2020 · 2020
Earlier work this paper cites.
Towards scalable multi-domain conversational agents: The schema-guided dialogue dataset. In AAAI Conference on Artificial Intelligence
Abhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta, and Pranav Khaitan. 2020 · 2020
Earlier work this paper cites.
Brain power
Vijay Balasubramanian. 2021 · 2021
Earlier work this paper cites.
HybridQA: A Dataset of Multi-Hop Question Answering over Tabular and Textual Data. In Proceedings of the International Conference on Learning Representations (ICLR)
Wenhu Chen, Zhihao He, Yu Su, Yunyao Yu, William Wang, and Xifeng Yan. 2021 · 2021
Earlier work this paper cites.
Editing Factual Knowledge in Language Models. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing
Nicola De Cao, Wilker Aziz, and Ivan Titov. 2021 · 2021
Earlier work this paper cites.
Efficient Attentions for Long Document Summarization. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies
Luyang Huang, Shuyang Cao, Nikolaus Parulian, Heng Ji, and Lu Wang. 2021 · 2021
Earlier work this paper cites.
Unsupervised dense information retrieval with contrastive learning
Gautier Izacard, Mathilde Caron, Lucas Hosseini, Sebastian Riedel, Piotr Bojanowski, Armand Joulin, and Edouard Grave. 2021 · 2021
Earlier work this paper cites.
Learning Transferable Visual Models From Natural Language Supervision. In International Conference on Machine Learning
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Earlier work this paper cites.
Niket Tandon, Aman Madaan, Peter Clark, and Yiming Yang. 2021 · 2021
Earlier work this paper cites.
Long Range Arena : A Benchmark for Efficient Transformers. In International Conference on Learning Representations
Yi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen, Dara Bahri, Philip Pham, Jinfeng Rao, Liu Yang, Sebastian Ruder, and Donald Metzler. 2021 · 2021
Earlier work this paper cites.
Dual-system episodic control: Integrating episodic memory and reinforcement learning
Jane X Wang, Zeb Kurth-Nelson, Dharshan Kumaran, Dhruva Tirumala, Hubert Soyer, Joel Z Leibo, Demis Hassabis, and Matthew Botvinick. 2021 · 2021
Earlier work this paper cites.
Beyond Goldfish Memory: Long-term Open-domain Conversation
Jing Xu, Arthur Szlam, and Jason Weston. 2021 · 2021
Earlier work this paper cites.
Keep Me Updated! Memory Management in Long-term Conversations. In Findings of the Association for Computational Linguistics: EMNLP 2022
Sanghwan Bae, Donghyun Kwak, Soyoung Kang, Min Young Lee, Sungdong Kim, Yuin Jeong, Hyeri Kim, Sang-Woo Lee, Woomyoung Park, and Nako Sung. 2022 · 2022
Earlier work this paper cites.
LangChain
Harrison Chase. 2022 · 2022
Earlier work this paper cites.
Towards Teachable Reasoning Systems: Using a Dynamic Memory of User Feedback for Continual System Improvement. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing
Bhavana Dalvi Mishra, Oyvind Tafjord, and Peter Clark. 2022 · 2022
Earlier work this paper cites.
Calibrating Factual Knowledge in Pretrained Language Models. In Findings of the Association for Computational Linguistics: EMNLP 2022 . 5937–5947
Qingxiu Dong, Damai Dai, Yifan Song, Jingjing Xu, Zhifang Sui, and Lei Li. 2022 · 2022
Earlier work this paper cites.
Dynamic Global Memory for Document-level Argument Extraction. In Proceedings of the ACL
Xinya Du, Sha Li, and Heng Ji. 2022 · 2022
Earlier work this paper cites.
PerKGQA: Question answering over personalized knowledge graphs. In Findings of the Association for Computational Linguistics: NAACL 2022
Ritam Dutt, Kasturi Bhattacharjee, Rashmi Gangadharaiah, Dan Roth, and Carolyn Rose. 2022 · 2022
Earlier work this paper cites.
Mechanisms of memory updating: State dependency vs. reconsolidation
Christopher Kiley and Colleen M Parks. 2022 · 2022
Earlier work this paper cites.
LlamaIndex
Jerry Liu. 2022 · 2022
Earlier work this paper cites.
Incremental prompting: Episodic memory prompt for lifelong event detection
Minqian Liu, Shiyu Chang, and Lifu Huang. 2022 · 2022
Earlier work this paper cites.
UniTranSeR: A Unified Transformer Semantic Representation Framework for Multimodal Task-Oriented Dialog System. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Zhiyuan Ma, Jianjun Li, Guohui Li, and Yongjing Cheng. 2022 · 2022
Earlier work this paper cites.
DSI++: Updating transformer memory with new documents
Sanket Vaibhav Mehta, Jai Gupta, Yi Tay, Mostafa Dehghani, Vinh Q Tran, Jinfeng Rao, Marc Najork, Emma Strubell, and Donald Metzler. 2022 · 2022
Earlier work this paper cites.
Locating and Editing Factual Associations in GPT
Kevin Meng, David Bau, Alex Andonian, et al · 2022
Earlier work this paper cites.
Mass-editing Memory in a Transformer
Kevin Meng, Arnab Sen Sharma, Alex Andonian, et al · 2022
Earlier work this paper cites.
Capturing Global Structural Information in Long Document Question Answering with Compressive Graph Selector Network. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing
Yuxiang Nie, Heyan Huang, Wei Wei, and Xian-Ling Mao. 2022 · 2022
Earlier work this paper cites.
UniK-QA: Unified Representations of Structured and Unstructured Knowledge for Open-Domain Question Answering. In Findings of the Association for Computational Linguistics: NAACL 2022
Barlas Oguz, Xilun Chen, Vladimir Karpukhin, Stan Peshterliev, Dmytro Okhonko, Michael Schlichtkrull, Sonal Gupta, Yashar Mehdad, and Scott Yih. 2022 · 2022
Earlier work this paper cites.
Chatgpt: Optimizing language models for dialogue
OpenAI. 2022 · 2022
Earlier work this paper cites.
Memory replay with data compression for continual learning
Liyuan Wang, Xingxing Zhang, Kuo Yang, Longhui Yu, Chongxuan Li, Lanqing Hong, Shifeng Zhang, Zhenguo Li, Yi Zhong, and Jun Zhu. 2022 · 2022
Earlier work this paper cites.
Yuhuai Wu, Markus N Rabe, DeLesley Hutchins, et al · 2022
Earlier work this paper cites.
An Efficient Memory-Augmented Transformer for Knowledge-Intensive NLP Tasks. In Proceedings of the EMNLP
Yuxiang Wu, Yu Zhao, Baotian Hu, Pasquale Minervini, Pontus Stenetorp, and Sebastian Riedel. 2022b · 2022
Earlier work this paper cites.
Long Time No See! Open-domain Conversation with Long-term Persona Memory
Xinchao Xu, Zhibin Gou, Wenquan Wu, et al · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Earlier work this paper cites.
Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers. In NeurIPS
Sotiris Anagnostidis, Dario Pavllo, and Luca et al. Biggio. 2023 · 2023
Earlier work this paper cites.
Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Jinze Bai, Shuai Bai, Shusheng Yang, Shijie Wang, Sinan Tan, Peng Wang, Junyang Lin, Chang Zhou, and Jingren Zhou. 2023 · 2023
Earlier work this paper cites.
Unlearn What You Want to Forget: Efficient Unlearning for LLMs. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 12041–12052
Jiaao Chen and Diyi Yang. 2023 · 2023
Earlier work this paper cites.
Adapting Language Models to Compress Contexts. In Proceedings of the EMNLP
Alexis Chevalier, Alexander Wettig, Anirudh Ajith, and Danqi Chen. 2023 · 2023
Earlier work this paper cites.
LongNet: Scaling Transformers to 1,000,000,000 Tokens
Jiayu Ding, Shuming Ma, Li Dong, Xingxing Zhang, Shaohan Huang, Wenhui Wang, Nanning Zheng, and Furu Wei. 2023 · 2023
Earlier work this paper cites.
PaLM-E: An Embodied Multimodal Language Model
Danny Driess, Fei Xia, Mehdi SM Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Vuong, Tianhe Yu, et al · 2023
Earlier work this paper cites.
CogAgent: A Visual Language Model for GUI Agents
Wenyi Hong, Weihan Wang, Qingsong Lv, Jiazheng Xu, Wenmeng Yu, Junhui Ji, Yan Wang, Zihan Wang, Yuxuan Zhang, Juanzi Li, Bin Xu, Yuxiao Dong, Ming Ding, and Jie Tang. 2023 · 2023
Earlier work this paper cites.
Groundnlq@ ego4d natural language queries challenge 2023
Zhijian Hou, Lei Ji, Difei Gao, Wanjun Zhong, Kun Yan, Chao Li, Wing-Kwong Chan, Chong-Wah Ngo, Nan Duan, and Mike Zheng Shou. 2023 · 2023
Earlier work this paper cites.
ChatDB: Augmenting LLMs with Databases as Their Symbolic Memory
Chenxu Hu, Jie Fu, Chenzhuang Du, Simian Luo, Junbo Zhao, and Hang Zhao. 2023 · 2023
Earlier work this paper cites.
Learning Retrieval Augmentation for Personalized Dialogue Generation. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
Qiushi Huang, Shuai Fu, Xubo Liu, Wenwu Wang, Tom Ko, Yu Zhang, and Lilian Tang. 2023a · 2023
Earlier work this paper cites.
Advancing transformer architecture in long-context large language models: A comprehensive survey
Yunpeng Huang, Jingwei Xu, Junyu Lai, Zixu Jiang, Taolue Chen, Zenan Li, Yuan Yao, Xiaoxing Ma, Lijuan Yang, Hao Chen, et al · 2023
Earlier work this paper cites.
Jihyoung Jang, Minseong Boo, and Hyounghun Kim. 2023a · 2023
Earlier work this paper cites.
Knowledge Unlearning for Mitigating Privacy Risks in Language Models. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Anna Rogers, Jordan Boyd-Graber, and Naoaki Okazaki (Eds.). Association for Computational Linguistics, Toronto, Canada, 14389–14408
Joel Jang, Dongkeun Yoon, Sohee Yang, Sungmin Cha, Moontae Lee, Lajanugen Logeswaran, and Minjoon Seo. 2023b · 2023
Earlier work this paper cites.
LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
Huiqiang Jiang, Qianhui Wu, Chin-Yew Lin, Yuqing Yang, and Lili Qiu. 2023a · 2023
Earlier work this paper cites.
Active Retrieval Augmented Generation. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
Zhengbao Jiang, Frank Xu, Luyu Gao, Zhiqing Sun, Qian Liu, Jane Dwivedi-Yu, Yiming Yang, Jamie Callan, and Graham Neubig. 2023b · 2023
Earlier work this paper cites.
Needle In A Haystack - Pressure Testing LLMs
Gregory Kamradt. 2023 · 2023
Earlier work this paper cites.
Efficient Memory Management for Large Language Model Serving with PagedAttention. In Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles
Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph E. Gonzalez, Hao Zhang, and Ion Stoica. 2023 · 2023
Earlier work this paper cites.
Learning to Reason and Memorize with Self-Notes
Jack Lanchantin, Shubham Toshniwal, and Jason et al. Weston. 2023 · 2023
Earlier work this paper cites.
MoT: Memory-of-Thought Enables ChatGPT to Self-Improve. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
Xiaonan Li and Xipeng Qiu. 2023 · 2023
Earlier work this paper cites.
Compressing Context to Enhance Inference Efficiency of Large Language Models. In EMNLP
Yucheng Li, Bo Dong, and Frank et al. Guerin. 2023 · 2023
Earlier work this paper cites.
TCRA-LLM: Token Compression Retrieval Augmented Large Language Model for Inference Cost Reduction. In EMNLP Findings
Junyi Liu, Liangzhi Li, and Tong et al. Xiang. 2023d · 2023
Earlier work this paper cites.
RECAP: Retrieval-Enhanced Context-Aware Prefix Encoder for Personalized Dialogue Response Generation. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Shuai Liu, Hyundong Cho, Marjorie Freedman, Xuezhe Ma, and Jonathan May. 2023a · 2023
Earlier work this paper cites.
RECAP: retrieval-enhanced context-aware prefix encoder for personalized dialogue response generation
Shuai Liu, Hyundong J Cho, Marjorie Freedman, Xuezhe Ma, and Jonathan May. 2023b · 2023
Earlier work this paper cites.
Memochat: Tuning LLMs to Use Memos for Consistent Long-range Open-domain Conversation
Junru Lu, Siyu An, and Mingbao et al. Lin. 2023 · 2023
Earlier work this paper cites.
Generative replay inspired by hippocampal memory indexing for continual language learning. In Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics
Aru Maekawa, Hidetaka Kamigaito, Kotaro Funakoshi, and Manabu Okumura. 2023 · 2023
Earlier work this paper cites.
EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding. In Advances in Neural Information Processing Systems (NeurIPS)
Karttikeya Mangalam, Raiymbek Akshulakov, and Jitendra Malik. 2023 · 2023
Earlier work this paper cites.
Mass-Editing Memory in a Transformer. In The Eleventh International Conference on Learning Representations
Kevin Meng, Arnab Sen Sharma, Alex J Andonian, Yonatan Belinkov, and David Bau. 2023 · 2023
Earlier work this paper cites.
Conversation Understanding using Relational Temporal Graph Neural Networks with Auxiliary Cross-Modality Interaction. In EMNLP
Cam-Van Thi Nguyen, Anh-Tuan Mai, and The-Son et al. Le. 2023 · 2023
Earlier work this paper cites.
MemGPT: Towards LLMs as Operating Systems
Charles Packer, Vivian Fang, Shishir G Patil, et al · 2023
Earlier work this paper cites.
Generative Agents: Interactive Simulacra of Human Behavior. In Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology . 1–22
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2023 · 2023
Earlier work this paper cites.
In-context Unlearning: Language Models as Few-shot Unlearners
Martin Pawelczyk, Seth Neel, and Himabindu Lakkaraju. 2023 · 2023
Earlier work this paper cites.
Agent-OM: Leveraging LLM Agents for Ontology Matching
Zhangcheng Qiang, Weiqing Wang, and Kerry Taylor. 2023 · 2023
Earlier work this paper cites.
Lamp: When large language models meet personalization
Alireza Salemi, Sheshera Mysore, Michael Bendersky, and Hamed Zamani. 2023 · 2023
Earlier work this paper cites.
FlexGen: high-throughput generative inference of large language models with a single GPU. In International Conference on Machine Learning
Ying Sheng, Lianmin Zheng, Binhang Yuan, Zhuohan Li, Max Ryabinin, Beidi Chen, Percy Liang, Christopher Ré, Ion Stoica, and Ce Zhang. 2023 · 2023
Earlier work this paper cites.
Large Language Models Can Be Easily Distracted by Irrelevant Context. In International Conference on Machine Learning
Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed H. Chi, Nathanael Schärli, and Denny Zhou. 2023 · 2023
Earlier work this paper cites.
A unified approach to domain incremental learning with memory: Theory and algorithm
Haizhou Shi and Hao Wang. 2023 · 2023
Earlier work this paper cites.
Xin Su, Tiep Le, Steven Bethard, and Phillip Howard. 2023 · 2023
Earlier work this paper cites.
Enhancing Personalized Dialogue Generation with Contrastive Latent Variables: Combining Sparse and Dense Persona. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Anna Rogers, Jordan Boyd-Graber, and Naoaki Okazaki (Eds.). Association for Computational Linguistics, Toronto, Canada, 5456–5468
Yihong Tang, Bo Wang, Miao Fang, Dongming Zhao, Kun Huang, Ruifang He, and Yuexian Hou. 2023a · 2023
Earlier work this paper cites.
Enhancing Personalized Dialogue Generation with Contrastive Latent Variables: Combining Sparse and Dense Persona. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Yihong Tang, Bo Wang, Miao Fang, Dongming Zhao, Kun Huang, Ruifang He, and Yuexian Hou. 2023b · 2023
Earlier work this paper cites.
Focused Transformer: Contrastive Training for Context Scaling. In NeurIPS
Szymon Tworkowski, Konrad Staniszewski, and Mikołaj et al. Pacek. 2023 · 2023
Earlier work this paper cites.
KGA: A General Machine Unlearning Framework Based on Knowledge Gap Alignment
Lingzhi Wang, Tong Chen, Wei Yuan, et al · 2023
Earlier work this paper cites.
Resolving knowledge conflicts in large language models
Yike Wang, Shangbin Feng, Heng Wang, Weijia Shi, Vidhisha Balachandran, Tianxing He, and Yulia Tsvetkov. 2023b · 2023
Earlier work this paper cites.
LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos
Ying Wang, Yanlai Yang, and Mengye Ren. 2023c · 2023
Earlier work this paper cites.
DEPN: Detecting and Editing Privacy Neurons in Pretrained Language Models. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 2875–2886
Xinwei Wu, Junzhuo Li, Minghui Xu, Weilong Dong, Shuangzhi Wu, Chao Bian, and Deyi Xiong. 2023 · 2023
Earlier work this paper cites.
MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Zhiyang Xu, Ying Shen, and Lifu Huang. 2023 · 2023
Earlier work this paper cites.
ReAct: Synergizing Reasoning and Acting in Language Models. In The Eleventh International Conference on Learning Representations (ICLR)
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao. 2023 · 2023
Earlier work this paper cites.
TRAMS: Training-free Memory Selection for Long-range Language Modeling. In Findings of the Association for Computational Linguistics: EMNLP 2023
Haofei Yu, Cunxiang Wang, Yue Zhang, and Wei Bi. 2023 · 2023
Earlier work this paper cites.
NExT-Chat: An LMM for Chat, Detection and Segmentation
Ao Zhang, Yuan Yao, Wei Ji, Zhiyuan Liu, and Tat-Seng Chua. 2023b · 2023
Earlier work this paper cites.
Can We Edit Factual Knowledge by In-Context Learning?. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 4862–4876
Ce Zheng, Lei Li, Qingxiu Dong, Yuxuan Fan, Zhiyong Wu, Jingjing Xu, and Baobao Chang. 2023 · 2023
Earlier work this paper cites.
MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 15686–15702
Zexuan Zhong, Zhengxuan Wu, Christopher D Manning, Christopher Potts, and Danqi Chen. 2023 · 2023
Earlier work this paper cites.
Context-faithful Prompting for Large Language Models. In Findings of EMNLP
Wenxuan Zhou, Sheng Zhang, Hoifung Poon, and Muhao Chen. 2023 · 2023
Cited alongside, same era.
Keyformer: KV Cache reduction through key tokens selection for Efficient Generative Inference. In Proceedings of Machine Learning and Systems
Muhammad Adnan, Akhil Arunkumar, Gaurav Jain, Prashant J. Nair, Ilya Soloveychik, and Purushotham Kamath. 2024 · 2024
Cited alongside, same era.
CHAI: Clustered Head Attention for Efficient LLM Inference. In ICML
Saurabh Agarwal, Bilge Acun, and Basil et al. Hosmer. 2024 · 2024
Cited alongside, same era.
L-Eval: Instituting Standardized Evaluation for Long Context Language Models. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Chenxin An, Shansan Gong, Ming Zhong, Xingjian Zhao, Mukai Li, Jun Zhang, Lingpeng Kong, and Xipeng Qiu. 2024a · 2024
Cited alongside, same era.
Make Your LLM Fully Utilize the Context. In Advances in Neural Information Processing Systems
Large language model unlearning
Yuanshun Yao, Xiaojun Xu, and Yang Liu. 2024b · 2024
Later among the works it cites.
Long-Context Language Modeling with Parallel Context Encoding. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Howard Yen, Tianyu Gao, and Danqi Chen. 2024 · 2024
Later among the works it cites.
CompAct: Compressing Retrieved Documents Actively for Question Answering. In EMNLP
Chanwoong Yoon, Taewhoo Lee, and Hyeon et al. Hwang. 2024 · 2024
Later among the works it cites.
KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches. In Findings of EMNLP . 4623–4648
Jiayi Yuan, Hongyi Liu, Shaochen Zhong, Yu-Neng Chuang, Songchen Li, Guanchu Wang, Duy Le, Hongye Jin, Vipin Chaudhary, Zhaozhuo Xu, Zirui Liu, and Xia Hu. 2024 · 2024
Later among the works it cites.
FragRel: Exploiting Fragment-level Relations in the External Memory of Large Language Models. In Findings of the Association for Computational Linguistics: ACL 2024 , Lun-Wei Ku, Andre Martins, and Vivek Srikumar (Eds.). Association for Computational Linguistics, Bangkok, Thailand, 16348–16361
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shengnan An, Zexiong Ma, Zeqi Lin, Nanning Zheng, Jian-Guang Lou, and Weizhu Chen. 2024b · 2024
Cited alongside, same era.
Introducing the Model Context Protocol
Anthropic. 2024 · 2024
Cited alongside, same era.
LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Yushi Bai, Xin Lv, Jiajie Zhang, Hongchang Lyu, Jiankai Tang, Zhidian Huang, Zhengxiao Du, Xiao Liu, Aohan Zeng, Lei Hou, Yuxiao Dong, Jie Tang, and Juanzi Li. 2024 · 2024
Cited alongside, same era.
Titans: Learning to memorize at test time
Ali Behrouz, Peilin Zhong, and Vahab Mirrokni. 2024 · 2024
Cited alongside, same era.
Memory Layers at Scale
Vincent-Pierre Berges, Barlas Oğuz, Daniel Haziza, Wen-tau Yih, Luke Zettlemoyer, and Gargi Ghosh. 2024 · 2024
Cited alongside, same era.
AWESOME: GPU Memory-constrained Long Document Summarization using Memory Mechanism and Global Salient Content. In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers)
Shuyang Cao and Lu Wang. 2024 · 2024
Cited alongside, same era.
Retaining Key Information under High Compression Ratios: Query-Guided Compressor for LLMs. In Proceedings of the ACL
Zhiwei Cao, Qian Cao, Yu Lu, Ningxin Peng, Luyang Huang, Shanbo Cheng, and Jinsong Su. 2024 · 2024
Cited alongside, same era.
Improving Factuality with Explicit Working Memory
Mingda Chen, Yang Li, Karthik Padthe, et al · 2024
Cited alongside, same era.
Xihang Yue, Linchao Zhu, and Yi Yang. 2024 · 2024
Later among the works it cites.
LLM-based Medical Assistant Personalization with Short- and Long-Term Memory Coordination. In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers)
Kai Zhang, Yangyang Kang, Fubang Zhao, and Xiaozhong Liu. 2024d · 2024
Later among the works it cites.
A Comprehensive Study of Knowledge Editing for Large Language Models
Ningyu Zhang, Yunzhi Yao, Bozhong Tian, et al · 2024
Later among the works it cites.
DAFNet: Dynamic Auxiliary Fusion for Sequential Model Editing in Large Language Models. In Findings of the Association for Computational Linguistics: ACL 2024 . 1588–1602
Taolin Zhang, Qizhou Chen, Dongyang Li, Chengyu Wang, Xiaofeng He, Longtao Huang, Jun Huang, et al · 2024
Later among the works it cites.
∞ \infty Bench: Extending Long Context Evaluation Beyond 100K Tokens. In ACL . 15262–15277
Xinrong Zhang, Yingfa Chen, Shengding Hu, Zihang Xu, Junhao Chen, Moo Hao, Xu Han, Zhen Thai, Shuo Wang, Zhiyuan Liu, and Maosong Sun. 2024b · 2024
Later among the works it cites.
A survey on the memory mechanism of large language model based agents
Zeyu Zhang, Xiaohe Bo, Chen Ma, Rui Li, Xu Chen, Quanyu Dai, Jieming Zhu, Zhenhua Dong, and Ji-Rong Wen. 2024a · 2024
Later among the works it cites.
DIVKNOWQA: Assessing the Reasoning Ability of LLMs via Open-Domain Question Answering over Knowledge Base and Text. In Findings of the Association for Computational Linguistics: NAACL 2024
Wenting Zhao, Ye Liu, Tong Niu, Yao Wan, Philip S. Yu, Shafiq Joty, Yingbo Zhou, and Semih Yavuz. 2024b · 2024
Later among the works it cites.
Atom: Low-Bit Quantization for Efficient and Accurate LLM Serving. In MLSys
Yilong Zhao, Chien-Yu Lin, Kan Zhu, Zihao Ye, Lequn Chen, Size Zheng, Luis Ceze, Arvind Krishnamurthy, Tianqi Chen, and Baris Kasikci. 2024a · 2024
Later among the works it cites.
Longtao Zheng, Rundong Wang, Xinrun Wang, and Bo An. 2024 · 2024
Later among the works it cites.
MemoryBank: Enhancing Large Language Models with Long-term Memory. In Proceedings of the AAAI Conference on Artificial Intelligence
Wanjun Zhong, Lianghong Guo, Qiqi Gao, et al · 2024
Later among the works it cites.
VISTA: Visualized Text Embedding For Universal Multi-Modal Retrieval. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Junjie Zhou, Zheng Liu, Shitao Xiao, Bo Zhao, and Yongping Xiong. 2024 · 2024
Later among the works it cites.
MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems
Qingyao Ai, Yichen Tang, Changyue Wang, Jianming Long, Weihang Su, and Yiqun Liu. 2025 · 2025
Closest in time.
Claude: AI Thinking Partner
Anthropic. 2023 · 2025
Closest in time.
Equipping agents for the real world with Agent Skills
Anthropic. 2025 · 2025
Closest in time.
Siri: Apple’s Intelligent Voice Assistant
Apple Inc. 2025 · 2025
Closest in time.
LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks
Yushi Bai, Shangqing Tu, Jiajie Zhang, Hao Peng, Xiaozhi Wang, Xin Lv, Shulin Cao, Jiazheng Xu, Lei Hou, Yuxiao Dong, Jie Tang, and Juanzi Li. 2025 · 2025
Closest in time.
Open Problems in Machine Unlearning for AI Safety
Fazl Barez, Tingchen Fu, Ameya Prabhu, Stephen Casper, Amartya Sanyal, Adel Bibi, Aidan O’Gara, Robert Kirk, Ben Bucknall, Tim Fist, Luke Ong, Philip Torr, Kwok-Yan Lam, Robert Trager, David Krueger, Sören Mindermann, José Hernandez-Orallo, Mor Geva, and Yarin Gal. 2025 · 2025
Closest in time.
RocketKV: Accelerating Long-Context LLM Inference via Two-Stage KV Cache Compression. In Forty-second International Conference on Machine Learning
Payman Behnam, Yaosheng Fu, Ritchie Zhao, Po-An Tsai, Zhiding Yu, and Alexey Tumanov. 2025 · 2025
Closest in time.
Doubao-1.5-pro: A High-Efficiency Sparse MoE Multimodal AI Model
ByteDance Seed Team. 2025 · 2025
Closest in time.
Memory Decoder: A Pretrained, Plug-and-Play Memory for Large Language Models
Jiaqi Cao, Jiarui Wang, Rubin Wei, Qipeng Guo, Kai Chen, Bowen Zhou, and Zhouhan Lin. 2025 · 2025
Closest in time.
Towards Robust and Parameter-Efficient Knowledge Unlearning for LLMs. In The Thirteenth International Conference on Learning Representations
Sungmin Cha, Sungjun Cho, Dasol Hwang, and Moontae Lee. 2025 · 2025
Closest in time.
Character.AI: A Platform for Creating and Interacting with AI Characters
Character Technologies, Inc. 2023 · 2025
Closest in time.
HaluMem: Evaluating Hallucinations in Memory Systems of Agents
Ding Chen, Simin Niu, Kehang Li, Peng Liu, Xiangping Zheng, Bo Tang, Xinchi Li, Feiyu Xiong, and Zhiyu Li. 2025 · 2025
Closest in time.
Towards Scalable Exact Machine Unlearning Using Parameter-Efficient Fine-Tuning. In The Thirteenth International Conference on Learning Representations
Somnath Basu Roy Chowdhury, Krzysztof Marcin Choromanski, Arijit Sehanobish, Kumar Avinava Dubey, and Snigdha Chaturvedi. 2025 · 2025
Closest in time.
Sft memorizes, rl generalizes: A comparative study of foundation model post-training
Tianzhe Chu, Yuexiang Zhai, Jihan Yang, Shengbang Tong, Saining Xie, Dale Schuurmans, Quoc V Le, Sergey Levine, and Yi Ma. 2025 · 2025
Closest in time.
CodeBuddy: AI-Powered Coding Assistant
Codebuddy AI Inc. 2025 · 2025
Closest in time.
Coze: Build your own AI agent
Coze. 2024 · 2025
Closest in time.
Everything is Editable: Extend Knowledge Editing to Unstructured Data in Large Language Models. In The Thirteenth International Conference on Learning Representations
Jingcheng Deng, Zihao Wei, Liang Pang, Hanxing Ding, Huawei Shen, and Xueqi Cheng. 2025 · 2025
Closest in time.
Streaming Video Question-Answering with In-context Video KV-Cache Retrieval. In ICLR
Shangzhe Di, Zhelun Yu, and Guanghao et al. Zhang. 2025 · 2025
Closest in time.
Unified Parameter-Efficient Unlearning for LLMs. In The Thirteenth International Conference on Learning Representations
Chenlu Ding, Jiancan Wu, Yancheng Yuan, Jinda Lu, Kai Zhang, Alex Su, Xiang Wang, and Xiangnan He. 2025 · 2025
Closest in time.
MemGuide: Intent-Driven Memory Selection for Goal-Oriented Multi-Session LLM Agents
Yiming Du, Bingbing Wang, Yang He, Bin Liang, Baojun Wang, Zhongyang Li, Lin Gui, Jeff Z. Pan, Ruifeng Xu, and Kam-Fai Wong. 2025 · 2025
Closest in time.
Memp: Exploring Agent Procedural Memory
Runnan Fang, Yuan Liang, Xiaobin Wang, Jialong Wu, Shuofei Qiao, Pengjun Xie, Fei Huang, Huajun Chen, and Ningyu Zhang. 2025b · 2025
Closest in time.
Not All Heads Matter: A Head-Level KV Cache Compression Method with Integrated Retrieval and Reasoning. In International Conference on Learning Representations
Yu Fu, Zefan Cai, Abedelkadir Asi, Wayne Xiong, Yue Dong, and Wen Xiao. 2025 · 2025
Closest in time.
Yubin Ge, Salvatore Romeo, Jason Cai, Raphael Shu, Monica Sunkara, Yassine Benajiba, and Yi Zhang. 2025 · 2025
Closest in time.
GitHub Copilot: Your AI pair programmer
GitHub and OpenAI. 2021 · 2025
Closest in time.
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics
Marc Glocker, Peter Hönig, Matthias Hirschmanner, and Markus Vincze. 2025 · 2025
Closest in time.
Gemini: Multimodal AI Assistant
Google AI / DeepMind. 2024 · 2025
Closest in time.
Towards Lifelong Model Editing via Simulating Ideal Editor. In Forty-second International Conference on Machine Learning
Yaming Guo, Siyang Guo, Hengshu Zhu, and Ying Sun. 2025 · 2025
Closest in time.
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models
Bernal Jiménez Gutiérrez, Yiheng Shu, Weijian Qi, Sizhe Zhou, and Yu Su. 2025 · 2025
Closest in time.
Concept-Reversed Winograd Schema Challenge: Evaluating and Improving Robust Reasoning in Large Language Models via Abstraction. In Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 2: Short Papers) . Association for Computational Linguistics
Kaiqiao Han, Tianqing Fang, Zhaowei Wang, Yangqiu Song, and Mark Steedman. [n. d.] · 2025
Closest in time.
Radar: Fast Long-Context Decoding for Any Transformer. In The Thirteenth International Conference on Learning Representations
Yongchang Hao, Mengyao Zhai, Hossein Hajimirsadeghi, Sepidehsadat Hosseini, and Frederick Tung. 2025 · 2025
Closest in time.
Graphiti: Bridging Graph and Relational Database Queries
Yang He, Ruijie Fang, Isil Dillig, and Yuepeng Wang. 2025 · 2025
Closest in time.
Decoupling Memories, Muting Neurons: Towards Practical Machine Unlearning for Large Language Models. In Findings of the Association for Computational Linguistics: ACL 2025 , Wanxiang Che, Joyce Nabende, Ekaterina Shutova, and Mohammad Taher Pilehvar (Eds.). Association for Computational Linguistics, Vienna, Austria, 13978–13999
Lishuai Hou, Zixiong Wang, Gaoyang Liu, Chen Wang, Wei Liu, and Kai Peng. 2025 · 2025
Closest in time.
Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
Yuanzhe Hu, Yu Wang, and Julian McAuley. 2025 · 2025
Closest in time.
Masking in Multi-hop QA: An Analysis of How Language Models Perform with Context Permutation
Wenyu Huang, Pavlos Vougiouklis, Mirella Lapata, and Jeff Z. Pan. 2025 · 2025
Closest in time.
Xiaoyi: Huawei’s AI Smart Assistant
Huawei Technologies Co., Ltd. 2025 · 2025
Closest in time.
Cursor: An AI-Powered Code Editor
Anysphere Inc. 2024 · 2025
Closest in time.
LangGraph: Build Resilient Language Agents as Graphs
LangChain Inc. 2025 · 2025
Closest in time.
SEPS: A Separability Measure for Robust Unlearning in LLMs. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, and Violet Peng (Eds.). Association for Computational Linguistics, Suzhou, China, 5556–5587
Wonje Jeung, Sangyeon Yoon, and Albert No. 2025 · 2025
Closest in time.
Bowen Jiang, Yuan Yuan, Maohao Shen, Zhuoqun Hao, Zhangchen Xu, Zichen Chen, Zijun Liu, Anirudh Ravi Vijjini, Jiaming He, et al · 2025
Closest in time.
Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG. In The Thirteenth International Conference on Learning Representations
Bowen Jin, Jinsung Yoon, Jiawei Han, and Sercan O Arik. 2025 · 2025
Closest in time.
You Only Read Once (YORO): Learning to Internalize Database Knowledge for Text-to-SQL. In Proceedings of NAACL-HLT (Long Papers) . 1889–1901
Hideo Kobayashi, Wuwei Lan, Peng Shi, Shuaichen Chang, Jiang Guo, Henghui Zhu, Zhiguo Wang, and Patrick Ng. 2025 · 2025
Closest in time.
STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning
Mingcong Lei, Yiming Zhao, Ge Wang, Zhixin Mai, Shuguang Cui, Yatong Han, and Jinke Ren. 2025 · 2025
Closest in time.
Selective Attention Improves Transformer. In ICLR
Yaniv Leviathan, Matan Kalman, and Yossi Matias. 2025 · 2025
Closest in time.
Machine Unlearning: Taxonomy, Metrics, Applications, Challenges, and Prospects
Na Li, Chunyi Zhou, Yansong Gao, Hui Chen, Zhi Zhang, Boyu Kuang, and Anmin Fu. 2025b · 2025
Closest in time.
MemOS: An Operating System for Memory-Augmented Generation (MAG) in Large Language Models
Zhiyu Li, Shichao Song, Hanyu Wang, Simin Niu, Ding Chen, Jiawei Yang, Chenyang Xi, Huayi Lai, Jihao Zhao, Yezhaohui Wang, et al · 2025
Closest in time.
A Survey of Personalized Large Language Models: Progress and Future Directions
Jiahong Liu, Zexuan Qiu, Zhongyang Li, Quanyu Dai, Jieming Zhu, Minda Hu, Menglin Yang, and Irwin King. 2025a · 2025
Closest in time.
Echo: A Large Language Model with Temporal Episodic Memory
WenTao Liu, Ruohua Zhang, and Aimin et al. Zhou. 2025c · 2025
Closest in time.
Threats, Attacks, and Defenses in Machine Unlearning: A Survey
Ziyao Liu, Huanyi Ye, Chen Chen, Yongsen Zheng, and Kwok-Yan Lam. 2025b · 2025
Closest in time.
Seeing, Listening, Remembering, and Reasoning: A Multimodal Agent with Long-Term Memory
Lin Long, Yichen He, Wentao Ye, Yiyuan Pan, Yuan Lin, Hang Li, Junbo Zhao, and Wei Li. 2025 · 2025
Closest in time.
Replika: The AI companion who cares
Luka, Inc. 2025 · 2025
Closest in time.
Elias Lumer, Anmol Gulati, Vamse Kumar Subbiah, Pradeep Honaganahalli Basavaraju, and James A Burke. 2025 · 2025
Closest in time.
Memobase: Profile-Based Long-Term Memory for AI Applications
memodb io. 2025 · 2025
Closest in time.
Enhancing Reasoning with Collaboration and Memory
Julie Michelman, Nasrin Baratalipour, and Matthew Abueg. 2025 · 2025
Closest in time.
Me.bot: Your AI Second Brain
Mindverse AI. 2025 · 2025
Closest in time.
Fast Exact Unlearning for In-Context Learning Data for LLMs. In Forty-second International Conference on Machine Learning
Andrei Ioan Muresanu, Anvith Thudi, Michael R. Zhang, and Nicolas Papernot. 2025 · 2025
Closest in time.
Dynamic Retriever for In-Context Knowledge Editing via Policy Optimization. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, and Violet Peng (Eds.). Association for Computational Linguistics, Suzhou, China, 16755–16768
Mahmud Wasif Nafee, Maiqi Jiang, Haipeng Chen, and Yanfu Zhang. 2025 · 2025
Closest in time.
MemU: An open-source memory framework for AI companions
NevaMind-AI. 2025 · 2025
Closest in time.
Kai Tzu-iunn Ong, Namyoung Kim, Minju Gwak, Hyungjoo Chae, Taeyoon Kwon, Yohan Jo, Seung-won Hwang, Dongha Lee, and Jinyoung Yeo. 2025 · 2025
Closest in time.
OpenAI Platform Documentation: Embeddings Guide
OpenAI. 2025 · 2025
Closest in time.
ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory
Siru Ouyang, Jun Yan, I-Hung Hsu, Yanfei Chen, Ke Jiang, Zifeng Wang, Rujun Han, Long T. Le, Samira Daruki, Xiangru Tang, Vishy Tirumalashetty, George Lee, Mahsan Rofouei, Hangfei Lin, Jiawei Han, Chen-Yu Lee, and Tomas Pfister. 2025 · 2025
Closest in time.
Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents
Mathis Pink, Qinyuan Wu, Vy Ai Vo, Javier Turek, Jianing Mu, Alexander Huth, and Mariya Toneva. 2025 · 2025
Closest in time.
UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Yujia Qin, Yining Ye, Junjie Fang, Haoming Wang, Shihao Liang, Shizuo Tian, Junda Zhang, Jiahao Li, Yunxin Li, Shijue Huang, Wanjun Zhong, Kuanye Li, Jiale Yang, Yu Miao, Woyu Lin, Longxiang Liu, Xu Jiang, Qianli Ma, Jingyu Li, Xiaojun Xiao, Kai Cai, Chuang Li, Yaowei Zheng, Chaolin Jin, Chen Li, Xiao Zhou, Minchao Wang, Haoli Chen, Zhaojian Li, Haihua Yang, Haifeng Liu, Feng Lin, Tao Peng, Xin Liu, and Guang Shi. 2025 · 2025
Closest in time.
Zep: A Temporal Knowledge Graph Architecture for Agent Memory
Preston Rasmussen, Pavlo Paliychuk, Travis Beauvais, Jack Ryan, and Daniel Chalef. 2025 · 2025
Closest in time.
MemInsight: Autonomous Memory Augmentation for LLM Agents
Rana Salama, Jason Cai, and Michelle et al. Yuan. 2025 · 2025
Closest in time.
LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models. In Forty-second International Conference on Machine Learning
Dachuan Shi, Yonggan Fu, Xiangchi Yuan, Zhongzhi Yu, Haoran You, Sixu Li, Xin Dong, Jan Kautz, Pavlo Molchanov, and Yingyan Celine Lin. 2025 · 2025
Closest in time.
Tongyi DeepResearch: A New Era of Open-Source AI Researchers
Tongyi DeepResearch Team. 2025 · 2025
Closest in time.
ima.copilot: Intelligent Workbench Powered by Tencent’s Hunyuan Model
Tencent. 2025 · 2025
Closest in time.
Knowledge Editing through Chain-of-Thought. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, and Violet Peng (Eds.). Association for Computational Linguistics, Suzhou, China, 10684–10704
Changyue Wang, Weihang Su, Qingyao Ai, Yichen Tang, and Yiqun Liu. 2025e · 2025
Closest in time.
O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving Agents
Piaohong Wang, Motong Tian, Jiaxian Li, Yuan Liang, Yuqing Wang, Qianben Chen, Tiannan Wang, Zhicong Lu, Jiawei Ma, Yuchen Eleanor Jiang, and Wangchunshu Zhou. 2025g · 2025
Closest in time.
Recursively Summarizing Enables Long-Term Dialogue Memory in Large Language Models
Qingyue Wang, Yanan Fu, Yanan Cao, Shi Wang, Zhiliang Tian, and Liang Ding. 2025c · 2025
Closest in time.
Recursively summarizing enables long-term dialogue memory in large language models
Qingyue Wang, Yanhe Fu, Yanan Cao, Shuai Wang, Zhiliang Tian, and Liang Ding. 2025d · 2025
Closest in time.
Rongzheng Wang, Qizhi Chen, Yihong Huang, Yizhuo Ma, Muquan Li, Jiakai Li, Ke Qin, Guangchun Luo, and Shuang Liang. 2025a · 2025
Closest in time.
Mirix: Multi-agent memory system for llm-based agents
Yu Wang and Xi Chen. 2025 · 2025
Closest in time.
Mem- { \{ \ \backslash alpha } \} : Learning Memory Construction via Reinforcement Learning
Yu Wang, Ryuichi Takanobu, Zhiqi Liang, Yuzhen Mao, Yuanzhe Hu, Julian McAuley, and Xiaojian Wu. 2025f · 2025
Closest in time.
Episodic Memory Representation for Long-form Video Understanding
Yun Wang, Long Zhang, Jingren Liu, Jiaqi Yan, Zhanjie Zhang, Jiahao Zheng, Xun Yang, Dapeng Wu, Xiangyu Chen, and Xuelong Li. 2025k · 2025
Closest in time.
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
Zhaowei Wang, Wenhao Yu, Xiyu Ren, Jipeng Zhang, Yu Zhao, Rohit Saxena, Liang Cheng, Ginny Wong, Simon See, Pasquale Minervini, et al · 2025
Closest in time.
MLP Memory: Language Modeling with Retriever-pretrained External Memory
Rubin Wei, Jiaqi Cao, Jiarui Wang, Jushi Kai, Qipeng Guo, Bowen Zhou, and Zhouhan Lin. 2025 · 2025
Closest in time.
Interpersonal Memory Matters: A New Task for Proactive Dialogue Utilizing Conversational History
Bowen Wu, Wenqing Wang, and Haoran et al. Li. 2025b · 2025
Closest in time.
Kelle: Co-design KV Caching and eDRAM for Efficient LLM Serving in Edge Computing
Tianhua Xia and Sai Qian Zhang. 2025 · 2025
Closest in time.
WORLDMEM: Long-term Consistent World Simulation with Memory
Zeqi Xiao, Yushi Lan, and Yifan et al. Zhou. 2025 · 2025
Closest in time.
ReLearn: Unlearning via Learning for Large Language Models. In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Wanxiang Che, Joyce Nabende, Ekaterina Shutova, and Mohammad Taher Pilehvar (Eds.). Association for Computational Linguistics, Vienna, Austria, 5967–5987
Haoming Xu, Ningyuan Zhao, Liming Yang, Sendong Zhao, Shumin Deng, Mengru Wang, Bryan Hooi, Nay Oo, Huajun Chen, and Ningyu Zhang. 2025b · 2025
Closest in time.
A-MEM: Agentic Memory for LLM Agents
Wujiang Xu, Zujie Liang, Kai Mei, Hang Gao, Juntao Tan, and Yongfeng Zhang. 2025a · 2025
Closest in time.
Sikuan Yan, Xiufeng Yang, Zuchao Huang, Ercong Nie, Zifeng Ding, Zonggen Li, Xiaowen Ma, Hinrich Schütze, Volker Tresp, and Yunpu Ma. 2025 · 2025
Closest in time.
AgentFold: Long-Horizon Web Agents with Proactive Context Management
Rui Ye, Zhongwang Zhang, Kuan Li, Huifeng Yin, Zhengwei Tao, Yida Zhao, Liangcai Su, Liwen Zhang, Zile Qiao, Xinyu Wang, Pengjun Xie, Fei Huang, Siheng Chen, Jingren Zhou, and Yong Jiang. 2025 · 2025
Closest in time.
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits. In Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) , Luis Chiruzzo, Alan Ritter, and Lu Wang (Eds.). Association for Computational Linguistics, Albuquerque, New Mexico, 12656–12669
Paul Youssef, Zhixue Zhao, Jörg Schlötterer, and Christin Seifert. 2025 · 2025
Closest in time.
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Hongli Yu, Tinghong Chen, Jiangtao Feng, Jiangjie Chen, Weinan Dai, Qiying Yu, Ya-Qin Zhang, Wei-Ying Ma, Jingjing Liu, Mingxuan Wang, et al · 2025
Closest in time.
Context as memory: Scene-consistent interactive long video generation with memory retrieval
Jiwen Yu, Jianhong Bai, Yiran Qin, Quande Liu, Xintao Wang, Pengfei Wan, Di Zhang, and Xihui Liu. 2025a · 2025
Closest in time.
G-Memory: Tracing Hierarchical Memory for Multi-Agent Systems
Guibin Zhang, Muxin Fu, Guancheng Wan, Miao Yu, Kun Wang, and Shuicheng Yan. 2025b · 2025
Closest in time.
Agent learning via early experience
Kai Zhang, Xiangchao Chen, Bo Liu, Tianci Xue, Zeyi Liao, Zhihan Liu, Xiyao Wang, Yuting Ning, Zhaorun Chen, Xiaohan Fu, et al · 2025
Closest in time.
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Qizheng Zhang, Changran Hu, Shubhangi Upasani, Boyuan Ma, Fenglu Hong, Vamsidhar Kamanuru, Jay Rainton, Chen Wu, Mengmeng Ji, Hanchen Li, et al · 2025
Closest in time.
LightMem: Lightweight and Efficient Memory-Augmented Generation
Qingyang Zhang, Ningyu Zhang, et al · 2025
Closest in time.
Self-Memory Alignment: Mitigating Factual Hallucinations with Generalized Improvement
Siyuan Zhang, Yichi Zhang, and Yinpeng et al. Dong. 2025e · 2025
Closest in time.
EventWeave: A Dynamic Framework for Capturing Core and Supporting Events in Dialogue Systems
Zhengyi Zhao, Shubo Zhang, Yiming Du, Bin Liang, Baojun Wang, Zhongyang Li, Binyang Li, and Kam-Fai Wong. 2025 · 2025
Closest in time.
AgentFly: Fine-tuning LLM Agents without Fine-tuning LLMs
Huichi Zhou, Yihang Chen, Siyuan Guo, Xue Yan, Kin Hei Lee, Zihan Wang, Ka Yiu Lee, Guchun Zhang, Kun Shao, Linyi Yang, et al · 2025
Closest in time.
M2Edit: Locate and Edit Multi-Granularity Knowledge in Multimodal Large Language Model. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, and Violet Peng (Eds.). Association for Computational Linguistics, Suzhou, China, 29017–29030
Yang Zhou, Pengfei Cao, Yubo Chen, Qingbin Liu, Dianbo Sui, Xi Chen, Kang Liu, and Jun Zhao. 2025a · 2025
Closest in time.
MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents
Zijian Zhou, Ao Qu, Zhaoxuan Wu, Sunghwan Kim, Alok Prakash, Daniela Rus, Jinhua Zhao, Bryan Kian Hsiang Low, and Paul Pu Liang. 2025c · 2025
Closest in time.
Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection. In The Thirteenth International Conference on Learning Representations
Yun Zhu, Jia-Chen Gu, Caitlin Sikora, Ho Ko, Yinxiao Liu, Chu-Cheng Lin, Lei Shu, Liangchen Luo, Lei Meng, Bang Liu, and Jindong Chen. 2025 · 2025
Closest in time.
Latent Collaboration in Multi-Agent Systems
Jiaru Zou, Xiyuan Yang, Ruizhong Qiu, Gaotang Li, Katherine Tieu, Pan Lu, Ke Shen, Hanghang Tong, Yejin Choi, Jingrui He, James Zou, Mengdi Wang, and Ling Yang. 2025 · 2025
Closest in time.
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022
2027
Closest in time.