Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have made significant progress in open-ended dialogue, yet their inability to retain and retrieve relevant information from long-term interactions limits their effectiveness in applications requiring sustained personalization.
Statistical Theory of Extreme Values and Some Practical Applications: A Series of Lectures
E. Gumbel · 1954
Earlier work this paper cites.
Eliza—a computer program for the study of natural language communication between man and machine
J. Weizenbaum · 1966
Earlier work this paper cites.
The process of retrieval from very long-term memory
M. D. Williams and J. D. Hollan · 1981
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
M. McCloskey and N. J. Cohen · 1989
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
PARADISE: A framework for evaluating spoken dialogue agents
M. A. Walker, D. J. Litman, C. A. Kamm, and A. Abella · 1997
Earlier work this paper cites.
Managing long term communications: conversation and contact management
S. Whittaker, Q. Jones, and L. Terveen · 2002
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
S. Banerjee and A. Lavie · 2005
Earlier work this paper cites.
A* sampling
C. J. Maddison, D. Tarlow, and T. Minka · 2014
Earlier work this paper cites.
Categorical reparameterization with gumbel-softmax
E. Jang, S. Gu, and B. Poole · 2017
Earlier work this paper cites.
Bertscore: Evaluating text generation with BERT
T. Zhang, V. Kishore, F. Wu, K. Q. Weinberger, and Y. Artzi · 2020
Earlier work this paper cites.
A cooperative memory network for personalized task-oriented dialogue systems with incomplete user profiles
J. Pei, P. Ren, and M. de Rijke · 2021
Earlier work this paper cites.
Keep me updated! memory management in long-term conversations
S. Bae, D. Kwak, S. Kang, M. Y. Lee, S. Kim, Y. Jeong, H. Kim, S.-W. Lee, W. Park, and N. Sung · 2022
Earlier work this paper cites.
Unsupervised dense information retrieval with contrastive learning
G. Izacard, M. Caron, L. Hosseini, S. Riedel, P. Bojanowski, A. Joulin, and E. Grave · 2022
Earlier work this paper cites.
Beyond goldfish memory: Long-term open-domain conversation
J. Xu, A. Szlam, and J. Weston · 2022
Earlier work this paper cites.
Optimizing natural language processing, large language models (llms) for efficient customer service, and hyper-personalization to enable sustainable growth and revenue
S. Kolasani · 2023
Earlier work this paper cites.
Prompted LLMs as chatbot modules for long open-domain conversation
G. Lee, V. Hartmann, J. Park, D. Papailiopoulos, and K. Lee · 2023
Cited alongside, same era.
Towards general text embeddings with multi-stage contrastive learning
Z. Li, X. Zhang, Y. Zhang, D. Long, P. Xie, and M. Zhang · 2023
Cited alongside, same era.
Memochat: Tuning llms to use memos for consistent long-range open-domain conversation
J. Lu, S. An, M. Lin, G. Pergola, Y. He, D. Yin, X. Sun, and Y. Wu · 2023
Cited alongside, same era.
Large language models can be easily distracted by irrelevant context
F. Shi, X. Chen, K. Misra, N. Scales, D. Dohan, E. H. Chi, N. Schärli, and D. Zhou · 2023
Cited alongside, same era.
Recursively summarizing enables long-term dialogue memory in large language models
Q. Wang, L. Ding, Y. Cao, Z. Tian, S. Wang, D. Tao, and L. Guo · 2023
Ai for education (ai4edu): Advancing personalized education with llm and adaptive learning
Q. Wen, J. Liang, C. Sierra, R. Luckin, R. Tong, Z. Liu, P. Cui, and J. Tang · 2024
Later among the works it cites.
Length extrapolation of transformers: A survey from the perspective of positional encoding
L. Zhao, X. Feng, X. Feng, W. Zhong, D. Xu, Q. Yang, H. Liu, B. Qin, and T. Liu · 2024
Later among the works it cites.
Dape: Data-adaptive positional encoding for length extrapolation
C. Zheng, Y. Gao, H. Shi, M. Huang, J. Li, J. Xiong, X. Ren, M. Ng, X. Jiang, Z. Li, and Y. Li · 2024
Later among the works it cites.
Memorybank: Enhancing large language models with long-term memory
W. Zhong, L. Guo, Q. Gao, H. Ye, and Y. Wang · 2024
Later among the works it cites.
Mem0: Building production-ready ai agents with scalable long-term memory, 2025
P. Chhikara, D. Khant, S. Aryan, T. Singh, and D. Yadav · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Attribute or abstain: Large language models as long document assistants
J. Buchmann, X. Liu, and I. Gurevych · 2024
Cited alongside, same era.
When large language models meet personalization: perspectives of challenges and opportunities
J. Chen, Z. Liu, X. Huang, C. Wu, Q. Liu, G. Jiang, Y. Pu, Y. Lei, X. Chen, X. Wang, K. Zheng, D. Lian, and E. Chen · 2024
Cited alongside, same era.
Can LLM be a personalized judge?
Y. R. Dong, T. Hu, and N. Collier · 2024
Cited alongside, same era.
Intelligent agents with llm-based process automation
Y. Guan, D. Wang, Z. Chu, S. Wang, F. Ni, R. Song, and C. Zhuang · 2024
Cited alongside, same era.
Grounding and evaluation for large language models: Practical challenges and lessons learned (survey)
K. Kenthapadi, M. Sameki, and A. Taly · 2024
Cited alongside, same era.
Lost in the middle: How language models use long contexts
N. F. Liu, K. Lin, J. Hewitt, A. Paranjape, M. Bevilacqua, F. Petroni, and P. Liang · 2024
Cited alongside, same era.
Evaluating very long-term conversational memory of LLM agents
A. Maharana, D.-H. Lee, S. Tulyakov, M. Bansal, F. Barbieri, and Y. Fang · 2024
Cited alongside, same era.
Retrieve, summarize, plan: Advancing multi-hop question answering with an iterative approach
Z. Jiang, M. Sun, L. Liang, and Z. Zhang · 2025
Closest in time.
Hello again! LLM-powered personalized agent for long-term dialogue
H. Li, C. Yang, A. Zhang, Y. Deng, X. Wang, and T.-S. Chua · 2025
Closest in time.
SCBench: A KV cache-centric analysis of long-context methods
Y. LI, H. Jiang, Q. Wu, X. Luo, S. Ahn, C. Zhang, A. H. Abdi, D. Li, J. Gao, Y. Yang, and L. Qiu · 2025
Closest in time.
Chunkkv: Semantic-preserving kv cache compression for efficient long-context llm inference
X. Liu, Z. Tang, P. Dong, Z. Li, B. Li, X. Hu, and X. Chu · 2025
Closest in time.
Towards lifelong dialogue agents via timeline-based memory management
K. T.-i. Ong, N. Kim, M. Gwak, H. Chae, T. Kwon, Y. Jo, S.-w. Hwang, D. Lee, and J. Yeo · 2025
Closest in time.
Secom: On memory construction and retrieval for personalized conversational agents
Z. Pan, Q. Wu, H. Jiang, X. Luo, H. Cheng, D. Li, Y. Yang, C.-Y. Lin, H. V. Zhao, L. Qiu, and J. Gao · 2025
Closest in time.
Zep: A temporal knowledge graph architecture for agent memory, 2025
P. Rasmussen, P. Paliychuk, T. Beauvais, J. Ryan, and D. Chalef · 2025
Closest in time.
Agent workflow memory
Z. Z. Wang, J. Mao, D. Fried, and G. Neubig · 2025
Closest in time.
Longmemeval: Benchmarking chat assistants on long-term interactive memory
D. Wu, H. Wang, W. Yu, Y. Zhang, K.-W. Chang, and D. Yu · 2025
Closest in time.
A-mem: Agentic memory for llm agents, 2025
W. Xu, K. Mei, H. Gao, J. Tan, Z. Liang, and Y. Zhang · 2025
Closest in time.