Fetching the paper…
Reading the bibliography…
Episodic memory -- the ability to recall specific events grounded in time and space -- is a cornerstone of human cognition, enabling not only coherent storytelling, but also planning and decision-making.
Generating long sequences with sparse transformers
Rewon Child, Scott Gray, Alec Radford, and Ilya Sutskever · 1904
Earlier work this paper cites.
Generalization through memorization: Nearest neighbor language models
Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis · 1911
Earlier work this paper cites.
A theory of cognitive dissonance , volume 2
Leon Festinger · 1962
Earlier work this paper cites.
The hippocampus as a spatial map: preliminary evidence from unit activity in the freely-moving rat
John O’Keefe and Jonathan Dostrovsky · 1971
Earlier work this paper cites.
Episodic and semantic memory
Endel Tulving et al · 1972
Earlier work this paper cites.
Encoding specificity and retrieval processes in episodic memory
Endel Tulving and Donald M Thomson · 1973
Earlier work this paper cites.
Frequency and repetition effects in lexical memory
Don L Scarborough, Charles Cortese, and Hollis S Scarborough · 1977
Earlier work this paper cites.
The hippocampus as a cognitive map
John O’Keefe and Lynn Nadel · 1978
Earlier work this paper cites.
Recognizing: The judgment of previous occurrence
George Mandler · 1980
Earlier work this paper cites.
Reality monitoring
Marcia K Johnson and Carol L Raye · 1981
Earlier work this paper cites.
The rivermead behavioural memory test (rbmt)
BA Wilson, Janet Cockburn, and Alan Baddeley · 1985
Earlier work this paper cites.
The hippocampal memory indexing theory
Timothy J Teyler and Pascal DiScenna · 1986
Earlier work this paper cites.
Source monitoring
Marcia K Johnson, Shahin Hashtroudi, and D Stephen Lindsay · 1993
Earlier work this paper cites.
The autobiographical memory interview (ami) in organic and psychogenic amnesia
Michael D Kopelman · 1994
Earlier work this paper cites.
The hippocampus and memory for orderly stimulus relations
Jeffery A Dusek and Howard Eichenbaum · 1997
Earlier work this paper cites.
California verbal learning test–
Dean C Delis, Joel H Kramer, Edith Kaplan, and Beth A Ober · 2000
Earlier work this paper cites.
Episodic future thinking
Cristina M Atance and Daniela K O’Neill · 2001
Earlier work this paper cites.
The human hippocampus and spatial and episodic memory
Neil Burgess, Eleanor A Maguire, and John O’Keefe · 2002
Earlier work this paper cites.
Aging and autobiographical memory: dissociating episodic from semantic retrieval
Brian Levine, Eva Svoboda, Janine F Hay, Gordon Winocur, and Morris Moscovitch · 2002
Earlier work this paper cites.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan · 2004
Earlier work this paper cites.
Microstructure of a spatial map in the entorhinal cortex
Torkel Hafting, Marianne Fyhn, Sturla Molden, May-Britt Moser, and Edvard I Moser · 2005
Earlier work this paper cites.
Statistical comparisons of classifiers over multiple data sets
Janez Demšar · 2006
Earlier work this paper cites.
Gmat: Global memory augmentation for transformers
Ankit Gupta and Jonathan Berant · 2006
Earlier work this paper cites.
Linformer: Self-attention with linear complexity
Sinong Wang, Belinda Z Li, Madian Khabsa, Han Fang, and Hao Ma · 2006
Earlier work this paper cites.
Leveraging passage retrieval with generative models for open domain question answering
Gautier Izacard and Edouard Grave · 2007
Earlier work this paper cites.
The hippocampal indexing theory and episodic memory: updating the index
Timothy J Teyler and Jerry W Rudy · 2007
Earlier work this paper cites.
Freebase: a collaboratively created graph database for structuring human knowledge
Kurt Bollacker, Colin Evans, Praveen Paritosh, Tim Sturge, and Jamie Taylor · 2008
Cited alongside, same era.
Detecting navigational deficits in cognitive aging and alzheimer disease using virtual reality
Laura A Cushman, Karen Stein, and Charles J Duffy · 2008
Cited alongside, same era.
Rethinking attention with performers
Krzysztof Choromanski, Valerii Likhosherstov, David Dohan, Xingyou Song, Andreea Gane, Tamas Sarlos, Peter Hawkins, Jared Davis, Afroz Mohiuddin, Lukasz Kaiser, et al · 2009
Cited alongside, same era.
Kilt: a benchmark for knowledge intensive language tasks
Fabio Petroni, Aleksandra Piktus, Angela Fan, Patrick Lewis, Majid Yazdani, Nicola De Cao, James Thorne, Yacine Jernite, Vladimir Karpukhin, Jean Maillard, et al · 2009
Cited alongside, same era.
Hippocampal “time cells” bridge the gap in memory for discontiguous events
Quantifying memorization across neural language models
Nicholas Carlini, Daphne Ippolito, Matthew Jagielski, Katherine Lee, Florian Tramer, and Chiyuan Zhang · 2022
Later among the works it cites.
Rethinking the distinction between episodic and semantic memory: Insights from the past, present, and future
Felipe De Brigard, Sharda Umanath, and Muireann Irish · 2022
Later among the works it cites.
Time-aware language models as temporal knowledge bases
Bhuwan Dhingra, Jeremy R Cole, Julian Martin Eisenschlos, Daniel Gillick, Jacob Eisenstein, and William W Cohen · 2022
Later among the works it cites.
Large language models struggle to learn long-tail knowledge
Nikhil Kandpal, Haikang Deng, Adam Roberts, Eric Wallace, and Colin Raffel · 2022
Later among the works it cites.
An efficient memory-augmented transformer for knowledge-intensive nlp tasks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Christopher J MacDonald, Kyle Q Lepage, Uri T Eden, and Howard Eichenbaum · 2011
Cited alongside, same era.
Extracting training data from large language models
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, et al · 2012
Cited alongside, same era.
Learning dense representations of phrases at scale
Jinhyuk Lee, Mujeen Sung, Jaewoo Kang, and Danqi Chen · 2012
Cited alongside, same era.
Modifying memories in transformer models
Chen Zhu, Ankit Singh Rawat, Manzil Zaheer, Srinadh Bhojanapalli, Daliang Li, Felix Yu, and Sanjiv Kumar · 2012
Cited alongside, same era.
Online evaluation of novel choices by simultaneous representation of multiple memories
Helen C Barron, Raymond J Dolan, and Timothy EJ Behrens · 2013
Cited alongside, same era.
Hippocampal place-cell sequences depict future paths to remembered goals
Brad E Pfeiffer and David J Foster · 2013
Cited alongside, same era.
Cortical representations are reinstated by the hippocampus during memory retrieval
Kazumasa Z Tanaka, Aleksandr Pevzner, Anahita B Hamidi, Yuki Nakazawa, Jalina Graham, and Brian J Wiltgen · 2014
Cited alongside, same era.
Navigating through digital folders uses the same brain structures as real world navigation
Yael Benn, Ofer Bergman, Liv Glazer, Paris Arent, Iain D Wilkinson, Rosemary Varley, and Steve Whittaker · 2015
Cited alongside, same era.
Yuxiang Wu, Yu Zhao, Baotian Hu, Pasquale Minervini, Pontus Stenetorp, and Sebastian Riedel · 2022
Later among the works it cites.
Why does claude sometimes follow up his response with "human:" and then some response i presume he’s thinking i will respond with?
adbertram · 2023
Later among the works it cites.
Emergent and predictable memorization in large language models
Stella Biderman, USVSN Sai Prashanth, Lintang Sutawika, Hailey Schoelkopf, Quentin Anthony, Shivanshu Purohit, and Edward Raf · 2023
Later among the works it cites.
Longnet: Scaling transformers to 1,000,000,000 tokens
Jiayu Ding, Shuming Ma, Li Dong, Xingxing Zhang, Shaohan Huang, Wenhui Wang, and Furu Wei · 2023
Later among the works it cites.
Sources of hallucination by large language models on inference tasks
Nick McKenna, Tianyi Li, Liang Cheng, Mohammad Javad Hosseini, Mark Johnson, and Mark Steedman · 2023
Later among the works it cites.
Towards benchmarking and improving the temporal reasoning capability of large language models
Qingyu Tan, Hwee Tou Ng, and Lidong Bing · 2023
Later among the works it cites.
Augmenting language models with long-term memory
Weizhi Wang, Li Dong, Hao Cheng, Xiaodong Liu, Xifeng Yan, Jianfeng Gao, and Furu Wei · 2023
Later among the works it cites.
Efficient streaming language models with attention sinks
Guangxuan Xiao, Yuandong Tian, Beidi Chen, Song Han, and Mike Lewis · 2023
Later among the works it cites.
Rishabh Agarwal, Avi Singh, Lei M Zhang, Bernd Bohnet, Stephanie Chan, Ankesh Anand, Zaheer Abbas, Azade Nova, John D Co-Reyes, Eric Chu, et al · 2024
Later among the works it cites.
Bernd Bohnet, Kevin Swersky, Rosanne Liu, Pranjal Awasthi, Azade Nova, Javier Snaider, Hanie Sedghi, Aaron T Parisi, Michael Collins, Angeliki Lazaridou, et al · 2024
Later among the works it cites.
Larimar: Large language models with episodic memory control
Payel Das, Subhajit Chaudhury, Elliot Nelson, Igor Melnyk, Sarath Swaminathan, Sihui Dai, Aurélie Lozano, Georgios Kollias, Vijil Chenthamarakshan, Soham Dan, et al · 2024
Later among the works it cites.
Human-like episodic memory for infinite context llms
Zafeirios Fountas, Martin A Benfeghoul, Adnan Oomerjee, Fenia Christopoulou, Gerasimos Lampouras, Haitham Bou-Ammar, and Jun Wang · 2024
Later among the works it cites.
Ruler: What’s the real context size of your long-context language models?
Cheng-Ping Hsieh, Simeng Sun, Samuel Kriman, Shantanu Acharya, Dima Rekesh, Fei Jia, and Boris Ginsburg · 2024
Later among the works it cites.
Needle In A Haystack - Pressure Testing LLMs
Gregory Kamradt · 2024
Later among the works it cites.
Realtime qa: what’s the answer right now?
Jungo Kasai, Keisuke Sakaguchi, Ronan Le Bras, Akari Asai, Xinyan Yu, Dragomir Radev, Noah A Smith, Yejin Choi, Kentaro Inui, et al · 2024
Later among the works it cites.
Babilong: Testing the limits of llms with long context reasoning-in-a-haystack
Yuri Kuratov, Aydar Bulatov, Petr Anokhin, Ivan Rodkin, Dmitry Sorokin, Artyom Sorokin, and Mikhail Burtsev · 2024
Later among the works it cites.
Needlebench: Can llms do retrieval and reasoning in 1 million context window?
Mo Li, Songyang Zhang, Yunxin Liu, and Kai Chen · 2024
Later among the works it cites.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Machel Reid, Nikolay Savinov, Denis Teplyashin, Dmitry Lepikhin, Timothy Lillicrap, Jean-baptiste Alayrac, Radu Soricut, Angeliki Lazaridou, Orhan Firat, Julian Schrittwieser, et al · 2024
Later among the works it cites.
Michelangelo: Long context evaluations beyond haystacks via latent structure queries
Kiran Vodrahalli, Santiago Ontanon, Nilesh Tripuraneni, Kelvin Xu, Sanil Jain, Rakesh Shivanna, Jeffrey Hui, Nishanth Dikkala, Mehran Kazemi, Bahare Fatemi, et al · 2024
Later among the works it cites.
Infinite bench: Extending long context evaluation beyond 100k tokens
Xinrong Zhang, Yingfa Chen, Shengding Hu, Zihang Xu, Junhao Chen, Moo Hao, Xu Han, Zhen Thai, Shuo Wang, Zhiyuan Liu, et al · 2024
Later among the works it cites.
Code and Data for Episodic Memories Generation and Evaluation Benchmark for Large Language Models
Alexis Huet, Zied Ben Houidi, and Dario Rossi · 2025
Closest in time.