Fetching the paper…
Reading the bibliography…
Memory is one of the most essential cognitive functions serving as a repository of world knowledge and episodes of activities.
A comprehensive, application-oriented study of catastrophic forgetting in dnns
Benedikt Pfülb and Alexander Gepperth. 2019 · 1905
Earlier work this paper cites.
Do Neural Language Representations Learn Physical Commonsense?
Maxwell Forbes, Ari Holtzman, and Yejin Choi. 2019 · 1908
Earlier work this paper cites.
The magical number seven, plus or minus two: Some limits on our capacity for processing information
George A Miller. 1956 · 1956
Earlier work this paper cites.
The present status of interference theory
Leo Postman. 1961 · 1959
Earlier work this paper cites.
The interference theory of forgetting
John Ceraso. 1967 · 1967
Earlier work this paper cites.
The control of short-term memory
Richard C Atkinson and Richard M Shiffrin. 1971 · 1971
Earlier work this paper cites.
The long and the short of long–term memory—a molecular framework
Philip Goelet, Vincent F Castellucci, Samuel Schacher, and Eric R Kandel. 1986 · 1986
Earlier work this paper cites.
Mechanisms of memory
Larry R Squire. 1986 · 1986
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
Michael McCloskey and Neal J Cohen. 1989 · 1989
Earlier work this paper cites.
Using semi-distributed representations to overcome catastrophic forgetting in connectionist networks
Robert M French. 1991 · 1991
Earlier work this paper cites.
On the form of forgetting
John T Wixted and Ebbe B Ebbesen. 1991 · 1991
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
James L McClelland, Bruce L McNaughton, and Randall C O’Reilly. 1995 · 1995
Earlier work this paper cites.
Catastrophic forgetting, rehearsal and pseudorehearsal
Anthony Robins. 1995 · 1995
Earlier work this paper cites.
The development of memory
Susan E Gathercole. 1998 · 1998
Earlier work this paper cites.
Episodic memory, semantic memory, and amnesia
Larry R Squire and Stuart M Zola. 1998 · 1998
Earlier work this paper cites.
Episodic and declarative memory: role of the hippocampus
Endel Tulving and Hans J Markowitsch. 1998 · 1998
Earlier work this paper cites.
Catastrophic forgetting in connectionist networks
Robert M French. 1999 · 1999
Earlier work this paper cites.
Cognitive architectures: Where do we go from here?
Wlodzislaw Duch, Richard Jayadi Oentaryo, and Michel Pasquier. 2008 · 2008
Earlier work this paper cites.
Critical periods and catastrophic interference effects in the development of self-organizing feature maps
Fiona M Richardson and Michael SC Thomas. 2008 · 2008
Earlier work this paper cites.
Long-term retention of basic science knowledge: a review study
Eugène JFM Custers. 2010 · 2010
Earlier work this paper cites.
Attention and arousal: Cognition and performance
Michael Eysenck. 2012 · 2012
Cited alongside, same era.
Memory: A contribution to experimental psychology
Hermann Ebbinghaus. 2013 · 2013
Cited alongside, same era.
An empirical investigation of catastrophic forgetting in gradient-based neural networks
Ian J Goodfellow, Mehdi Mirza, Da Xiao, Aaron Courville, and Yoshua Bengio. 2013 · 2013
Cited alongside, same era.
Wikidata: a free collaborative knowledgebase
Denny Vrandečić and Markus Krötzsch. 2014 · 2014
Cited alongside, same era.
Replication and analysis of ebbinghaus’ forgetting curve
Jaap MJ Murre and Joeri Dros. 2015 · 2015
Cited alongside, same era.
A bio-inspired incremental learning architecture for applied perceptual problems
Alexander Gepperth and Cem Karaoguz. 2016 · 2016
X-FACTR: Multilingual factual knowledge retrieval from pretrained language models
Zhengbao Jiang, Antonios Anastasopoulos, Jun Araki, Haibo Ding, and Graham Neubig. 2020a · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, Peter J Liu, et al. 2020 · 2020
Later among the works it cites.
Evaluating commonsense in pre-trained language models
Xuhui Zhou, Yue Zhang, Leyang Cui, and Dandan Huang. 2020 · 2020
Later among the works it cites.
Knowledgeable or educated guess? revisiting language models as knowledge bases
Boxi Cao, Hongyu Lin, Xianpei Han, Le Sun, Lingyong Yan, Meng Liao, Tong Xue, and Jin Xu. 2021 · 2021
Later among the works it cites.
Extracting training data from large language models
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, et al. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning without forgetting
Zhizhong Li and Derek Hoiem. 2017 · 2017
Cited alongside, same era.
An overview of multi-task learning in deep neural networks
Sebastian Ruder. 2017 · 2017
Cited alongside, same era.
Continual learning through synaptic intelligence
Friedemann Zenke, Ben Poole, and Surya Ganguli. 2017 · 2017
Cited alongside, same era.
Lifelong machine learning
Zhiyuan Chen and Bing Liu. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Open sesame: Getting inside BERT’s linguistic knowledge
Yongjie Lin, Yi Chern Tan, and Robert Frank. 2019 · 2019
Cited alongside, same era.
A continual learning survey: Defying forgetting in classification tasks
Matthias De Lange, Rahaf Aljundi, Marc Masana, Sarah Parisot, Xu Jia, Aleš Leonardis, Gregory Slabaugh, and Tinne Tuytelaars. 2021 · 2021
Later among the works it cites.
Measuring and improving consistency in pretrained language models
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, Eduard Hovy, Hinrich Schütze, and Yoav Goldberg. 2021 · 2021
Later among the works it cites.
Multilingual LAMA: Investigating knowledge in multilingual pretrained language models
Nora Kassner, Philipp Dufter, and Hinrich Schütze. 2021 · 2021
Later among the works it cites.
R Thomas McCoy, Paul Smolensky, Tal Linzen, Jianfeng Gao, and Asli Celikyilmaz. 2021 · 2021
Later among the works it cites.
An empirical investigation of the role of pre-training in lifelong learning
Sanket Vaibhav Mehta, Darshan Patil, Sarath Chandar, and Emma Strubell. 2021 · 2021
Later among the works it cites.
Effect of scale on catastrophic forgetting in neural networks
Vinay Venkatesh Ramasesh, Aitor Lewkowycz, and Ethan Dyer. 2021 · 2021
Later among the works it cites.
A survey on multi-task learning
Yu Zhang and Qiang Yang. 2021 · 2021
Later among the works it cites.
Can prompt probe pretrained language models? understanding the invisible risks from a causal view
Boxi Cao, Hongyu Lin, Xianpei Han, Fangchao Liu, and Le Sun. 2022 · 2022
Later among the works it cites.
Quantifying memorization across neural language models
Nicholas Carlini, Daphne Ippolito, Matthew Jagielski, Katherine Lee, Florian Tramer, and Chiyuan Zhang. 2022 · 2022
Later among the works it cites.
Towards continual knowledge learning of language models
Joel Jang, Seonghyeon Ye, Sohee Yang, Joongbo Shin, Janghoon Han, Gyeonghun Kim, Stanley Jungkyu Choi, and Minjoon Seo. 2022 · 2022
Later among the works it cites.
Memorisation versus generalisation in pre-trained language models
Michael Tänzer, Sebastian Ruder, and Marek Rei. 2022 · 2022
Later among the works it cites.
Memorization without overfitting: Analyzing the training dynamics of large language models
Kushal Tirumala, Aram H Markosyan, Luke Zettlemoyer, and Armen Aghajanyan. 2022 · 2022
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Closest in time.
The life cycle of knowledge in big language models: A survey
Boxi Cao, Hongyu Lin, Xianpei Han, and Le Sun. 2024 · 2024
Closest in time.