Fetching the paper…
Reading the bibliography…
Deep neural networks perform well on classification tasks where data streams are i.i.d.
Sparse distributed memory
Kanerva, P · 1988
Earlier work this paper cites.
Local learning algorithms
Bottou, L. and Vapnik, V · 1992
Earlier work this paper cites.
Sparse distributed memory and related models
Kanerva, P · 1992
Earlier work this paper cites.
Is learning the n-th thing any easier than learning the first?
Thrun, S · 1995
Earlier work this paper cites.
A model of inductive bias learning
Baxter, J · 2000
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A., Hinton, G., et al · 2009
Earlier work this paper cites.
An analysis of single-layer networks in unsupervised feature learning
Coates, A., Ng, A., and Lee, H · 2011
Earlier work this paper cites.
Learning to learn
Thrun, S. and Pratt, L · 2012
Earlier work this paper cites.
Error-driven incremental learning in deep convolutional neural network for large-scale image classification
Xiao, T., Zhang, J., Yang, K., Peng, Y., and Zhang, Z · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
Yosinski, J., Clune, J., Bengio, Y., and Lipson, H · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Hinton, G., Vinyals, O., Dean, J., et al · 2015
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., et al · 2015
Earlier work this paper cites.
End-to-end memory networks
Sukhbaatar, S., Weston, J., Fergus, R., et al · 2015
Earlier work this paper cites.
Chandar, S., Ahn, S., Larochelle, H., Vincent, P., Tesauro, G., and Bengio, Y · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A. A., Milan, K., Quan, J., Ramalho, T., Grabska-Barwinska, A., et al · 2017
Earlier work this paper cites.
Learning without forgetting
Li, Z. and Hoiem, D · 2017
Earlier work this paper cites.
icarl: Incremental classifier and representation learning
Rebuffi, S.-A., Kolesnikov, A., Sperl, G., and Lampert, C. H · 2017
Earlier work this paper cites.
Continual learning with deep generative replay
Shin, H., Lee, J. K., Kim, J., and Kim, J · 2017
Earlier work this paper cites.
Neural discrete representation learning
Van Den Oord, A., Vinyals, O., et al · 2017
Earlier work this paper cites.
Continual learning through synaptic intelligence
Zenke, F., Poole, B., and Ganguli, S · 2017
Earlier work this paper cites.
Lifelong machine learning
Chen, Z. and Liu, B · 2018
Earlier work this paper cites.
Overcoming catastrophic forgetting with hard attention to the task
Serra, J., Suris, D., Miron, M., and Karatzoglou, A · 2018
Earlier work this paper cites.
Online meta-learning
Finn, C., Rajeswaran, A., Kakade, S., and Levine, S · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for nlp
Houlsby, N., Giurgiu, A., Jastrzebski, S., Morrone, B., De Laroussilhe, Q., Gesmundo, A., Attariyan, M., and Gelly, S · 2019
Cited alongside, same era.
Unsupervised deep learning by neighbourhood discovery
Huang, J., Dong, Q., Gong, S., and Zhu, X · 2019
Cited alongside, same era.
Large memory layers with product keys
Lample, G., Sablayrolles, A., Ranzato, M., Denoyer, L., and Jégou, H · 2019
Cited alongside, same era.
Latent retrieval for weakly supervised open domain question answering
Lee, K., Chang, M.-W., and Toutanova, K · 2019
Cited alongside, same era.
Generating diverse high-fidelity images with vq-vae-2
Razavi, A., Van den Oord, A., and Vinyals, O · 2019
Cited alongside, same era.
Sketch based memory for neural networks
Panigrahy, R., Wang, X., and Zaheer, M · 2021
Later among the works it cites.
Combined scaling for open-vocabulary image classification
Pham, H., Dai, Z., Ghiasi, G., Kawaguchi, K., Liu, H., Yu, A. W., Yu, J., Chen, Y.-T., Luong, M.-T., Wu, Y., et al · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Later among the works it cites.
Gradient projection memory for continual learning
Saha, G., Garg, I., and Roy, K · 2021
Later among the works it cites.
Algorithmic insights on continual learning from fruit flies
Shen, Y., Dasgupta, S., and Navlakha, S · 2021
Later among the works it cites.
Translation-equivariant image quantizer for bi-directional image-text generation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Van de Ven, G. M. and Tolias, A. S · 2019
Cited alongside, same era.
Continual learning of context-dependent processing in neural networks
Zeng, G., Chen, Y., Cui, B., and Yu, S · 2019
Cited alongside, same era.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Cited alongside, same era.
Unsupervised learning of visual features by contrasting cluster assignments
Caron, M., Misra, I., Mairal, J., Goyal, P., Bojanowski, P., and Joulin, A · 2020
Cited alongside, same era.
Big self-supervised models are strong semi-supervised learners
Chen, T., Kornblith, S., Swersky, K., Norouzi, M., and Hinton, G. E · 2020
Cited alongside, same era.
Sequential mastery of multiple visual tasks: Networks naturally learn to learn and forget to forget
Davidson, G. and Mozer, M. C · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., et al · 2020
Cited alongside, same era.
Shin, W., Lee, G., Lee, J., Lee, J., and Choi, E · 2021
Later among the works it cites.
Ernie 3.0: Large-scale knowledge enhanced pre-training for language understanding and generation
Sun, Y., Wang, S., Feng, S., Ding, S., Pang, C., Shang, J., Liu, J., Chen, X., Zhao, Y., Lu, Y., et al · 2021
Later among the works it cites.
Efficient feature transformations for discriminative and generative continual learning
Verma, V. K., Liang, K. J., Mehta, N., Rai, P., and Carin, L · 2021
Later among the works it cites.
Emergent symbols through binding in external memory
Webb, T. W., Sinha, I., and Cohen, J. D · 2021
Later among the works it cites.
Soundstream: An end-to-end neural audio codec
Zeghidour, N., Luebs, A., Omran, A., Skoglund, J., and Tagliasacchi, M · 2021
Later among the works it cites.
Improving language models by retrieving from trillions of tokens
Borgeaud, S., Mensch, A., Hoffmann, J., Cai, T., Rutherford, E., Millican, K., Van Den Driessche, G. B., Lespiau, J.-B., Damoc, B., Clark, A., et al · 2022
Closest in time.
Efficient architecture search for continual learning
Gao, Q., Luo, Z., Klabjan, D., and Zhang, F · 2022
Closest in time.
Retrieval-augmented reinforcement learning
Goyal, A., Friesen, A., Banino, A., Weber, T., Ke, N. R., Badia, A. P., Guez, A., Mirza, M., Humphreys, P. C., Konyushova, K., et al · 2022
Closest in time.
Robustness implies generalization via data-dependent generalization bounds
Kawaguchi, K., Deng, Z., Luh, K., and Huang, J · 2022
Closest in time.
Fine-tuning can distort pretrained features and underperform out-of-distribution
Kumar, A., Raghunathan, A., Jones, R., Ma, T., and Liang, P · 2022
Closest in time.
Foundational models for continual learning: An empirical study of latent replay
Ostapenko, O., Lesort, T., Rodríguez, P., Arefin, M. R., Douillard, A., Rish, I., and Charlin, L · 2022
Closest in time.
Online task-free continual learning with dynamic sparse distributed memory
Pourcel, J., Vu, N.-S., and French, R. M · 2022
Closest in time.
Learning to imagine: Diversify memory for incremental learning using unlabeled data
Tang, Y.-M., Peng, Y.-X., and Zheng, W.-S · 2022
Closest in time.
Deeper insights into vits robustness towards common corruptions
Tian, R., Wu, Z., Dai, Q., Hu, H., and Jiang, Y · 2022
Closest in time.
Trockman, A. and Kolter, J. Z · 2022
Closest in time.
Learning to prompt for continual learning
Wang, Z., Zhang, Z., Lee, C.-Y., Zhang, H., Sun, R., Ren, X., Su, G., Perot, V., Dy, J., and Pfister, T · 2022
Closest in time.
Vector-quantized image modeling with improved vqgan
Yu, J., Li, X., Koh, J. Y., Zhang, H., Pang, R., Qin, J., Ku, A., Xu, Y., Baldridge, J., and Wu, Y · 2022
Closest in time.
Sparse distributed memory is a continual learner
Bricken, T., Davies, X., Singh, D., Krotov, D., and Kreiman, G · 2023
Closest in time.