Fetching the paper…
Reading the bibliography…
In this work, we study the problem of continual learning (CL) where the goal is to learn a model on a sequence of tasks, such that the data from the previous tasks becomes unavailable while learning on the current task data.
Amari, S.i.: Neural learning in structured parameter spaces - natural riemannian gradient. In: Mozer, M., Jordan, M., Petsche, T. (eds.) Advances in Neural Information Processing Systems. vol. 9. MIT Press (1996)
1996
Earlier work this paper cites.
French, R.M.: Catastrophic forgetting in connectionist networks. Trends in cognitive sciences 3
1999
Earlier work this paper cites.
Friedman, J., Hastie, T., Tibshirani, R., et al.: The elements of statistical learning. Springer series in statistics New York (2001)
2001
Earlier work this paper cites.
Abraham, W.C., Robins, A.: Memory retention–the synaptic stability versus plasticity dilemma. Trends in neurosciences 28
2005
Earlier work this paper cites.
Spall, J.C.: Monte carlo computation of the fisher information matrix in nonstandard settings. Journal of Computational and Graphical Statistics 14
2005
Earlier work this paper cites.
Spall, J.C.: Improved methods for monte carlo estimation of the fisher information matrix. In: 2008 American Control Conference (2008)
2008
Earlier work this paper cites.
Krizhevsky, A., Hinton, G., et al.: Learning multiple layers of features from tiny images (2009)
2009
Earlier work this paper cites.
Wah, C., Branson, S., Welinder, P., Perona, P., Belongie, S.: The Caltech-UCSD Birds-200-2011 Dataset. Tech. Rep. CNS-TR-2011-001, California Institute of Technology (2011)
2011
Earlier work this paper cites.
Wah, C., Branson, S., Welinder, P., et al.: The caltech-ucsd birds-200-2011 dataset (2011)
2011
Earlier work this paper cites.
Krause, J., Stark, M., Deng, J., Fei-Fei, L.: 3d object representations for fine-grained categorization. In: Proceedings of the IEEE International Conference on Computer Vision Workshops (2013)
2013
Earlier work this paper cites.
Krause, J., Stark, M., Deng, J., Fei-Fei, L.: 3d object representations for fine-grained categorization. In: 2013 IEEE International Conference on Computer Vision Workshops (2013)
2013
Earlier work this paper cites.
Pascanu, R., Bengio, Y.: Revisiting natural gradient for deep networks (2014)
2014
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: IEEE Conf. Comput. Vis. Pattern Recog. (2015)
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., Wojna, Z.: Rethinking the inception architecture for computer vision. In: IEEE Conf. Comput. Vis. Pattern Recog. (2016)
2016
Earlier work this paper cites.
Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., Wojna, Z.: Rethinking the inception architecture for computer vision. In: IEEE Conf. Comput. Vis. Pattern Recog. (2016)
2016
Earlier work this paper cites.
Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A.A., Milan, K., Quan, J., Ramalho, T., Grabska-Barwinska, A., et al.: Overcoming catastrophic forgetting in neural networks. PNAS 114
2017
Earlier work this paper cites.
Lakshminarayanan, B., Pritzel, A., Blundell, C.: Simple and scalable predictive uncertainty estimation using deep ensembles. In: Adv. Neural Inform. Process. Syst. (2017)
2017
Earlier work this paper cites.
Li, Z., Hoiem, D.: Learning without forgetting. TPAMI 40
2017
Earlier work this paper cites.
Zenke, F., Poole, B., Ganguli, S.: Continual learning through synaptic intelligence. In: Int. Conf. Machine Learn. (2017)
2017
Earlier work this paper cites.
Aljundi, R., Babiloni, F., Elhoseiny, M., Rohrbach, M., Tuytelaars, T.: Memory aware synapses: Learning what (not) to forget. In: Eur. Conf. Comput. Vis. (2018)
2018
Earlier work this paper cites.
Izmailov, P., Podoprikhin, D., Garipov, T., Vetrov, D., Wilson, A.G.: Averaging weights leads to wider optima and better generalization. In: Conference on Uncertainty in Artificial Intelligence (UAI) (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Serra, J., Suris, D., Miron, M., Karatzoglou, A.: Overcoming catastrophic forgetting with hard attention to the task. In: Int. Conf. Machine Learn. (2018)
2018
Earlier work this paper cites.
Dhar, P., Singh, R.V., Peng, K.C., Wu, Z., Chellappa, R.: Learning without memorizing. In: IEEE Conf. Comput. Vis. Pattern Recog. (2019)
2019
Earlier work this paper cites.
Ovadia, Y., Fertig, E., Ren, J., Nado, Z., Sculley, D., Nowozin, S., Dillon, J.V., Lakshminarayanan, B., Snoek, J.: Can you trust your model’s uncertainty? evaluating predictive uncertainty under dataset shift. In: Adv. Neural Inform. Process. Syst. (2019)
2019
Earlier work this paper cites.
van de Ven, G.M., Tolias, A.S.: Three scenarios for continual learning (2019)
2019
Earlier work this paper cites.
Wu, Y., Chen, Y., Wang, L., Ye, Y., Liu, Z., Guo, Y., Fu, Y.: Large scale incremental learning. In: IEEE Conf. Comput. Vis. Pattern Recog. (2019)
2019
Earlier work this paper cites.
Buzzega, P., Boschini, M., Porrello, A., Abati, D., Calderara, S.: Dark experience for general continual learning: a strong, simple baseline. Adv. Neural Inform. Process. Syst. 33
2020
Earlier work this paper cites.
Chizat, L., Oyallon, E., Bach, F.: On lazy training in differentiable programming (2020)
2020
Cited alongside, same era.
Frankle, J., Dziugaite, G.K., Roy, D., Carbin, M.: Linear mode connectivity and the lottery ticket hypothesis. In: Int. Conf. Machine Learn. (2020)
2020
Cited alongside, same era.
Li, T., Sahu, A.K., Talwalkar, A., Smith, V.: Federated learning: Challenges, methods, and future directions. IEEE Signal Processing Magazine 37
2020
Cited alongside, same era.
Liao, Z., Drummond, T., Reid, I., Carneiro, G.: Approximate fisher information matrix to characterize the training of deep neural networks. IEEE Transactions on Pattern Analysis and Machine Intelligence 42
2020
Cited alongside, same era.
Mirzadeh, S.I., Farajtabar, M., Gorur, D., Pascanu, R., Ghasemzadeh, H.: Linear mode connectivity in multitask and continual learning (2020)
2020
Masana, M., Liu, X., Twardowski, B., Menta, M., Bagdanov, A.D., van de Weijer, J.: Class-incremental learning: survey and performance evaluation on image classification. IEEE Trans. Pattern Anal. Mach. Intell. (2022)
2022
Later among the works it cites.
Matena, M.S., Raffel, C.A.: Merging models with fisher-weighted averaging. Adv. Neural Inform. Process. Syst. 35
2022
Later among the works it cites.
Murata, K., Ito, S., Ohara, K.: Learning and transforming general representations to break down stability-plasticity dilemma. In: Proceedings of the Asian Conference on Computer Vision (2022)
2022
Later among the works it cites.
Pham, Q., Liu, C., Steven, H.: Continual normalization: Rethinking batch normalization for online continual learning. In: Int. Conf. Learn. Represent. (2022)
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Neyshabur, B., Sedghi, H., Zhang, C.: What is being transferred in transfer learning? In: Adv. Neural Inform. Process. Syst. (2020)
2020
Cited alongside, same era.
Prabhu, A., Torr, P.H., Dokania, P.K.: Gdumb: A simple approach that questions our progress in continual learning. In: Eur. Conf. Comput. Vis. (2020)
2020
Cited alongside, same era.
Stickland, A.C., Murray, I.: Diverse ensembles improve calibration. In: International Conference on Machine Learning (ICML) Workshop on Uncertainty and Robustness in Deep Learning (2020)
2020
Cited alongside, same era.
Titsias, M.K., Schwarz, J., de G. Matthews, A.G., Pascanu, R., Teh, Y.W.: Functional regularisation for continual learning with gaussian processes (2020)
2020
Cited alongside, same era.
Zhang, J., Zhang, J., Ghosh, S., Li, D., Zhu, J., Zhang, H., Wang, Y.: Regularize, expand and compress: Nonexpansive continual learning. In: WACV (2020)
2020
Cited alongside, same era.
2021
Cited alongside, same era.
Dogucu, M., Johnson, A., Ott, M.: bayesrules: Datasets and Supplemental Functions from Bayes Rules! Book (2021), r package version 0.0.2.9000
2021
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
Wang, L., Zhang, X., Li, Q., Zhu, J., Zhong, Y.: Coscl: Cooperation of small continual learners is stronger than a big one (2022)
2022
Later among the works it cites.
2022
Later among the works it cites.
Wang, Z., Zhang, Z., Ebrahimi, S., Sun, R., Zhang, H., Lee, C.Y., Ren, X., Su, G., Perot, V., Dy, J., et al.: Dualprompt: Complementary prompting for rehearsal-free continual learning. In: Eur. Conf. Comput. Vis. Springer (2022)
2022
Later among the works it cites.
Wang, Z., Zhang, Z., Lee, C.Y., Zhang, H., Sun, R., Ren, X., Su, G., Perot, V., Dy, J., Pfister, T.: Learning to prompt for continual learning. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Wortsman, M., Ilharco, G., Gadre, S.Y., Roelofs, R., Gontijo-Lopes, R., Morcos, A.S., Namkoong, H., Farhadi, A., Carmon, Y., Kornblith, S., et al.: Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time. In: Int. Conf. Machine Learn. PMLR (2022)
2022
Later among the works it cites.
Wortsman, M., Ilharco, G., Li, M., Kim, J.W., Hajishirzi, H., Farhadi, A., Namkoong, H., Schmidt, L.: Robust fine-tuning of zero-shot models. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Wu, T.Y., Swaminathan, G., Li, Z., Ravichandran, A., Vasconcelos, N., Bhotika, R., Soatto, S.: Class-incremental learning with strong pre-trained models. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Yang, B., Deng, X., Shi, H., Li, C., Zhang, G., Xu, H., Zhao, S., Lin, L., Liang, X.: Continual object detection via prototypical task correlation guided gating mechanism. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Later among the works it cites.
Janson, P., Zhang, W., Aljundi, R., Elhoseiny, M.: A simple baseline that questions the use of pretrained-models in continual learning (2023)
2023
Closest in time.
Kim, D., Han, B.: On the stability-plasticity dilemma of class-incremental learning. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Closest in time.
McMahan, H.B., Moore, E., Ramage, D., Hampson, S., y Arcas, B.A.: Communication-efficient learning of deep networks from decentralized data (2023)
2023
Closest in time.
Mehta, S.V., Patil, D., Chandar, S., Strubell, E.: An empirical investigation of the role of pre-training in lifelong learning (2023)
2023
Closest in time.
Oquab, M., Darcet, T., Moutakanni, T., Vo, H.V., Szafraniec, M., Khalidov, V., Fernandez, P., Haziza, D., Massa, F., El-Nouby, A., Howes, R., Huang, P.Y., Xu, H., Sharma, V., Li, S.W., Galuba, W., Rabbat, M., Assran, M., Ballas, N., Synnaeve, G., Misra, I., Jegou, H., Mairal, J., Labatut, P., Joulin, A., Bojanowski, P.: Dinov2: Learning robust visual features without supervision (2023)
2023
Closest in time.
Panos, A., Kobe, Y., Reino, D.O., Aljundi, R., Turner, R.E.: First session adaptation: A strong replay-free baseline for class-incremental learning. In: Int. Conf. Comput. Vis. (2023)
2023
Closest in time.
Ramé, A., Kirchmeyer, M., Rahier, T., Rakotomamonjy, A., Gallinari, P., Cord, M.: Diverse weight averaging for out-of-distribution generalization (2023)
2023
Closest in time.
Rypeść, G., Cygert, S., Khan, V., Trzcinski, T., Zieliński, B.M., Twardowski, B.: Divide and not forget: Ensemble of selectively trained experts in continual learning. In: Int. Conf. Learn. Represent. (2023)
2023
Closest in time.
2023
Closest in time.
Wang, L., Zhang, X., Li, Q., Zhang, M., Su, H., Zhu, J., Zhong, Y.: Incorporating neuro-inspired adaptability for continual learning in artificial intelligence. Nature Machine Intelligence 5
2023
Closest in time.
Wang, L., Zhang, X., Su, H., Zhu, J.: A comprehensive survey of continual learning: Theory, method and application (2023)
2023
Closest in time.
Zhang, G., Wang, L., Kang, G., Chen, L., Wei, Y.: Slca: Slow learner with classifier alignment for continual learning on a pre-trained model. In: Int. Conf. Comput. Vis. (2023)
2023
Closest in time.
Zhou, D.W., Ye, H.J., Zhan, D.C., Liu, Z.: Revisiting class-incremental learning with pre-trained models: Generalizability and adaptivity are all you need (2023)
2023
Closest in time.
McDonnell, M.D., Gong, D., Parvaneh, A., Abbasnejad, E., van den Hengel, A.: Ranpac: Random projections and pre-trained models for continual learning. Adv. Neural Inform. Process. Syst. 36
2024
Closest in time.
Ren, W., Li, X., Wang, L., Zhao, T., Qin, W.: Analyzing and reducing catastrophic forgetting in parameter efficient tuning (2024)
2024
Closest in time.