Fetching the paper…
Reading the bibliography…
Class-Incremental Learning (CIL) is a practical and challenging problem for achieving general artificial intelligence.
Y. LeCun, “The mnist database of handwritten digits,” http://yann. lecun. com/exdb/mnist/ , 1998
1998
Earlier work this paper cites.
N. Japkowicz and S. Stephen, “The class imbalance problem: A systematic study,” Intelligent data analysis , vol. 6, no. 5, pp. 429–449, 2002
2002
Earlier work this paper cites.
E. T. K. Sang and F. De Meulder, “Introduction to the conll-2003 shared task: Language-independent named entity recognition,” in Proceedings of the Seventh Conference on Natural Language Learning at HLT-NAACL 2003 , 2003, pp. 142–147
2003
Earlier work this paper cites.
E. Hovy, M. Marcus, M. Palmer, L. Ramshaw, and R. Weischedel, “Ontonotes: the 90% solution,” in Proceedings of the human language technology conference of the NAACL, Companion Volume: Short Papers , 2006, pp. 57–60
2006
Earlier work this paper cites.
A. Krizhevsky et al. , “Learning multiple layers of features from tiny images,” 2009
2009
Earlier work this paper cites.
H. He and E. A. Garcia, “Learning from imbalanced data,” IEEE Transactions on knowledge and data engineering , vol. 21, no. 9, pp. 1263–1284, 2009
2009
Earlier work this paper cites.
J. Pearl, Causality . Cambridge university press, 2009
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in 2009 IEEE conference on computer vision and pattern recognition . Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert, “icarl: Incremental classifier and representation learning,” in Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , 2017, pp. 2001–2010
2010
Earlier work this paper cites.
S. N. Murphy, G. Weber, M. Mendis, V. Gainer, H. C. Chueh, S. Churchill, and I. Kohane, “Serving the enterprise and beyond with informatics for integrating biology and the bedside (i2b2),” Journal of the American Medical Informatics Association , vol. 17, no. 2, pp. 124–130, 2010
2010
Earlier work this paper cites.
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng, “Reading digits in natural images with unsupervised feature learning,” 2011
2011
Earlier work this paper cites.
Y. Bulatov, “Notmnist dataset,” Google (Books/OCR), Tech. Rep.[Online]. Available: http://yaroslavvb. blogspot. it/2011/09/notmnist-dataset. html , vol. 2, 2011
2011
Earlier work this paper cites.
J. Liu, P. Pasupat, S. Cyphers, and J. R. Glass, “Asgard: A portable architecture for multilingual dialogue systems,” in IEEE International Conference on Acoustics, Speech and Signal Processing , 2013, pp. 8386–8390
2013
Earlier work this paper cites.
J. Liu, P. Pasupat, Y. Wang, S. Cyphers, and J. R. Glass, “Query understanding enhanced by hierarchical parsing structures,” in IEEE Workshop on Automatic Speech Recognition and Understanding , 2013, pp. 72–77
2013
Earlier work this paper cites.
2015
Earlier work this paper cites.
X. Zhang, J. Zhao, and Y. LeCun, “Character-level convolutional networks for text classification,” Advances in neural information processing systems , vol. 28, 2015
2015
Earlier work this paper cites.
M. Glymour, J. Pearl, and N. P. Jewell, Causal inference in statistics: A primer . John Wiley & Sons, 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska et al. , “Overcoming catastrophic forgetting in neural networks,” Proceedings of the national academy of sciences , vol. 114, no. 13, pp. 3521–3526, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
Z. Li and D. Hoiem, “Learning without forgetting,” IEEE transactions on pattern analysis and machine intelligence , vol. 40, no. 12, pp. 2935–2947, 2017
2017
Earlier work this paper cites.
F. Zenke, B. Poole, and S. Ganguli, “Continual learning through synaptic intelligence,” in International conference on machine learning . PMLR, 2017, pp. 3987–3995
2017
Earlier work this paper cites.
D. Lopez-Paz and M. Ranzato, “Gradient episodic memory for continual learning,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Serra, D. Suris, M. Miron, and A. Karatzoglou, “Overcoming catastrophic forgetting with hard attention to the task,” in International Conference on Machine Learning . PMLR, 2018, pp. 4548–4557
2018
Earlier work this paper cites.
A. Chaudhry, P. K. Dokania, T. Ajanthan, and P. H. Torr, “Riemannian walk for incremental learning: Understanding forgetting and intransigence,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 532–547
2018
Earlier work this paper cites.
P. Sprechmann, S. M. Jayakumar, J. W. Rae, A. Pritzel, A. P. Badia, B. Uria, O. Vinyals, D. Hassabis, R. Pascanu, and C. Blundell, “Memory-based parameter adaptation,” in International Conference on Learning Representations , 2018
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
S. Hou, X. Pan, C. C. Loy, Z. Wang, and D. Lin, “Learning a unified classifier incrementally via rebalancing,” in Proceedings of the IEEE/CVF conference on Computer Vision and Pattern Recognition , 2019, pp. 831–839
2019
Earlier work this paper cites.
Y. Wu, Y. Chen, L. Wang, Y. Ye, Z. Liu, Y. Guo, and Y. Fu, “Large scale incremental learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 374–382
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Minneapolis, Minnesota: Association for Computational Linguistics, Jun. 2019, pp. 4171–4186. [Online]. Available: https://aclanthology.org/N19-1423
2019
Earlier work this paper cites.
C. de Masson D’Autume, S. Ruder, L. Kong, and D. Yogatama, “Episodic memory in lifelong language learning,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
J. Rajasegaran, M. Hayat, S. H. Khan, F. S. Khan, and L. Shao, “Random path selection for continual learning,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. De Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-efficient transfer learning for nlp,” in International Conference on Machine Learning . PMLR, 2019, pp. 2790–2799
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Cited alongside, same era.
Z. Wang, Z. Zhang, C.-Y. Lee, H. Zhang, R. Sun, X. Ren, G. Su, V. Perot, J. Dy, and T. Pfister, “Learning to prompt for continual learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 139–149
2022
Later among the works it cites.
J. Zheng, Z. Liang, H. Chen, and Q. Ma, “Distilling causal effect from miscellaneous other-class for continual named entity recognition,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 3602–3615
2022
Later among the works it cites.
V. V. Ramasesh, A. Lewkowycz, and E. Dyer, “Effect of scale on catastrophic forgetting in neural networks,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Belouadah and A. Popescu, “Il2m: Class incremental learning with dual memory,” in Proceedings of the IEEE/CVF international conference on computer vision , 2019, pp. 583–592
2019
Cited alongside, same era.
A. Barbu, D. Mayo, J. Alverio, W. Luo, C. Wang, D. Gutfreund, J. Tenenbaum, and B. Katz, “Objectnet: A large-scale bias-controlled dataset for pushing the limits of object recognition models,” Advances in neural information processing systems , vol. 32, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga et al. , “Pytorch: An imperative style, high-performance deep learning library,” Advances in neural information processing systems , vol. 32, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
A. Prabhu, P. H. Torr, and P. K. Dokania, “Gdumb: A simple approach that questions our progress in continual learning,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part II 16 . Springer, 2020, pp. 524–540
2020
Cited alongside, same era.
P. Buzzega, M. Boschini, A. Porrello, D. Abati, and S. Calderara, “Dark experience for general continual learning: a strong, simple baseline,” Advances in neural information processing systems , vol. 33, pp. 15 920–15 930, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Later among the works it cites.
E. Arani, F. Sarfraz, and B. Zonooz, “Learning fast, learning slow: A general continual learning method based on complementary learning system,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
G. Kim, C. Xiao, T. Konishi, Z. Ke, and B. Liu, “A theoretical study on solving continual learning,” Advances in Neural Information Processing Systems , vol. 35, pp. 5065–5079, 2022
2022
Later among the works it cites.
G. Kim, B. Liu, and Z. Ke, “A multi-head model for continual learning via out-of-distribution replay,” in Conference on Lifelong Learning Agents . PMLR, 2022, pp. 548–563
2022
Later among the works it cites.
Z. Wang, Z. Zhang, S. Ebrahimi, R. Sun, H. Zhang, C.-Y. Lee, X. Ren, G. Su, V. Perot, J. Dy et al. , “Dualprompt: Complementary prompting for rehearsal-free continual learning,” in Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XXVI . Springer, 2022, pp. 631–648
2022
Later among the works it cites.
B. Ermis, G. Zappella, M. Wistuba, A. Rawal, and C. Archambeau, “Memory efficient continual learning with transformers,” Advances in Neural Information Processing Systems , vol. 35, pp. 10 629–10 642, 2022
2022
Later among the works it cites.
Y. Wang, Z. Huang, and X. Hong, “S-prompts learning with pre-trained transformers: An occam’s razor for domain incremental learning,” in Advances in Neural Information Processing Systems , 2022
2022
Later among the works it cites.
Z. Ke, Y. Shao, H. Lin, H. Xu, L. Shu, and B. Liu, “Adapting a language model while preserving its general knowledge,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 10 177–10 188
2022
Later among the works it cites.
Z. Ke, Y. Shao, H. Lin, T. Konishi, G. Kim, and B. Liu, “Continual pre-training of language models,” in The Eleventh International Conference on Learning Representations , 2022
2022
Later among the works it cites.
J. Jang, S. Ye, S. Yang, J. Shin, J. Han, K. Gyeonghun, S. J. Choi, and M. Seo, “Towards continual knowledge learning of language models,” in International Conference on Learning Representations , 2022
2022
Later among the works it cites.
Y. Xia, Q. Wang, Y. Lyu, Y. Zhu, W. Wu, S. Li, and D. Dai, “Learn and review: Enhancing continual named entity recognition via reviewing synthetic samples,” in Findings of the Association for Computational Linguistics: ACL 2022 , 2022, pp. 2291–2300
2022
Later among the works it cites.
Y. Zhang, Z. Yin, J. Shao, and Z. Liu, “Benchmarking omni-vision representation through the lens of visual realms,” in European Conference on Computer Vision , 2022, pp. 594–611
2022
Later among the works it cites.
F.-Y. Wang, D.-W. Zhou, H.-J. Ye, and D.-C. Zhan, “Foster: Feature boosting and compression for class-incremental learning,” in European conference on computer vision . Springer, 2022, pp. 398–414
2022
Later among the works it cites.
OpenAI, “Gpt-4 technical report,” ArXiv , vol. abs/2303.08774, 2023
2023
Later among the works it cites.
M. Tao, Y. Feng, and D. Zhao, “Can bert refrain from forgetting on sequential tasks? a probing study,” in The Eleventh International Conference on Learning Representations , 2023
2023
Later among the works it cites.
J. Chen, T. Nguyen, D. Gorur, and A. Chaudhry, “Is forgetting less a good inductive bias for forward transfer?” in The Eleventh International Conference on Learning Representations , 2023
2023
Later among the works it cites.
D. Kim and B. Han, “On the stability-plasticity dilemma of class-incremental learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 20 196–20 204
2023
Later among the works it cites.
2023
Later among the works it cites.
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig, “Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing,” ACM Computing Surveys , vol. 55, no. 9, pp. 1–35, 2023
2023
Later among the works it cites.
A. Razdaibiedina, Y. Mao, R. Hou, M. Khabsa, M. Lewis, and A. Almahairi, “Progressive prompts: Continual learning for language models,” in The Eleventh International Conference on Learning Representations , 2023
2023
Later among the works it cites.
Y. Zhang, P. Li, M. Sun, and Y. Liu, “Continual knowledge distillation for neural machine translation,” in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Toronto, Canada: Association for Computational Linguistics, Jul. 2023, pp. 7978–7996. [Online]. Available: https://aclanthology.org/2023.acl-long.443
2023
Later among the works it cites.
H. Xia, P. Wang, T. Liu, B. Lin, Y. Cao, and Z. Sui, “Enhancing continual relation extraction via classifier decomposition,” in Findings of the Association for Computational Linguistics: ACL 2023 . Toronto, Canada: Association for Computational Linguistics, Jul. 2023, pp. 10 053–10 062. [Online]. Available: https://aclanthology.org/2023.findings-acl.638
2023
Later among the works it cites.
2023
Later among the works it cites.
M. Mundt, K. W. Cooper, D. S. Dhami, A. Ribeiro, J. S. Smith, A. Bellot, and T. Hayes, “Continual causality: A retrospective of the inaugural aaai-23 bridge program,” in AAAI Bridge Program on Continual Causality . PMLR, 2023, pp. 1–10
2023
Later among the works it cites.
T. Gong, T. Gerstenberg, R. Mayrhofer, and N. R. Bramley, “Active causal structure learning in continuous time,” Cognitive Psychology , vol. 140, p. 101542, 2023
2023
Later among the works it cites.
F.-Y. Wang, D.-W. Zhou, L. Liu, H.-J. Ye, Y. Bian, D.-C. Zhan, and P. Zhao, “Beef: Bi-compatible class-incremental learning via energy-based expansion and fusion,” in The Eleventh International Conference on Learning Representations , 2023
2023
Later among the works it cites.
G. Kim, C. Xiao, T. Konishi, and B. Liu, “Learnability and algorithm for continual learning,” in International Conference on Machine Learning, ICML 2023, 23-29 July 2023, Honolulu, Hawaii, USA , ser. Proceedings of Machine Learning Research, A. Krause, E. Brunskill, K. Cho, B. Engelhardt, S. Sabato, and J. Scarlett, Eds., vol. 202. PMLR, 2023, pp. 16 877–16 896. [Online]. Available: https://proceedings.mlr.press/v202/kim23x.html
2023
Later among the works it cites.