Cutmix: Regularization strategy to train strong classifiers with localizable features
Sangdoo Yun, Dongyoon Han, Seong Joon Oh, Sanghyuk Chun, Junsuk Choe, and Youngjoon Yoo · 2019
Later among the works it cites.
M2kd: Multi-model and multi-level knowledge distillation for incremental learning
Original
Peng Zhou, Long Mai, Jianming Zhang, Ning Xu, Zuxuan Wu, and Larry S. Davis · 2019
Later among the works it cites.
Modeling the background for incremental learning in semantic segmentation
Fabio Cermelli, Massimiliano Mancini, Samuel Rota Bulò, Elisa Ricci, and Barbara Caputo · 2020
Later among the works it cites.
Routing networks with co-training for continual learning
Mark Patrick Collier, Effrosyni Kokiopoulou, Andrea Gesmundo, and Jesse Berent · 2020
Later among the works it cites.
Podnet: Pooled outputs distillation for small-tasks incremental learning
Arthur Douillard, Matthieu Cord, Charles Ollion, Thomas Robert, and Eduardo Valle · 2020
Later among the works it cites.
Orthogonal gradient descent for continual learning
Mehrdad Farajtabar, Navid Azizan, Alex Mott, and Ang Li · 2020
Later among the works it cites.
Remind your neural network to prevent catastrophic forgetting
Tyler L. Hayes, Kushal Kafle, Robik Shrestha, Manoj Acharya, and Christopher Kanan · 2020
Later among the works it cites.
Memory-efficient incremental learning through feature adaptation
Ahmet Iscen, Jeffrey Zhang, Svetlana Lazebnik, and Cordelia Schmid · 2020
Later among the works it cites.
Continual learning with extended kronecker-factored approximate curvature
Janghyeon Lee, Hyeong Gwon Hong, Donggyu Joo, and Junmo Kim · 2020
Later among the works it cites.
Batchensemble: An alternative approach to efficient ensemble and lifelong learning
Yeming Wen, Dustin Tran, and Jimmy Ba · 2020
Later among the works it cites.
Maintaining discrimination and fairness in class incremental learning
Bowen Zhao, Xi Xiao, Guojun Gan, Bin Zhang, and Shutao Xia · 2020
Later among the works it cites.
Convit: Improving vision transformers with soft convolutional inductive biases
Stéphane d’Ascoli, Hugo Touvron, Matthew Leavitt, Giulio Morcos, Ari annd Biroli, and Levent Sagun · 2021
Closest in time.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Closest in time.
Plop: Learning without forgetting for continual semantic segmentation
Arthur Douillard, Yifu Chen, Arnaud Dapogny, and Matthieu Cord · 2021
Closest in time.
Tackling catastrophic forgetting and background shift in continual semantic segmentation
Arthur Douillard, Yifu Chen, Arnaud Dapogny, and Matthieu Cord · 2021
Closest in time.
Continuum: Simple management of complex continual learning scenarios
Arthur Douillard and Timothée Lesort · 2021
Closest in time.
Asam: Adaptive sharpness-aware minimization for scale-invariant learning of deep neural networks
Jungmin Kwon, Jeongseop Kim, Hyunseo Park, and In Kwon Choi · 2021
Closest in time.
Preserving earlier knowledge in continual learning with the help of all previous feature extractors
Zhuoyun Li, Changhong Zhong, Sijia Liu, Ruixuan Wang, and Wei-Shi Zheng · 2021
Closest in time.
Sharpness-aware minimization in large-batch training: Training vision transformer in minutes
Yong Liu, Siqi Mai, Xiangning Chen, Cho-Jui Hsieh, and Yang You · 2021
Closest in time.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Closest in time.
Gradient projection memory for continual learning
Gobinda Saha, Isha Garg, and Kaushik Roy · 2021
Closest in time.
Ensembles and encoders for task-free continual learning
Murray Shanahan, Christos Kaplanis, and Jovana Mitrović · 2021
Closest in time.
Training data-efficient image transformers and distillation through attention
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Hervé Jégou · 2021
Closest in time.
Going deeper with image transformers
Hugo Touvron, Matthieu Cord, Alexandre Sablayrolles, Gabriel Synnaeve, and Hervé Jégou · 2021
Closest in time.
Rehearsal revealed: The limits and merits of revisiting samples in continual learning
Eli Verwimp, Matthias De Lange, and Tinne Tuytelaars · 2021
Closest in time.
Der: Dynamically expandable representation for class incremental learning
Shipeng Yan, Jiangwei Xie, and Xuming He · 2021
Closest in time.
Perceiver io: A general architecture for structured inputs & outputs
Andrew Jaegle, Sebastian Borgeaud, Jean-Baptiste Alayrac, Carl Doersch, Catalin Ionescu, David Ding, Skanda Koppula, Daniel Zoran, Andrew Brock, Evan Shelhamer, Olivier Hénaff, Matthew M. Botvinick, Andrew Zisserman, Oriol Vinyals, and João Carreira · 2022
Closest in time.