Fetching the paper…
Reading the bibliography…
Modern multilingual automatic speech recognition (ASR) systems like Whisper have made it possible to transcribe audio in multiple languages with a single model.
J. S. Vitter, “Random sampling with a reservoir,” ACM Trans. Math. Softw. , vol. 11, no. 1, pp. 37–57, 1985
1985
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural Comput. , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
R. M. French, “Catastrophic forgetting in connectionist networks,” Trends in Cognitive Sciences , vol. 3, no. 4, pp. 128–135, 1999
1999
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in International Conference on Machine Learning (ICML) , 2006, pp. 369–376
2006
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images,” Technical report, 2009
2009
Earlier work this paper cites.
M. Mermillod, A. Bugaiska, and P. Bonin, “The stability-plasticity dilemma: Investigating the continuum from catastrophic forgetting to age-limited learning effects,” Frontiers in psychology , vol. 4, p. 504, 2013
2013
Earlier work this paper cites.
I. J. Goodfellow, M. Mirza, D. Xiao, A. Courville, and Y. Bengio, “An empirical investigation of catastrophic forgetting in gradient-based neural networks,” in International Conference on Learning Representations (ICLR) , 2014
2014
Earlier work this paper cites.
G. Hinton, O. Vinyals, and J. Dean, “Distilling the knowledge in a neural network,” in NeurIPS Deep Learning and Representation Learning Workshop , 2015
2015
Earlier work this paper cites.
M. P. Lewis, G. F. Simon, and C. D. Fennig, “Ethnologue: Languages of the world, nineteenth edition,” Online version: https://www.ethnologue.com , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
H. Shin, J. K. Lee, J. Kim, and J. Kim, “Continual learning with deep generative replay,” in International Conference on Neural Information Processing Systems (NeurIPS) , vol. 30, 2017, pp. 2994–3003
2017
Earlier work this paper cites.
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert, “iCaRL: Incremental classifier and representation learning,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 5533–5542
2017
Earlier work this paper cites.
R. Aljundi, P. Chakravarty, and T. Tuytelaars, “Expert gate: Lifelong learning with a network of experts,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 7120–7129
2017
Earlier work this paper cites.
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska, D. Hassabis, C. Clopath, D. Kumaran, and R. Hadsell, “Overcoming catastrophic forgetting in neural networks,” Proceedings of the National Academy of Sciences , vol. 114, no. 13, pp. 3521–3526, 2017
2017
Earlier work this paper cites.
F. Zenke, B. Poole, and S. Ganguli, “Continual learning through synaptic intelligence,” in International Conference on Machine Learning (ICML) , 2017, pp. 3987–3995
2017
Earlier work this paper cites.
V. Lomonaco and D. Maltoni, “CORe50: a new dataset and benchmark for continuous object recognition,” in Conference on Robot Learning , 2017, pp. 17–26
2017
Earlier work this paper cites.
S.-A. Rebuffi, H. Bilen, and A. Vedaldi, “Learning multiple visual domains with residual adapters,” International Conference on Neural Information Processing Systems (NeurIPS) , vol. 30, p. 506–516, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in International Conference on Neural Information Processing Systems (NeurIPS) , 2017, pp. 6000–6010
2017
Earlier work this paper cites.
D. Lopez-Paz and M. Ranzato, “Gradient episodic memory for continual learning,” in International Conference on Neural Information Processing Systems (NeurIPS) , 2017, pp. 6470–6479
2017
Earlier work this paper cites.
A. Mallya, D. Davis, and S. Lazebnik, “Piggyback: Adapting a single network to multiple tasks by learning to mask weights,” in European Conference on Computer Vision (ECCV) , 2018, pp. 72–88
2018
Earlier work this paper cites.
A. Mallya and S. Lazebnik, “PackNet: Adding multiple tasks to a single network by iterative pruning,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 7765–7773
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
T. Kudo and J. Richardson, “SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing,” in Conference on Empirical Methods in Natural Language Processing (EMNLP): System Demonstrations , 2018
2018
Earlier work this paper cites.
R. Aljundi, F. Babiloni, M. Elhoseiny, M. Rohrbach, and T. Tuytelaars, “Memory aware synapses: Learning what (not) to forget,” in European Conference on Computer Vision (ECCV) , 2018, pp. 144–161
2018
Earlier work this paper cites.
A. Chaudhry, P. K. Dokania, T. Ajanthan, and P. H. Torr, “Riemannian walk for incremental learning: Understanding forgetting and intransigence,” in European Conference on Computer Vision (ECCV) , 2018, pp. 532–547
2018
Earlier work this paper cites.
D. Rolnick, A. Ahuja, J. Schwarz, T. P. Lillicrap, and G. Wayne, “Experience replay for continual learning,” International Conference on Neural Information Processing Systems (NeurIPS) , vol. 32, pp. 350–360, 2019
2019
Cited alongside, same era.
A. Chaudhry, M. Ranzato, M. Rohrbach, and M. Elhoseiny, “Efficient lifelong learning with A-GEM,” in International Conference on Learning Representations (ICLR) , 2019
2019
Cited alongside, same era.
S. Hou, X. Pan, C. C. Loy, Z. Wang, and D. Lin, “Learning a unified classifier incrementally via rebalancing,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 831–839
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” Technical report, 2019
2019
Cited alongside, same era.
J. Gallardo, T. L. Hayes, and C. Kanan, “Self-supervised training enhances online continual learning,” in British Machine Vision Conference , 2021
2021
Later among the works it cites.
K. J. Liang, W. Hao, D. Shen, Y. Zhou, W. Chen, C. Chen, and L. Carin, “MixKD: Towards efficient distillation of large-scale language models,” in International Conference on Learning Representations (ICLR) , 2021
2021
Later among the works it cites.
2022
Later among the works it cites.
B. Li, R. Pang, Y. Zhang, T. N. Sainath, T. Strohman, P. Haghani, Y. Zhu, B. Farris, N. Gaur, and M. Prasad, “Massively multilingual ASR: A lifelong learning solution,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2022, pp. 6397–6401
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
I. Loshchilov and F. Hutter, “Decoupled weight decay regularization,” in International Conference on Learning Representations (ICLR) , 2019
2019
Cited alongside, same era.
D. Rao, F. Visin, A. Rusu, R. Pascanu, Y. W. Teh, and R. Hadsell, “Continual unsupervised representation learning,” in International Conference on Neural Information Processing Systems (NeurIPS) , vol. 32, 2019
2019
Cited alongside, same era.
2020
Cited alongside, same era.
A. Baevski, Y. Zhou, A. Mohamed, and M. Auli, “wav2vec 2.0: A framework for self-supervised learning of speech representations,” in International Conference on Neural Information Processing Systems (NeurIPS) , vol. 33, 2020, pp. 12 449–12 460
2020
Cited alongside, same era.
R. Ardila, M. Branson, K. Davis, M. Kohler, J. Meyer, M. Henretty, R. Morais, L. Saunders, F. Tyers, and G. Weber, “Common Voice: A massively-multilingual speech corpus,” in Twelfth Language Resources and Evaluation Conference , 2020, pp. 4218–4222
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. Sadhu and H. Hermansky, “Continual learning in automatic speech recognition.” in Interspeech , 2020, pp. 1246–1250
2020
Cited alongside, same era.
P. Buzzega, M. Boschini, A. Porrello, D. Abati, and S. Calderara, “Dark experience for general continual learning: a strong, simple baseline,” in International Conference on Neural Information Processing Systems (NeurIPS) , vol. 33, 2020, pp. 15 920–15 930
2020
Cited alongside, same era.
O. Ostapenko, T. Lesort, P. Rodriguez, M. R. Arefin, A. Douillard, I. Rish, and L. Charlin, “Continual learning with foundation models: An empirical study of latent replay,” in Conference on Lifelong Learning Agents , vol. 199, 2022, pp. 60–91
2022
Later among the works it cites.
S. Chen, C. Wang, Z. Chen, Y. Wu, S. Liu, Z. Chen, J. Li, N. Kanda, T. Yoshioka, X. Xiao, J. Wu, L. Zhou, S. Ren, Y. Qian, Y. Qian, J. Wu, M. Zeng, X. Yu, and F. Wei, “WavLM: Large-scale self-supervised pre-training for full stack speech processing,” IEEE Journal of Selected Topics in Signal Processing , vol. 16, no. 6, pp. 1505–1518, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Mai, R. Li, J. Jeong, D. Quispe, H. Kim, and S. Sanner, “Online continual learning in image classification: An empirical survey,” Neurocomputing , vol. 469, pp. 28–51, 2022
2022
Later among the works it cites.
M. Masana, X. Liu, B. Twardowski, M. Menta, A. D. Bagdanov, and J. van de Weijer, “Class-incremental learning: survey and performance evaluation on image classification,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
T. Srinivasan, T.-Y. Chang, L. Pinto Alva, G. Chochlakis, M. Rostami, and J. Thomason, “CLiMB: A continual learning benchmark for vision-and-language tasks,” in International Conference on Neural Information Processing Systems (NeurIPS) , vol. 35, 2022, pp. 29 440–29 453
2022
Later among the works it cites.
G. M. van de Ven, T. Tuytelaars, and A. S. Tolias, “Three types of incremental learning,” Nature Machine Intelligence , pp. 1–13, 2022
2022
Later among the works it cites.
C.-C. Chiu, J. Qin, Y. Zhang, J. Yu, and Y. Wu, “Self-supervised learning with random-projection quantizer for speech recognition,” in International Conference on Machine Learning (ICML) , 2022, pp. 3915–3924
2022
Later among the works it cites.
D. Hwang, K. C. Sim, Z. Huo, and T. Strohman, “Pseudo label is better than human label,” in Interspeech , 2022, pp. 1421–1425
2022
Later among the works it cites.
S. Kessler, B. Thomas, and S. Karout, “An adapter based pre-training for efficient and scalable self-supervised speech representation learning,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2022, pp. 3179–3183
2022
Later among the works it cites.
Z. Wang, Z. Zhang, C.-Y. Lee, H. Zhang, R. Sun, X. Ren, G. Su, V. Perot, J. Dy, and T. Pfister, “Learning to prompt for continual learning,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 139–149
2022
Later among the works it cites.
D. Madaan, J. Yoon, Y. Li, Y. Liu, and S. J. Hwang, “Representational continuity for unsupervised continual learning,” in International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
N. Goyal, C. Gao, V. Chaudhary, P.-J. Chen, G. Wenzek, D. Ju, S. Krishnan, M. Ranzato, F. Guzmán, and A. Fan, “The FLoRes-101 evaluation benchmark for low-resource and multilingual machine translation,” Transactions of the Association for Computational Linguistics , vol. 10, pp. 522–538, 2022
2022
Later among the works it cites.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
B. Li, D. Hwang, Z. Huo, J. Bai, G. Prakash, T. N. Sainath, K. Chai Sim, Y. Zhang, W. Han, T. Strohman, and F. Beaufays, “Efficient domain adaptation for speech foundation models,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2023, pp. 1–5
2023
Closest in time.
2023
Closest in time.
A. Conneau, M. Ma, S. Khanuja, Y. Zhang, V. Axelrod, S. Dalmia, J. Riesa, C. Rivera, and A. Bapna, “FLEURS: Few-shot learning evaluation of universal representations of speech,” in IEEE Spoken Language Technology Workshop (SLT) , 2023, pp. 798–805
2023
Closest in time.