Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, Agarwal S, Herbert-Voss A, Krueger G, Henighan T, Child R, Ramesh A, Ziegler D, Wu J, Winter C, Hesse C, Chen M, Sigler E, Litwin M, Gray S, Chess B, Clark J, Berner C, McCandlish S, Radford A, Sutskever I, Amodei D (2020) Language Models are Few-Shot Learners. In: Proceedings of NeurIPS-2020, vol 33, pp 1877–1901
1901
Earlier work this paper cites.
Zhang X, Shapiro P, Kumar G, McNamee P, Carpuat M, Duh K (2019c) Curriculum learning for domain adaptation in neural machine translation. In: Proceedings of NAACL, pp 1903–1915
1915
Earlier work this paper cites.
Pi T, Li X, Zhang Z, Meng D, Wu F, Xiao J, Zhuang Y (2016) Self-paced boost learning for classification. In: Proceedings of IJCAI, pp 1932–1938
1938
Earlier work this paper cites.
Richter S, DeCarlo R (1983) Continuation methods: Theory and applications. IEEE Transactions on Automatic Control 28(6):660–665
1983
Earlier work this paper cites.
McCloskey M, Cohen NJ (1989) Catastrophic interference in connectionist networks: The sequential learning problem. In: Psychology of Learning and Motivation, vol 24, Elsevier, pp 109–165
1989
Earlier work this paper cites.
Chow J, Udpa L, Udpa S (1991) Homotopy continuation methods for neural networks. In: Proceedings of ISCAS, pp 2483–2486
1991
Earlier work this paper cites.
Elman JL (1993) Learning and development in neural networks: the importance of starting small. Cognition 48(1):71–99
1993
Earlier work this paper cites.
Mitchell TM (1997) Machine Learning. McGraw-Hill, New York
1997
Earlier work this paper cites.
Allgower EL, Georg K (2003) Introduction to numerical continuation methods. SIAM, DOI 10.1137/1.9780898719154.fm
2003
Earlier work this paper cites.
Bengio Y, Louradour J, Collobert R, Weston J (2009) Curriculum learning. In: Proceedings of ICML, pp 41–48
2009
Earlier work this paper cites.
Spitkovsky VI, Alshawi H, Jurafsky D (2009) Baby Steps: How “Less is More” in unsupervised dependency parsing. In: Proceedings of NIPS Workshop on Grammar Induction, Representation of Language and Language Learning
2009
Earlier work this paper cites.
Kumar MP, Packer B, Koller D (2010) Self-paced learning for latent variable models. In: Proceedings of NIPS, pp 1189–1197
2010
Earlier work this paper cites.
Kumar MP, Turki H, Preston D, Koller D (2011) Learning specific-class segmentation from diverse data. In: Proceedings of ICCV, pp 1800–1807
2011
Earlier work this paper cites.
Lee YJ, Grauman K (2011) Learning the easy things first: Self-paced visual category discovery. In: Proceedings of CVPR, pp 1721–1728
2011
Earlier work this paper cites.
Krizhevsky A, Sutskever I, Hinton GE (2012) ImageNet Classification with Deep Convolutional Neural Networks. In: Proceedings of NIPS, pp 1106–1114
2012
Earlier work this paper cites.
Supancic JS, Ramanan D (2013) Self-paced learning for long-term tracking. In: Proceedings of CVPR, pp 2379–2386
2013
Earlier work this paper cites.
Zhang XL, Wu J (2013) Denoising deep neural networks based voice activity detection. In: Proceedings of ICASSP, pp 853–857
2013
Earlier work this paper cites.
Simonyan K, Zisserman A (2014) Very Deep Convolutional Networks for Large-Scale Image Recognition. In: Proceedings of ICLR
2014
Earlier work this paper cites.
Zaremba W, Sutskever I (2014) Learning to execute. arXiv preprint arXiv:14104615
2014
Earlier work this paper cites.
Chen X, Gupta A (2015) Webly supervised learning of convolutional networks. In: Proceedings of ICCV, pp 1431–1439
2015
Earlier work this paper cites.
Jiang L, Meng D, Zhao Q, Shan S, Hauptmann AG (2015) Self-paced curriculum learning. In: Proceedings of AAAI, pp 2694–2700
2015
Earlier work this paper cites.
Pentina A, Sharmanska V, Lampert CH (2015) Curriculum learning of multiple tasks. In: Proceedings of CVPR, pp 5492–5500
2015
Earlier work this paper cites.
Ronneberger O, Fischer P, Brox T (2015) U-Net: Convolutional Networks for Biomedical Image Segmentation. In: Proceedings of MICCAI, pp 234–241
2015
Earlier work this paper cites.
Russakovsky O, Deng J, Su H, Krause J, Satheesh S, Ma S, Huang Z, Karpathy A, Khosla A, Bernstein M, Berg AC, Fei-Fei L (2015) ImageNet Large Scale Visual Recognition Challenge. International Journal of Computer Vision 115(3):211–252
2015
Earlier work this paper cites.
Shi Y, Larson M, Jonker CM (2015) Recurrent neural network language model adaptation with curriculum learning. Computer Speech & Language 33(1):136–154
2015
Earlier work this paper cites.
Szegedy C, Liu W, Jia Y, Sermanet P, Reed S, Anguelov D, Erhan D, Vanhoucke V, Rabinovich A (2015) Going Deeper With Convolutions. Proceedings of CVPR
2015
Earlier work this paper cites.
Xu C, Tao D, Xu C (2015) Multi-view self-paced learning for clustering. In: Proceedings of IJCAI, pp 3974–3980
2015
Earlier work this paper cites.
Zhao Q, Meng D, Jiang L, Xie Q, Xu Z, Hauptmann AG (2015) Self-paced learning for matrix factorization. In: Proceedings of AAAI, vol 3, p 4
2015
Earlier work this paper cites.
Amodei D, Ananthanarayanan S, Anubhai R, Bai J, Battenberg E, Case C, Casper J, Catanzaro B, Cheng Q, Chen G, et al. (2016) Deep speech 2: End-to-end speech recognition in English and Mandarin. In: Proceedings of ICML, pp 173–182
2016
Earlier work this paper cites.
Cirik V, Hovy E, Morency LP (2016) Visualizing and understanding curriculum learning for long short-term memory networks. arXiv preprint arXiv:161106204
2016
Earlier work this paper cites.
Gong C, Tao D, Maybank SJ, Liu W, Kang G, Yang J (2016) Multi-modal curriculum learning for semi-supervised image classification. IEEE Transactions on Image Processing 25(7):3249–3260
2016
Earlier work this paper cites.
Graves A, Wayne G, Reynolds M, Harley T, Danihelka I, Grabska-Barwińska A, Colmenarejo SG, Grefenstette E, Ramalho T, Agapiou J, et al. (2016) Hybrid computing using a neural network with dynamic external memory. Nature 538(7626):471–476
2016
Earlier work this paper cites.
He K, Zhang X, Ren S, Sun J (2016) Deep Residual Learning for Image Recognition. In: Proceedings of CVPR, pp 770–778
2016
Earlier work this paper cites.
Ionescu R, Alexe B, Leordeanu M, Popescu M, Papadopoulos DP, Ferrari V (2016) How hard can it be? estimating the difficulty of visual search in an image. In: Proceedings of CVPR, pp 2157–2166
2016
Earlier work this paper cites.
Kim Y, Jernite Y, Sontag D, Rush AM (2016) Character-Aware Neural Language Models. In: Proceedings of AAAI, pp 2741–2749
2016
Earlier work this paper cites.
Li H, Gong M, Meng D, Miao Q (2016) Multi-objective self-paced learning. In: Proceedings of AAAI, pp 1802–1808
2016
Earlier work this paper cites.
Liang J, Jiang L, Meng D, Hauptmann AG (2016) Learning to detect concepts from webly-labeled video data. In: Proceedings of IJCAI, pp 1746–1752
2016
Earlier work this paper cites.
Narvekar S, Sinapov J, Leonetti M, Stone P (2016) Source task creation for curriculum learning. In: Proceedings of AAMAS, pp 566–574
2016
Earlier work this paper cites.
Sachan M, Xing E (2016) Easy questions first? a case study on curriculum learning for question answering. In: Proceedings of ACL, pp 453–463
2016
Earlier work this paper cites.
Shi M, Ferrari V (2016) Weakly supervised object localization using size estimates. In: Proceedings of ECCV, Springer, pp 105–121
2016
Earlier work this paper cites.
Shrivastava A, Gupta A, Girshick R (2016) Training region-based object detectors with online hard example mining. In: Proceedings of CVPR, pp 761–769
2016
Earlier work this paper cites.
Tsvetkov Y, Faruqui M, Ling W, MacWhinney B, Dyer C (2016) Learning the Curriculum with Bayesian Optimization for Task-Specific Word Representation Learning. In: Proceedings of ACL, pp 130–139
2016
Earlier work this paper cites.
Braun S, Neil D, Liu SC (2017) A curriculum learning method for improved noise robustness in automatic speech recognition. In: Proceedings of EUSIPCO, pp 548–552
2017
Earlier work this paper cites.
Chang HS, Learned-Miller E, McCallum A (2017) Active bias: Training more accurate neural networks by emphasizing high variance samples. In: Proceedings of NIPS, pp 1002–1012
2017
Earlier work this paper cites.
Fan Y, He R, Liang J, Hu B (2017) Self-paced learning: An implicit regularization perspective. In: Proceedings of AAAI, pp 1877–1883
2017
Earlier work this paper cites.
Florensa C, Held D, Wulfmeier M, Zhang M, Abbeel P (2017) Reverse curriculum generation for reinforcement learning. In: Proceedings of CoRL, vol 78, pp 482–495
2017
Earlier work this paper cites.
Graves A, Bellemare MG, Menick J, Munos R, Kavukcuoglu K (2017) Automated curriculum learning for neural networks. In: Proceedings of ICML, vol 70, pp 1311–1320
2017
Earlier work this paper cites.
Gui L, Baltrušaitis T, Morency LP (2017) Curriculum learning for facial expression recognition. In: Proceedings of FG, pp 505–511
2017
Earlier work this paper cites.
Jesson A, Guizard N, Ghalehjegh SH, Goblot D, Soudan F, Chapados N (2017) CASED: curriculum adaptive sampling for extreme data imbalance. In: Proceedings of MICCAI, pp 639–646
2017
Earlier work this paper cites.
Kocmi T, Bojar O (2017) Curriculum learning and minibatch bucketing in neural machine translation. In: Proceedings of RANLP, pp 379–386
2017
Earlier work this paper cites.
Li H, Gong M (2017) Self-paced convolutional neural networks. In: Proceedings of IJCAI, pp 2110–2116
2017
Earlier work this paper cites.
Lin L, Wang K, Meng D, Zuo W, Zhang L (2017) Active self-paced learning for cost-effective and progressive face identification. IEEE Transactions on Pattern Analysis and Machine Intelligence 40(1):7–19
2017
Earlier work this paper cites.
Lotter W, Sorensen G, Cox D (2017) A multi-scale CNN and curriculum learning strategy for mammogram classification. In: Proceedings of DLMIA and ML-CDS, pp 169–177
2017
Earlier work this paper cites.
Ma F, Meng D, Xie Q, Li Z, Dong X (2017) Self-paced co-training. In: Proceedings of ICML, pp 2275–2284
2017
Earlier work this paper cites.
Morerio P, Cavazza J, Volpi R, Vidal R, Murino V (2017) Curriculum dropout. In: Proceedings of ICCV, pp 3544–3552
2017
Earlier work this paper cites.
Ranjan S, Hansen JH (2017) Curriculum learning based approaches for noise robust speaker recognition. IEEE/ACM Transactions on Audio, Speech, and Language Processing 26(1):197–210
2017
Earlier work this paper cites.
Ren Y, Zhao P, Sheng Y, Yao D, Xu Z (2017) Robust softmax regression for multi-class classification with self-paced learning. In: Proceedings of IJCAI, pp 2641–2647
2017
Earlier work this paper cites.