Fetching the paper…
Reading the bibliography…
Catastrophic forgetting is a problem faced by many machine learning models and algorithms.
Catastrophic interference in connectionist networks: The sequential learning problem
McCloskey, M. and Cohen, N. J · 1989
Earlier work this paper cites.
Connectionist models of recognition memory: constraints imposed by learning and forgetting functions
Ratcliff, R · 1990
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: Insights from the successes and failures of connectionist models of learning and memory
McClelland, J. L., McNaughton, B. L., and O’Reilly, R. C · 1995
Earlier work this paper cites.
Rank, trace-norm and max-norm
Srebro, Nathan and Shraibman, Adi · 2005
Earlier work this paper cites.
Biographies, bollywood, boom-boxes and blenders: Domain adaptation for sentiment classification
Blitzer, John, Dredze, Mark, and Pereira, Fernando · 2007
Earlier work this paper cites.
What is the best multi-stage architecture for object recognition?
Jarrett, Kevin, Kavukcuoglu, Koray, Ranzato, Marc’Aurelio, and LeCun, Yann · 2009
Cited alongside, same era.
Theano: a CPU and GPU math expression compiler
Bergstra, James, Breuleux, Olivier, Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Desjardins, Guillaume, Turian, Joseph, Warde-Farley, David, and Bengio, Yoshua · 2010
Cited alongside, same era.
Deep sparse rectifier neural networks
Glorot, Xavier, Bordes, Antoine, and Bengio, Yoshua · 2011
Cited alongside, same era.
Learning recurrent neural networks with Hessian-free optimization
Martens, James and Sutskever, Ilya · 2011
Cited alongside, same era.
Theano: new features and speed improvements
Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Bergstra, James, Goodfellow, Ian J., Bergeron, Arnaud, Bouchard, Nicolas, and Bengio, Yoshua · 2012
Cited alongside, same era.
Domain adaptation for large-scale sentiment classification: A deep learning approach
Random search for hyper-parameter optimization
Bergstra, James and Bengio, Yoshua · 2012
Later among the works it cites.
Improving neural networks by preventing co-adaptation of feature detectors
Hinton, Geoffrey E., Srivastava, Nitish, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan · 2012
Later among the works it cites.
Maxout networks
Goodfellow, Ian J., Warde-Farley, David, Mirza, Mehdi, Courville, Aaron, and Bengio, Yoshua · 2013
Closest in time.
Improving neural networks with dropout
Srivastava, Nitish · 2013
Closest in time.
Compete to compute
Srivastava, Rupesh K, Masci, Jonathan, Kazerounian, Sohrob, Gomez, Faustino, and Schmidhuber, Jürgen · 2013
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Glorot, Xavier, Bordes, Antoine, and Bengio, Yoshua
Cited in the paper.
Pylearn2: a machine learning research library
Goodfellow, Ian J., Warde-Farley, David, Lamblin, Pascal, Dumoulin, Vincent, Mirza, Mehdi, Pascanu, Razvan, Bergstra, James, Bastien, Frédéric, and Bengio, Yoshua
Cited in the paper.
McClelland, James L
Cited in the paper.