Fetching the paper…
Reading the bibliography…
Predictive models -- learned from observational data not covering the complete data distribution -- can rely on spurious correlations in the data for making predictions.
Random search and reproducibility for neural architecture search
Li, Liam and Ameet Talwalkar (2019) · 1902
Earlier work this paper cites.
Arjovsky, Martin, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz (2019) · 1907
Earlier work this paper cites.
Learning neural causal models from unknown interventions
Ke, Nan Rosemary, Olexa Bilaniuk, Anirudh Goyal, Stefan Bauer, Hugo Larochelle, Chris Pal, and Yoshua Bengio (2019) · 1910
Earlier work this paper cites.
Note on a method for calculating corrected sums of squares and products
Welford, BP (1962) · 1962
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Werbos, Paul J (1990) · 1990
Earlier work this paper cites.
An efficient gradient-based algorithm for on-line training of recurrent network trajectories
Williams, Ronald J and Jing Peng (1990) · 1990
Earlier work this paper cites.
Reinforcement learning for robots using neural networks
Lin, Long-Ji (1993) · 1993
Earlier work this paper cites.
Invariant risk minimization games
Ahuja, Kartik, Karthikeyan Shanmugam, Kush Varshney, and Amit Dhurandhar (2020) · 2002
Earlier work this paper cites.
Out-of-distribution generalization via risk extrapolation (rex)
Krueger, David, Ethan Caballero, Joern-Henrik Jacobsen, Amy Zhang, Jonathan Binas, Remi Le Priol, and Aaron Courville (2020) · 2003
Earlier work this paper cites.
On the role of tracking in stationary environments
Sutton, Richard S, Anna Koop, and David Silver (2007) · 2007
Earlier work this paper cites.
Sample-based learning and search with permanent and transient memories
Silver, David, Richard S Sutton, and Martin Müller (2008) · 2008
Earlier work this paper cites.
Causality
Pearl, Judea (2009) · 2009
Cited alongside, same era.
Gnu parallel - the command-line power tool
Tange, O. (2011) · 2011
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Ilya Sutskever, and Geoffrey E Hinton (2012) · 2012
Cited alongside, same era.
Representation search through generate and test
Mahmood, Ashique Rupam and Richard S Sutton (2013) · 2013
Cited alongside, same era.
Training recurrent neural networks
Sutskever, Ilya (2013) · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, Diederik P and Jimmy Ba (2014) · 2014
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, Chelsea, Pieter Abbeel, and Sergey Levine (2017) · 2017
Later among the works it cites.
Building machines that learn and think like people
Lake, Brenden M, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman (2017) · 2017
Later among the works it cites.
Meta-sgd: Learning to learn quickly for few-shot learning
Li, Zhenguo, Fengwei Zhou, Fei Chen, and Hang Li (2017) · 2017
Later among the works it cites.
Incremental off-policy reinforcement learning algorithms
Mahmood, Ashique (2017) · 2017
Later among the works it cites.
Horde: A scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
Sutton, Richard S, Joseph Modayil, Michael Delp Thomas Degris, Patrick M Pilarski, and Adam White (2017) · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Human-level control through deep reinforcement learning
Mnih, Volodymyr, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al. (2015) · 2015
Cited alongside, same era.
Scientific method
Andersen, Hanne and Brian Hepburn (2016) · 2016
Cited alongside, same era.
Lopez-Paz, David (2016) · 2016
Cited alongside, same era.
Meta-learning with memory-augmented neural networks
Santoro, Adam, Sergey Bartunov, Matthew Botvinick, Daan Wierstra, and Timothy Lillicrap (2016) · 2016
Cited alongside, same era.
Sparse attentive backtracking: Temporal credit assignment through reminding
Ke, Nan Rosemary, Anirudh Goyal, Olexa Bilaniuk, Jonathan Binas, Michael C Mozer, Chris Pal, and Yoshua Bengio (2018) · 2018
Later among the works it cites.
Simple random search of static linear policies is competitive for reinforcement learning
Mania, Horia, Aurelia Guy, and Benjamin Recht (2018) · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Sutton, Richard S and Andrew G Barto (2018) · 2018
Later among the works it cites.
A meta-transfer objective for learning to disentangle causal mechanisms
Bengio, Yoshua, Tristan Deleu, Nasim Rahaman, Rosemary Ke, Sébastien Lachapelle, Olexa Bilaniuk, Anirudh Goyal, and Christopher Pal (2019) · 2019
Later among the works it cites.
Meta-learning representations for continual learning
Javed, Khurram and Martha White (2019) · 2019
Later among the works it cites.