Fetching the paper…
Reading the bibliography…
Discovering and exploiting the causal structure in the environment is a crucial challenge for intelligent agents.
The perception of causality in infants
A. M. Leslie · 1982
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Causality: Models, Reasoning, and Inference
J. Pearl · 2000
Earlier work this paper cites.
Causation, prediction, and search
P. Spirtes, C. N. Glymour, R. Scheines, D. Heckerman, C. Meek, G. Cooper, and T. Richardson · 2000
Earlier work this paper cites.
Causal learning mechanisms in very young children: two-, three-, and four-year-olds infer causal relations from patterns of variation and covariation
A. Gopnik, D. M. Sobel, L. E. Schulz, and C. Glymour · 2001
Earlier work this paper cites.
A theory of causal learning in children: causal maps and bayes nets
A. Gopnik, C. Glymour, D. M. Sobel, L. E. Schulz, T. Kushnir, and D. Danks · 2004
Earlier work this paper cites.
Causal reasoning in rats
A. P. Blaisdell, K. Sawa, K. J. Leising, and M. R. Waldmann · 2006
Earlier work this paper cites.
Fundamentals of statistical causality
P. Dawid · 2007
Earlier work this paper cites.
Learning a theory of causality
N. D. Goodman, T. D. Ullman, and J. B. Tenenbaum · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Experiment selection for causal discovery
A. Hyttinen, F. Eberhardt, and P. O. Hoyer · 2013
Earlier work this paper cites.
Causal responsibility and counterfactuals
D. A. Lagnado, T. Gerstenberg, and R. Zultan · 2013
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Cited alongside, same era.
Causal reasoning in a prediction task with hidden causes
P. A. Ortega and D. D. Lee A. A. Stocker · 2015
Cited alongside, same era.
Learning causal graphs with small interventions
K. Shanmugam, M. Kocaoglu, A. G. Dimakis, and S. Vishwanath · 2015
Cited alongside, same era.
Neural module networks
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Cited alongside, same era.
Learning to learn by gradient descent by gradient descent
M. Andrychowicz, M. Denil, S. Gomez, M. W. Hoffman, D. Pfau, T. Schaul, B. Shillingford, and N. De Freitas · 2016
Cited alongside, same era.
r l 2 rl^{2} : Fast reinforcement learning via slow reinforcement learning
Formalizing neurath’s ship: Approximate algorithms for online causal learning
N. R. Bramley, P. Dayan, T. L. Griffiths, and D. A. Lagnado · 2017
Later among the works it cites.
Model-agnostic meta-learning for fast adaptation of deep networks
C. Finn, P. Abbeel, and S. Levine · 2017
Later among the works it cites.
Counterfactual data-fusion for online reinforcement learners
A. Forney, J. Pearl, and E. Bareinboim · 2017
Later among the works it cites.
Deep q-learning from demonstrations
T. Hester, M. Vecerik, O. Pietquin, M. Lanctot, T. Schaul, B. Piot, D. Horgan, J. Quan, A. Sendonaris, G. Dulac-Arnold, et al · 2017
Later among the works it cites.
Identifying best interventions through online importance sampling
R. Sen, K. Shanmugam, A. G. Dimakis, and S. Shakkottai · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel · 2016
Cited alongside, same era.
Causal bandits: Learning good interventions via causal inference
F. Lattimore, T. Lattimore, and M. D. Reid · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Causal inference in statistics: a primer
J. Pearl, M. Glymour, and N. P. Jewell · 2016
Cited alongside, same era.
Meta-learning with memory-augmented neural networks
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap · 2016
Cited alongside, same era.
Matching networks for one shot learning
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra, et al · 2016
Cited alongside, same era.
Learning to reinforcement learn
J. X. Wang, Z. Kurth-Nelson, D. Tirumala, H. Soyer, J. Z. Leibo, R. Munos, C. Blundell, D. Kumaran, and M. Botvinick · 2016
Cited alongside, same era.
Later among the works it cites.
Relational inductive biases, deep learning, and graph networks
P. W. Battaglia, J. B. Hamrick, V. Bapst, A. Sanchez-Gonzalez, V. Zambaldi, M. Malinowski, A. Tacchetti, D. Raposo, A. Santoro, R. Faulkner, et al · 2018
Later among the works it cites.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning, et al · 2018
Later among the works it cites.
Synthesizing programs for images using reinforced adversarial learning
Y. Ganin, T. Kulkarni, I. Babuschkin, S. M. Eslami, and O. Vinyals · 2018
Later among the works it cites.
Recasting gradient-based meta-learning as hierarchical bayes
E. Grant, C. Finn, S. Levine, T. Darrell, and T. Griffiths · 2018
Later among the works it cites.
Multi-task deep reinforcement learning with popart
M. Hessel, H. Soyer, L. Espeholt, W. Czarnecki, S. Schmitt, and H. van Hasselt · 2018
Later among the works it cites.
A multivariate additive noise model for complete causal discovery
P. K. Parida, T. Marwala, and S. Chakraverty · 2018
Later among the works it cites.
Prefrontal cortex as a meta-reinforcement learning system
J. X. Wang, Z. Kurth-Nelson, D. Kumaran, D. Tirumala, H. Soyer, J. Z. Leibo, D. Hassabis, and M. Botvinick · 2018
Later among the works it cites.