Fetching the paper…
Reading the bibliography…
Recent work in machine learning and cognitive science has suggested that understanding causal information is essential to the development of intelligence.
Curious model-building control systems
J. Schmidhuber · 1991
Earlier work this paper cites.
Q-learning
C. J. Watkins and P. Dayan · 1992
Earlier work this paper cites.
Detecting blickets: how young children use information about novel causal powers in categorization and induction
A. Gopnik and D. M. Sobel · 2000
Earlier work this paper cites.
Theory-based causal induction
T. L. Griffiths and J. B. Tenenbaum · 2009
Earlier work this paper cites.
Scientific thinking in young children: Theoretical advances, empirical research, and policy implications
A. Gopnik · 2012
Earlier work this paper cites.
When children are better (or at least more open-minded) learners than adults: Developmental differences in learning the forms of causal relationships
C. G. Lucas, S. Bridgers, T. L. Griffiths, and A. Gopnik · 2014
Earlier work this paper cites.
Unifying count-based exploration and intrinsic motivation
M. Bellemare, S. Srinivasan, G. Ostrovski, T. Schaul, D. Saxton, and R. Munos · 2016
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Earlier work this paper cites.
Deep exploration via bootstrapped dqn
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy · 2016
Earlier work this paper cites.
Changes in cognitive flexibility and hypothesis search across human life history from childhood to adolescence to adulthood
A. Gopnik, S. O’Grady, C. G. Lucas, T. L. Griffiths, A. Wente, S. Bridgers, R. Aboody, H. Fung, and R. E. Dahl · 2017
Earlier work this paper cites.
Count-based exploration in feature space for reinforcement learning
J. Martin, S. N. Sasikumar, T. Everitt, and M. Hutter · 2017
Earlier work this paper cites.
Count-based exploration with neural density models
G. Ostrovski, M. G. Bellemare, A. van den Oord, and R. Munos · 2017
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Earlier work this paper cites.
# exploration: A study of count-based exploration for deep reinforcement learning
H. Tang, R. Houthooft, D. Foote, A. Stooke, O. X. Chen, Y. Duan, J. Schulman, F. DeTurck, and P. Abbeel · 2017
Earlier work this paper cites.
Exploration by random network distillation
Y. Burda, H. Edwards, A. Storkey, and O. Klimov · 2018
Cited alongside, same era.
Babyai: A platform to study the sample efficiency of grounded language learning
M. Chevalier-Boisvert, D. Bahdanau, S. Lahlou, L. Willems, C. Saharia, T. H. Nguyen, and Y. Bengio · 2018
Cited alongside, same era.
Quantifying generalization in reinforcement learning
K. Cobbe, O. Klimov, C. Hesse, T. Kim, and J. Schulman · 2018
Cited alongside, same era.
Stable baselines
A. Hill, A. Raffin, M. Ernestus, A. Gleave, A. Kanervisto, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu · 2018
Cited alongside, same era.
Gotta learn fast: A new benchmark for generalization in rl
A. Nichol, V. Pfau, C. Hesse, O. Klimov, and J. Schulman · 2018
Exploring exploration: Comparing children with rl agents in unified environments
E. Kosoy, J. Collins, D. M. Chan, S. Huang, D. Pathak, P. Agrawal, J. Canny, A. Gopnik, and J. B. Hamrick · 2020
Later among the works it cites.
Adapting text embeddings for causal inference
V. Veitch, D. Sridhar, and D. Blei · 2020
Later among the works it cites.
A survey of exploration methods in reinforcement learning
S. Amin, M. Gomrokchi, H. Satija, H. van Hoof, and D. Precup · 2021
Later among the works it cites.
Decision transformer: Reinforcement learning via sequence modeling
L. Chen, K. Lu, A. Rajeswaran, K. Lee, A. Grover, M. Laskin, P. Abbeel, A. Srinivas, and I. Mordatch · 2021
Later among the works it cites.
Neural production systems
A. Goyal, A. Didolkar, N. R. Ke, C. Blundell, P. Beaudoin, N. Heess, M. C. Mozer, and Y. Bengio · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Phyre: A new benchmark for physical reasoning
A. Bakhtin, L. van der Maaten, J. Johnson, L. Gustafson, and R. Girshick · 2019
Cited alongside, same era.
A meta-transfer objective for learning to disentangle causal mechanisms
Y. Bengio, T. Deleu, N. Rahaman, R. Ke, S. Lachapelle, O. Bilaniuk, A. Goyal, and C. Pal · 2019
Cited alongside, same era.
Recurrent independent mechanisms
A. Goyal, A. Lamb, J. Hoffmann, S. Sodhani, S. Levine, Y. Bengio, and B. Schölkopf · 2019
Cited alongside, same era.
Self-supervised exploration via disagreement
D. Pathak, D. Gandhi, and A. Gupta · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever, et al · 2019
Cited alongside, same era.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine · 2019
Cited alongside, same era.
Causalworld: A robotic manipulation benchmark for causal structure and transfer learning
O. Ahmed, F. Träuble, A. Goyal, A. Neitz, Y. Bengio, B. Schölkopf, M. Wüthrich, and S. Bauer · 2020
Cited alongside, same era.
Later among the works it cites.
Systematic evaluation of causal discovery in visual model based reinforcement learning
N. R. Ke, A. R. Didolkar, S. Mittal, A. Goyal, G. Lajoie, S. Bauer, D. J. Rezende, M. C. Mozer, Y. Bengio, and C. Pal · 2021
Later among the works it cites.
Causalcity: Complex simulations with agency for causal discovery and reasoning
D. McDuff, Y. Song, J. Lee, V. Vineet, S. Vemprala, N. Gyde, H. Salman, S. Ma, K. Sohn, and A. Kapoor · 2021
Later among the works it cites.
Towards causal representation learning
B. Schölkopf, F. Locatello, S. Bauer, N. R. Ke, N. Kalchbrenner, A. Goyal, and Y. Bengio · 2021
Later among the works it cites.
Causal curiosity: Rl agents discovering self-supervised experiments for causal representation learning
S. A. Sontakke, A. Mehrjou, L. Itti, and B. Schölkopf · 2021
Later among the works it cites.
Acre: Abstract causal reasoning beyond covariation
C. Zhang, B. Jia, M. Edmonds, S.-C. Zhu, and Y. Zhu · 2021
Later among the works it cites.
Palm: Scaling language modeling with pathways
A. Chowdhery, S. Narang, J. Devlin, M. Bosma, G. Mishra, A. Roberts, P. Barham, H. W. Chung, C. Sutton, S. Gehrmann, et al · 2022
Closest in time.
Training compute-optimal large language models
J. Hoffmann, S. Borgeaud, A. Mensch, E. Buchatskaya, T. Cai, E. Rutherford, D. de Las Casas, L. A. Hendricks, J. Welbl, A. Clark, T. Hennigan, E. Noland, K. Millican, G. van den Driessche, B. Damoc, A. Guy, S. Osindero, K. Simonyan, E. Elsen, J. W. Rae, O. Vinyals, and L. Sifre · 2022
Closest in time.
Learning causal overhypotheses through exploration in children and computational models
E. Kosoy, A. Liu, J. Collins, D. M. Chan, J. B. Hamrick, N. R. Ke, S. H. Huang, B. Kaufmann, J. Canny, and A. Gopnik · 2022
Closest in time.
Teaching models to express their uncertainty in words
S. Lin, J. Hilton, and O. Evans · 2022
Closest in time.
S. Smith, M. Patwary, B. Norick, P. LeGresley, S. Rajbhandari, J. Casper, Z. Liu, S. Prabhumoye, G. Zerveas, V. Korthikanti, E. Zhang, R. Child, R. Y. Aminabadi, J. Bernauer, X. Song, M. Shoeybi, Y. He, M. Houston, S. Tiwary, and B. Catanzaro · 2022
Closest in time.