Fetching the paper…
Reading the bibliography…
As machine learning systems become more powerful they also become increasingly unpredictable and opaque.
Design of experiments
Fisher, R. A · 1936
Earlier work this paper cites.
Cybernetics or Control and Communication in the Animal and the Machine
Wiener, N · 1948
Earlier work this paper cites.
Theorie der linearen Wechselstromschaltungen , volume 1
Cauer, W · 1954
Earlier work this paper cites.
An introduction to cybernetics
Ashby, W. R · 1961
Earlier work this paper cites.
Mazes, maps, and memory
Olton, D. S · 1979
Earlier work this paper cites.
The utility driven dynamic error propagation network
Robinson, A. and Fallside, F · 1987
Earlier work this paper cites.
Generalization of backpropagation with application to a recurrent gas market model
Werbos, P. J · 1988
Earlier work this paper cites.
Nature’s capacities and their measurement
Cartwright, N. et al · 1994
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Causal inference without counterfactuals
Dawid, A. P · 2000
Earlier work this paper cites.
Causation, prediction, and search
Spirtes, P., Glymour, C. N., Scheines, R., and Heckerman, D · 2000
Earlier work this paper cites.
Reinforcement learning with long short-term memory
Bakker, B · 2001
Earlier work this paper cites.
Structure and strength in causal induction
Griffiths, T. L. and Tenenbaum, J. B · 2005
Earlier work this paper cites.
Causality
Pearl, J · 2009
Earlier work this paper cites.
Actual causation and the art of modeling
Halpern, J. Y. and Hitchcock, C · 2011
Earlier work this paper cites.
A philosophical treatise of universal induction
Rathmanner, S. and Hutter, M · 2011
Cited alongside, same era.
Information theory of decisions and actions
Tishby, N. and Polani, D · 2011
Cited alongside, same era.
Counterfactual graphical models for longitudinal mediation analysis with unobserved confounding
Shpitser, I · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Cited alongside, same era.
Markov decision processes: discrete stochastic dynamic programming
Puterman, M. L · 2014
Cited alongside, same era.
Bandits with unobserved confounders: A causal approach
Bareinboim, E., Forney, A., and Pearl, J · 2015
Cited alongside, same era.
Methods for interpreting and understanding deep neural networks
Montavon, G., Samek, W., and Müller, K.-R · 2018
Later among the works it cites.
The book of why: the new science of cause and effect
Pearl, J. and Mackenzie, D · 2018
Later among the works it cites.
Rabinowitz, N. C., Perbet, F., Song, H. F., Zhang, C., Eslami, S., and Botvinick, M · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Sutton, R. S. and Barto, A. G · 2018
Later among the works it cites.
Rigorous agent evaluation: An adversarial approach to uncover catastrophic failures
Uesato, J., Kumar, A., Szepesvari, C., Erez, T., Ruderman, A., Anderson, K., Dvijotham, K. D., Heess, N., and Kohli, P · 2018
Later among the works it cites.
Programmatically interpretable reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Statistical causality from a decision-theoretic perspective
Dawid, A. P · 2015
Cited alongside, same era.
Research priorities for robust and beneficial artificial intelligence
Russell, S., Dewey, D., and Tegmark, M · 2015
Cited alongside, same era.
Concrete problems in ai safety
Amodei, D., Olah, C., Steinhardt, J., Christiano, P., Schulman, J., and Mané, D · 2016
Cited alongside, same era.
Learning to learn by gradient descent by gradient descent
Andrychowicz, M., Denil, M., Gomez, S., Hoffman, M. W., Pfau, D., Schaul, T., Shillingford, B., and De Freitas, N · 2016
Cited alongside, same era.
Causal inference in statistics: A primer
Pearl, J., Glymour, M., and Jewell, N. P · 2016
Cited alongside, same era.
Learning to reinforcement learn
Wang, J. X., Kurth-Nelson, Z., Tirumala, D., Soyer, H., Leibo, J. Z., Munos, R., Blundell, C., Kumaran, D., and Botvinick, M · 2016
Cited alongside, same era.
Verma, A., Murali, V., Singh, R., Kohli, P., and Chaudhuri, S · 2018
Later among the works it cites.
Arjovsky, M., Bottou, L., Gulrajani, I., and Lopez-Paz, D · 2019
Later among the works it cites.
Path-specific counterfactual fairness
Chiappa, S · 2019
Later among the works it cites.
Understanding agent incentives using causal influence diagrams, part i: single action settings
Everitt, T., Ortega, P. A., Barnes, E., and Legg, S · 2019
Later among the works it cites.
Towards interpretable reinforcement learning using attention augmented agents
Mott, A., Zoran, D., Chrzanowski, M., Wierstra, D., and Rezende, D. J · 2019
Later among the works it cites.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Later among the works it cites.
The incentives that shape behaviour
Carey, R., Langlois, E., Everitt, T., and Legg, S · 2020
Later among the works it cites.
Explainable reinforcement learning: A survey
Puiutta, E. and Veith, E. M. S. P · 2020
Later among the works it cites.
Resolving spurious correlations in causal models of environments via interventions
Volodin, S., Wichers, N., and Nixon, J · 2020
Later among the works it cites.