Fetching the paper…
Reading the bibliography…
Corrigibility is a safety property for artificially intelligent agents.
John Von Neumann and Oskar Morgenstern, Theory of games and economic behavior , Princeton University Press, 1944
1944
Earlier work this paper cites.
Marcus Hutter, Universal algorithmic intelligence: A mathematical top→ down approach , Artificial general intelligence, Springer, 2007, pp. 227–290
2007
Earlier work this paper cites.
Stephen M Omohundro, The basic AI drives , AGI, vol. 171, 2008, pp. 483–492
2008
Earlier work this paper cites.
Nick Bostrom, Superintelligence: paths, dangers, strategies , 2014
2014
Earlier work this paper cites.
Stuart Armstrong, Motivated value selection for artificial agents , Workshops at the Twenty-Ninth AAAI Conference on Artificial Intelligence, 2015
2015
Earlier work this paper cites.
Nate Soares, Benja Fallenstein, Stuart Armstrong, and Eliezer Yudkowsky, Corrigibility , Workshops at the Twenty-Ninth AAAI Conference on Artificial Intelligence, 2015
2015
Earlier work this paper cites.
Tom Everitt, Daniel Filan, Mayank Daswani, and Marcus Hutter, Self-modification of policy and utility function in rational agents , International Conference on Artificial General Intelligence, Springer, 2016, pp. 1–11
2016
Cited alongside, same era.
Laurent Orseau and Stuart Armstrong, Safely interruptible agents , Proceedings of the Thirty-Second Conference on Uncertainty in Artificial Intelligence, AUAI Press, 2016, pp. 557–566
2016
Cited alongside, same era.
Dylan Hadfield-Menell, Anca Dragan, Pieter Abbeel, and Stuart Russell, The off-switch game , Workshops at the Thirty-First AAAI Conference on Artificial Intelligence, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Tom Everitt, Gary Lea, and Marcus Hutter, AGI safety literature review , Proceedings of the 27th International Joint Conference on Artificial Intelligence, AAAI Press, 2018, pp. 5441–5449
2018
Later among the works it cites.
Koen Holtman, AGI agent simulator , 2019, Open Source, Apache Licence 2.0. Available at https://github.com/kholtman/agisim
2019
Closest in time.
Yat Long Lo, Chung Yu Woo, and Ka Lok Ng, The necessary roadblock to artificial general intelligence: Corrigibility , AI Matters 5
2019
Closest in time.
Koen Holtman, AGI agent safety by iteratively improving the utility function , To be published (2020)
2020
Closest in time.
Koen Holtman, AGI agent safety by iteratively improving the utility function: Proofs, models, and reality , To be published. (2020)
2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
Ryan Carey, Incorrigibility in the CIRL framework , Proceedings of the 2018 AAAI/ACM Conference on AI, Ethics, and Society, ACM, 2018, pp. 30–35
2018
Cited alongside, same era.