Fetching the paper…
Reading the bibliography…
An ambitious goal for machine learning is to create agents that behave ethically: The capacity to abide by human moral norms would greatly expand the context in which autonomous agents could be practically and safely deployed, e.g.
A difficulty in the concept of social welfare
Arrow, K. J · 1950
Earlier work this paper cites.
Non-cooperative games
Nash, J · 1951
Earlier work this paper cites.
Groundwork of the metaphysic of morals. Translated and analysed by HJ Paton
Kant, I. and Paton, H. J · 1964
Earlier work this paper cites.
The problem of abortion and the doctrine of the double effect
Foot, P · 1967
Earlier work this paper cites.
Manipulation of schemes that mix voting with chance
Gibbard, A · 1977
Earlier work this paper cites.
The social rate of discount for nuclear waste storage: economics or ethics?
Schulze, W. D., Brookshire, D. S., and Sandler, T · 1981
Earlier work this paper cites.
Ethics, irreversibility, future generations and the social rate of discount
Pearce, D · 1983
Earlier work this paper cites.
Computational philosophy of science
Thagard, P · 1993
Earlier work this paper cites.
Average reward reinforcement learning: Foundations, algorithms, and empirical results
Mahadevan, S · 1996
Earlier work this paper cites.
Multi-criteria reinforcement learning
Gábor, Z., Kalmár, Z., and Szepesvári, C · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
Sutton, R. S. and Barto, A. G · 1998
Earlier work this paper cites.
Moral uncertainty and its consequences
Lockhart, T · 2000
Earlier work this paper cites.
Artificial morality: Virtuous robots for virtual games
Danielson, P · 2002
Earlier work this paper cites.
Q-decomposition for reinforcement learning agents
Russell, S. J. and Zimdars, A · 2003
Earlier work this paper cites.
Towards machine ethics: Implementing two action-based ethical theories
Anderson, M., Anderson, S., and Armen, C · 2005
Earlier work this paper cites.
Emergence of cooperation: State of the art
Nitschke, G · 2005
Earlier work this paper cites.
Why machine ethics?
Allen, C., Wallach, W., and Smit, I · 2006
Earlier work this paper cites.
Computers as prostheses for the imagination
Dennett, D · 2006
Earlier work this paper cites.
Ethics of the discount rate in the stern review on the economics of climate change
Beckerman, W., Hepburn, C., et al · 2007
Earlier work this paper cites.
A comprehensive survey of multiagent reinforcement learning
Bu, L., Babu, R., De Schutter, B., et al · 2008
Earlier work this paper cites.
Moral machines: Teaching robots right from wrong
Wallach, W. and Allen, C · 2008
Earlier work this paper cites.
Moral uncertainty—towards a solution?
Bostrom, N · 2009
Cited alongside, same era.
Probabilistic electoral methods, representative probability, and maximum entropy
Sewell, R., MacKay, D., and McLean, I · 2009
Cited alongside, same era.
Rectified linear units improve restricted boltzmann machines
Nair, V. and Hinton, G. E · 2010
Cited alongside, same era.
Machine metaethics
Anderson, S. L · 2011
Cited alongside, same era.
Computational meta-ethics
Lokhorst, G.-J. C · 2011
Cited alongside, same era.
On what matters , volume 1
Parfit, D · 2011
Cited alongside, same era.
Geometric reasons for normalising variance to aggregate preferences, 2013
Stoic ethics for artificial agents
Murray, G · 2017
Later among the works it cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L. R., Lai, M., Bolton, A., Chen, Y., Lillicrap, T. P., Hui, F., Sifre, L., van den Driessche, G., Graepel, T., and Hassabis, D · 2017
Later among the works it cites.
A survey of preference-based reinforcement learning methods
Wirth, C., Akrour, R., Neumann, G., and Fürnkranz, J · 2017
Later among the works it cites.
Everitt, T., Lea, G., and Hutter, M · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cotton-Barratt, O · 2013
Cited alongside, same era.
A survey of multi-objective sequential decision-making
Roijers, D. M., Vamplew, P., Whiteson, S., and Dazeley, R · 2013
Cited alongside, same era.
Moral uncertainty and the principle of equity among moral theories
Sepielli, A · 2013
Cited alongside, same era.
Normative uncertainty
MacAskill, W · 2014
Cited alongside, same era.
Towards an ethical robot: internal models, consequences and ethical action selection
Winfield, A. F., Blum, C., and Liu, W · 2014
Cited alongside, same era.
Deep ordinal reinforcement learning
Zap, A., Joppen, T., and Fürnkranz, J · 2014
Cited alongside, same era.
Toward non-intuition-based machine and artificial intelligence ethics: A deontological approach based on modal logic
Hooker, J. N. and Kim, T. W. N · 2018
Later among the works it cites.
Measuring and avoiding side effects using relative reachability
Krakovna, V., Orseau, L., Martic, M., and Legg, S · 2018
Later among the works it cites.
Quadratic voting: How mechanism design can radicalize democracy
Lalley, S. P. and Weyl, E. G · 2018
Later among the works it cites.
Trial without error: Towards safe reinforcement learning via human intervention
Saunders, W., Sastry, G., Stuhlmüller, A., and Evans, O · 2018
Later among the works it cites.
Human-aligned artificial intelligence is a multiobjective problem
Vamplew, P., Dazeley, R., Foale, C., Firmin, S., and Mummery, J · 2018
Later among the works it cites.
Hyperbolic discounting and learning over multiple horizons
Fedus, W., Gelada, C., Bengio, Y., Bellemare, M. G., and Larochelle, H · 2019
Later among the works it cites.
Doctrine of double effect
McIntyre, A · 2019
Later among the works it cites.
Voting methods
Pacuit, E · 2019
Later among the works it cites.
Learning human objectives by evaluating hypothetical behavior
Reddy, S., Dragan, A. D., Levine, S., Legg, S., and Leike, J · 2019
Later among the works it cites.
Game theory
Ross, D · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al · 2019
Later among the works it cites.
Artificial intelligence, values and alignment
Gabriel, I · 2020
Closest in time.
Computational philosophy
Grim, P. and Singer, D · 2020
Closest in time.
An axiomatic approach to axiological uncertainty
Riedener, S · 2020
Closest in time.
Conservative agency via attainable utility preservation
Turner, A. M., Hadfield-Menell, D., and Tadepalli, P · 2020
Closest in time.