Fetching the paper…
Reading the bibliography…
For an autonomous system to be helpful to humans and to pose no unwarranted risks, it needs to align its values with those of the humans in its environment in such a way that its actions contribute to the maximization of value for the humans.
Some moral and technical consequences of automation
Wiener, N · 1960
Earlier work this paper cites.
The optimal control of partially observable Markov processes over a finite horizon
Smallwood, R and Sondik, E · 1973
Earlier work this paper cites.
On the folly of rewarding A, while hoping for B
Kerr, S · 1975
Earlier work this paper cites.
Theory of the firm: Managerial behavior, agency costs and ownership structure
Jensen, M and Meckling, W · 1976
Earlier work this paper cites.
Learning binary relations and total orders
Goldman, S, Rivest, R, and Schapire, R · 1993
Earlier work this paper cites.
An investigation of the Therac-25 accidents
Leveson, N and Turner, C · 1993
Earlier work this paper cites.
On the complexity of teaching
Goldman, S and Kearns, M · 1995
Earlier work this paper cites.
Incentives in organizations
Gibbons, R · 1998
Earlier work this paper cites.
Learning agents for uncertain environments (extended abstract)
Russell, Stuart J · 1998
Earlier work this paper cites.
Sequential optimality and coordination in multiagent systems
Boutilier, Craig · 1999
Cited alongside, same era.
The complexity of decentralized control of Markov decision processes
Bernstein, D, Zilberstein, S, and Immerman, N · 2000
Cited alongside, same era.
Algorithms for inverse reinforcement learning
Ng, A and Russell, S · 2000
Cited alongside, same era.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P and Ng, A · 2004
Cited alongside, same era.
Maximum margin planning
Ratliff, N, Bagnell, J, and Zinkevich, M · 2006
Cited alongside, same era.
Bayesian inverse reinforcement learning
Ramachandran, D and Amir, E · 2007
Cited alongside, same era.
Maximum entropy inverse reinforcement learning
Multi-agent inverse reinforcement learning
Natarajan, S, Kunapuli, G, Judah, K, Tadepalli, P, and Kersting, Kand Shavlik, J · 2010
Later among the works it cites.
Artificial Intelligence
Russell, S. and Norvig, P · 2010
Later among the works it cites.
Computational rationalization: The inverse equilibrium problem
Waugh, K, Ziebart, B, and Bagnell, J · 2011
Later among the works it cites.
Algorithmic and human teaching of sequential decision tasks
Cakmak, M and Lopes, M · 2012
Later among the works it cites.
Generating legible motion
Dragan, A and Srinivasa, S · 2013
Later among the works it cites.
Decentralized stochastic control with partial history sharing: A common information approach
Nayyar, A, Mahajan, A, and Teneketzis, D · 2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ziebart, B, Maas, A, Bagnell, J, and Dey, A · 2008
Cited alongside, same era.
Recent developments in algorithmic teaching
Balbach, F and Zeugmann, T · 2009
Cited alongside, same era.
A game-theoretic approach to generating spatial descriptions
Golland, D, Liang, P, and Klein, D · 2010
Cited alongside, same era.
Superintelligence: Paths, dangers, strategies
Bostrom, N · 2014
Later among the works it cites.
A decision-theoretic model of assistance
Fern, A, Natarajan, S, Judah, K, and Tadepalli, P · 2014
Later among the works it cites.
Inverse game theory
Kuleshov, V and Schrijvers, O · 2015
Later among the works it cites.