Fetching the paper…
Reading the bibliography…
Assistance games (also known as cooperative inverse reinforcement learning games) have been proposed as a model for beneficial AI, wherein a robotic agent must act on behalf of a human principal but is initially uncertain about the humans payoff function.
The assistive multi-armed bandit
Chan, L., Hadfield-Menell, D., Srinivasa, S. S., and Dragan, A. D · 1901
Earlier work this paper cites.
Learning to interactively learn and assist
Woodward, M., Finn, C., and Hausman, K · 1906
Earlier work this paper cites.
Manipulation of voting schemes: a general result
Gibbard, A · 1973
Earlier work this paper cites.
Straightforwardness of game forms with lotteries as outcomes
Gibbard, A · 1978
Earlier work this paper cites.
Social choice theory
Sen, A · 1986
Earlier work this paper cites.
Deriving consensus in multiagent systems
Ephrati, E. and Rosenschein, J. S · 1996
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Ng, A. Y., Russell, S. J., et al · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A. Y · 2004
Earlier work this paper cites.
Interactive robots as social partners and peer tutors for children: A field trial
Kanda, T., Hirano, T., Eaton, D., and Ishiguro, H · 2004
Earlier work this paper cites.
Effects of repeated exposure to a humanoid robot on children with autism
Robins, B., Dautenhahn, K., Te Boekhorst, R., and Billard, A · 2004
Cited alongside, same era.
The distortion of cardinal preferences in voting
Procaccia, A. D. and Rosenschein, J. S · 2006
Cited alongside, same era.
Goal inference as inverse planning
Baker, C. L., Tenenbaum, J. B., and Saxe, R. R · 2007
Cited alongside, same era.
Bayesian inverse reinforcement learning
Ramachandran, D. and Amir, E · 2007
Cited alongside, same era.
Max-min fairness and its applications to routing and load-balancing in communication networks: a tutorial
Nace, D. and Pióro, M · 2008
Cited alongside, same era.
Maximum entropy inverse reinforcement learning
Ziebart, B. D., Maas, A. L., Bagnell, J. A., and Dey, A. K · 2008
Cited alongside, same era.
Optimal social choice functions: A utilitarian view
Boutilier, C., Caragiannis, I., Haber, S., Lu, T., Procaccia, A. D., and Sheffet, O · 2015
Later among the works it cites.
Multiobjective reinforcement learning: A comprehensive overview
Liu, C., Xu, X., and Hu, D · 2015
Later among the works it cites.
Concrete problems in ai safety
Amodei, D., Olah, C., Steinhardt, J., Christiano, P., Schulman, J., and Mané, D · 2016
Later among the works it cites.
Cooperative inverse reinforcement learning
Hadfield-Menell, D., Russell, S. J., Abbeel, P., and Dragan, A · 2016
Later among the works it cites.
Showing versus doing: Teaching by demonstration
Ho, M. K., Littman, M., MacGlashan, J., Cushman, F., and Austerweil, J. L · 2016
Later among the works it cites.
Robot planning with mathematical models of human state and action
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Markov decision processes: discrete stochastic dynamic programming
Puterman, M. L · 2014
Cited alongside, same era.
Fairness in multi-agent sequential decision-making
Zhang, C. and Shah, J. A · 2014
Cited alongside, same era.
An introduction to the theory of mechanism design
Börgers, T · 2015
Cited alongside, same era.
Dragan, A. D · 2017
Later among the works it cites.
Inverse reward design
Hadfield-Menell, D., Milli, S., Abbeel, P., Russell, S. J., and Dragan, A · 2017
Later among the works it cites.
Reinforcement learning with fairness constraints for resource distribution in human-robot teams
Claure, H., Chen, Y., Modi, J., Jung, M., and Nikolaidis, S · 2019
Later among the works it cites.