Fetching the paper…
Reading the bibliography…
Many real-life scenarios require humans to make difficult trade-offs: do we always follow all the traffic rules or do we violate the speed limit in an emergency? These scenarios force us to evaluate the trade-off between collective norms and our own personal objectives.
Benchmarking safe exploration in deep reinforcement learning
Ray, A.; Achiam, J.; and Amodei, D. 2019 · 1910
Earlier work this paper cites.
Decision field theory: a dynamic-cognitive approach to decision making in an uncertain environment
Busemeyer, J. R.; and Townsend, J. T. 1993 · 1993
Earlier work this paper cites.
Constrained Markov Decision Processes , volume 7
Altman, E. 1999 · 1999
Earlier work this paper cites.
Algorithms for Inverse Reinforcement Learning
Ng, A. Y.; and Russell, S. J. 2000 · 2000
Earlier work this paper cites.
Multialternative decision field theory: A dynamic connectionst model of decision making
Roe, R. M.; Busemeyer, J. R.; and Townsend, J. T. 2001 · 2001
Earlier work this paper cites.
Survey of decision field theory
Busemeyer, J. R.; and Diederich, A. 2002 · 2002
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P.; and Ng, A. Y. 2004 · 2004
Earlier work this paper cites.
Handbook of Constraint Programming
Rossi, F.; Van Beek, P.; and Walsh, T. 2006 · 2006
Earlier work this paper cites.
Maximum Entropy Inverse Reinforcement Learning
Ziebart, B. D.; Maas, A. L.; Bagnell, J. A.; and Dey, A. K. 2008 · 2008
Earlier work this paper cites.
Theoretical developments in decision field theory: Comment on Tsetsos, Usher, and Chater (2010)
Hotaling, J. M.; Busemeyer, J. R.; and Li, J. 2010 · 2010
Cited alongside, same era.
Modeling Interaction via the Principle of Maximum Causal Entropy
Ziebart, B. D.; Bagnell, J. A.; and Dey, A. K. 2010 · 2010
Cited alongside, same era.
Thinking, Fast and Slow
Kahneman, D. 2011 · 2011
Cited alongside, same era.
Research priorities for robust and beneficial artificial intelligence
Russell, S.; Dewey, D.; and Tegmark, M. 2015 · 2015
Cited alongside, same era.
Concrete problems in AI safety
Amodei, D.; Olah, C.; Steinhardt, J.; Christiano, P.; Schulman, J.; and Mané, D. 2016 · 2016
Cited alongside, same era.
Using Contextual Bandits with Behavioral Constraints for Constrained Online Movie Recommendation
Incorporating Behavioral Constraints in Online AI Systems
Balakrishnan, A.; Bouneffouf, D.; Mattei, N.; and Rossi, F. 2019 · 2019
Later among the works it cites.
Teaching AI agents ethical values using reinforcement learning and policy orchestration
Noothigattu, R.; Bouneffouf, D.; Mattei, N.; Chandra, R.; Madan, P.; Varshney, K. R.; Campbell, M.; Singh, M.; and Rossi, F. 2019 · 2019
Later among the works it cites.
Learning Preferences in a Cognitive Decision Model
Rahgooy, T.; and Venable, K. B. 2019 · 2019
Later among the works it cites.
Building Ethically Bounded AI
Rossi, F.; and Mattei, N. 2019 · 2019
Later among the works it cites.
Maximum Likelihood Constraint Inference for Inverse Reinforcement Learning
Scobee, D. R. R.; and Sastry, S. S. 2020 · 2020
Later among the works it cites.
Thinking Fast and Slow in AI
Booch, G.; Fabiano, F.; Horesh, L.; Kate, K.; Lenchner, J.; Linck, N.; Loreggia, A.; Murugesan, K.; Mattei, N.; Rossi, F.; and Srivastava, B. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Balakrishnan, A.; Bouneffouf, D.; Mattei, N.; and Rossi, F. 2018 · 2018
Cited alongside, same era.
Learning constraints from demonstrations
Chou, G.; Berenson, D.; and Ozay, N. 2018 · 2018
Cited alongside, same era.
Reinforcement Learning: An Introduction, 2nd Edition
Sutton, R. S.; and Barto, A. G. 2018 · 2018
Cited alongside, same era.
Learning Task Specifications from Demonstrations
Vazquez-Chanlatte, M.; Jha, S.; Tiwari, A.; Ho, M. K.; and Seshia, S. A. 2018 · 2018
Cited alongside, same era.
Closest in time.
Inverse Constrained Reinforcement Learning
Malik, S.; Anwar, U.; Aghasi, A.; and Ahmed, A. 2021 · 2021
Closest in time.
Ethically compliant sequential decision making
Svegliato, J.; Nashed, S. B.; and Zilberstein, S. 2021 · 2021
Closest in time.