Fetching the paper…
Reading the bibliography…
Many real-life scenarios require humans to make difficult trade-offs: do we always follow all the traffic rules or do we violate the speed limit in an emergency? These scenarios force us to evaluate the trade-off between collective rules and norms with our own personal objectives and desires.
Constrained Markov Decision Processes
Eitan Altman · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y. Ng and Stuart J. Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
Handbook of Constraint Programming
Francesca Rossi, Peter Van Beek, and Toby Walsh · 2006
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Brian D. Ziebart, Andrew L. Maas, J. Andrew Bagnell, and Anind K. Dey · 2008
Earlier work this paper cites.
Modeling interaction via the principle of maximum causal entropy
Brian D. Ziebart, J. Andrew Bagnell, and Anind K. Dey · 2010
Earlier work this paper cites.
Research priorities for robust and beneficial artificial intelligence
S. Russell, D. Dewey, and M. Tegmark · 2015
Earlier work this paper cites.
Reinforcement learning as a framework for ethical decision making
David Abel, James MacGlashan, and Michael L Littman · 2016
Earlier work this paper cites.
Concrete problems in AI safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané · 2016
Cited alongside, same era.
Using contextual bandits with behavioral constraints for constrained online movie recommendation
A. Balakrishnan, D. Bouneffouf, N. Mattei, and F. Rossi · 2018
Cited alongside, same era.
Learning constraints from demonstrations
Glen Chou, Dmitry Berenson, and Necmiye Ozay · 2018
Cited alongside, same era.
Preferences and ethical principles in decision making
Andrea Loreggia, Nicholas Mattei, Francesca Rossi, and K. Brent Venable · 2018
Cited alongside, same era.
Reinforcement Learning: An Introduction, 2nd Edition
Richard S. Sutton and Andrew G. Barto · 2018
Cited alongside, same era.
Teaching AI agents ethical values using reinforcement learning and policy orchestration
Ritesh Noothigattu, Djallel Bouneffouf, Nicholas Mattei, Rachita Chandra, Piyush Madan, Kush R. Varshney, Murray Campbell, Moninder Singh, and Francesca Rossi · 2019
Later among the works it cites.
Benchmarking safe exploration in deep reinforcement learning
Alex Ray, Joshua Achiam, and Dario Amodei · 2019
Later among the works it cites.
Building ethically bounded AI
Francesca Rossi and Nicholas Mattei · 2019
Later among the works it cites.
Maximum likelihood constraint inference for inverse reinforcement learning
Dexter R. R. Scobee and S. Shankar Sastry · 2020
Later among the works it cites.
Inverse constrained reinforcement learning
Shehryar Malik, Usman Anwar, Alireza Aghasi, and Ali Ahmed · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning task specifications from demonstrations
Marcell Vazquez-Chanlatte, Susmit Jha, Ashish Tiwari, Mark K. Ho, and Sanjit A. Seshia · 2018
Cited alongside, same era.
A low-cost ethics shaping approach for designing reinforcement learning agents
Yueh-Hua Wu and Shou-De Lin · 2018
Cited alongside, same era.
Incorporating behavioral constraints in online ai systems
Avinash Balakrishnan, Djallel Bouneffouf, Nicholas Mattei, and Francesca Rossi · 2019
Cited alongside, same era.
Manel Rodriguez-Soto, Maite Lopez-Sanchez, and Juan A Rodriguez-Aguilar · 2021
Later among the works it cites.
Ethically compliant sequential decision making
Justin Svegliato, Samer B Nashed, and Shlomo Zilberstein · 2021
Later among the works it cites.