Fetching the paper…
Reading the bibliography…
When making everyday decisions, people are guided by their conscience, an internal sense of right and wrong.
The Methods of Ethics
Henry Sidgwick · 1907
Earlier work this paper cites.
The Right and the Good
W. D. Ross · 1930
Earlier work this paper cites.
The Limits of Morality
Shelly Kagan · 1991
Earlier work this paper cites.
The structure of normative ethics
Shelly Kagan · 1992
Earlier work this paper cites.
Learning agents for uncertain environments (extended abstract)
S. Russell · 1998
Earlier work this paper cites.
A Theory of Justice
John Rawls · 1999
Earlier work this paper cites.
Morality: its nature and justification
Bernard Gert · 2005
Earlier work this paper cites.
A comprehensive survey on safe reinforcement learning
J. Garcia and F. Fernández · 2015
Earlier work this paper cites.
Cooperative inverse reinforcement learning
Dylan Hadfield-Menell, S. Russell, P. Abbeel, and A. Dragan · 2016
Earlier work this paper cites.
Deep reinforcement learning with a natural language action space
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf · 2016
Earlier work this paper cites.
Keep calm and explore: Language models for action generation in text-based games
Shunyu Yao, Rohan Rao, Matthew Hausknecht, and Karthik Narasimhan · 2016
Earlier work this paper cites.
Constrained policy optimization
Joshua Achiam, David Held, A. Tamar, and P. Abbeel · 2017
Cited alongside, same era.
Utilitarianism: a very short introduction
Katarzyna de. Lazari-Radek and Peter Singer · 2017
Cited alongside, same era.
J. Leike, Miljan Martic, Victoria Krakovna, Pedro A. Ortega, Tom Everitt, Andrew Lefrancq, Laurent Orseau, and S. Legg · 2017
Cited alongside, same era.
Textworld: A learning environment for text-based games
Marc-Alexandre Côté, Ákos Kádár, Xingdi Yuan, Ben Kybartas, Tavian Barnes, Emery Fine, J. Moore, Matthew J. Hausknecht, Layla El Asri, Mahmoud Adada, Wendy Tay, and Adam Trischler · 2018
Cited alongside, same era.
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumeé III, and Kate Crawford · 2018
Cited alongside, same era.
Learning dynamic belief graphs to generalize on text-based games
Ashutosh Adhikari, Xingdi Yuan, Marc-Alexandre Côté, Mikuláš Zelinka, Marc-Antoine Rondeau, Romain Laroche, Pascal Poupart, Jian Tang, Adam Trischler, and William L. Hamilton · 2020
Later among the works it cites.
Graph constrained reinforcement learning for natural language action spaces
Prithviraj Ammanabrolu and Matthew Hausknecht · 2020
Later among the works it cites.
How to avoid being eaten by a grue: Structured exploration strategies for textual worlds
Prithviraj Ammanabrolu, Ethan Tien, Matthew Hausknecht, and Mark O. Riedl · 2020
Later among the works it cites.
Interactive fiction games: A colossal adventure
Matthew Hausknecht, Prithviraj Ammanabrolu, Marc-Alexandre Côté, and Xingdi Yuan · 2020
Later among the works it cites.
Learning human objectives by evaluating hypothetical behavior
Siddharth Reddy, Anca Dragan, Sergey Levine, Shane Legg, and Jan Leike · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Playing text-adventure games with graph-based deep reinforcement learning
Prithviraj Ammanabrolu and Mark Riedl · 2019
Cited alongside, same era.
The text-based adventure ai competition
Timothy Atkinson, H. Baier, Tara Copplestone, S. Devlin, and J. Swan · 2019
Cited alongside, same era.
Nail: A general interactive fiction agent
Matthew J. Hausknecht, R. Loynd, Greg Yang, A. Swaminathan, and J. Williams · 2019
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach
Y. Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, M. Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
Benchmarking safe exploration in deep reinforcement learning
Alex Ray, Joshua Achiam, and Dario Amodei · 2019
Cited alongside, same era.
Safelife 1.0: Exploring side effects in complex environments
Carroll L Wainwright and Peter Eckersley · 2019
Cited alongside, same era.
Nicomachean Ethics
Aristotle
Cited in the paper.
Later among the works it cites.
Avoiding side effects in complex environments
Alex Turner, Neale Ratzlaff, and Prasad Tadepalli · 2020
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Closest in time.
Decision transformer: Reinforcement learning via sequence modeling
Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Michael Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch · 2021
Closest in time.
Reinforcement learning as one big sequence modeling problem
Michael Janner, Qiyang Li, and Sergey Levine · 2021
Closest in time.
Training value-aligned reinforcement learning agents using a normative prior
Md Sultan Al Nahian, Spencer Frazier, Brent Harrison, and Mark Riedl · 2021
Closest in time.
Open-ended learning leads to generally capable agents
Ended Learning Team, Adam Stooke, Anuj Mahajan, Catarina Barros, Charlie Deck, Jakob Bauer, Jakub Sygnowski, Maja Trebacz, Max Jaderberg, Michael Mathieu, et al · 2021
Closest in time.