Fetching the paper…
Reading the bibliography…
We introduce TextWorld, a sandbox learning environment for the training and evaluation of RL agents on text-based games.
The problem of the random walk
Karl Pearson · 1905
Earlier work this paper cites.
Three models for the description of language
N. Chomsky · 1956
Earlier work this paper cites.
Strips: A new approach to the application of theorem proving to problem solving
Richard E Fikes and Nils J Nilsson · 1971
Earlier work this paper cites.
Aventure, 1976
William Crowther and Donald Woods · 1976
Earlier work this paper cites.
Zork I, 1980
Infocom · 1980
Earlier work this paper cites.
Zork III, 1982
Infocom · 1982
Earlier work this paper cites.
Infidel, 1983
Michael Berlyn · 1983
Earlier work this paper cites.
The hitchhiker’s guide to the galaxy, 1984
Douglas Adams and Steve Meretzky · 1984
Earlier work this paper cites.
Seastalker, 1984
Stu Galley and Jim Lawrence · 1984
Earlier work this paper cites.
Wishbringer, 1985
Brian Moriarty · 1985
Earlier work this paper cites.
Inhumane, 1985
Andrew Plotkin · 1985
Earlier work this paper cites.
Sherlock: The riddle of the crown jewels, 1987
Bob Bates · 1987
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Martin L. Puterman · 1994
Earlier work this paper cites.
Reverberations, 1996
Russell Glasser · 1996
Earlier work this paper cites.
SimCity
Will Wright · 1996
Earlier work this paper cites.
Anchorhead, 1998
Michel Gentry · 1998
Cited alongside, same era.
Planning and acting in partially observable stochastic domains
Leslie Pack Kaelbling, Michael L Littman, and Anthony R Cassandra · 1998
Cited alongside, same era.
The enterprise incidents, 2002
Brian Desilets · 2002
Cited alongside, same era.
Goldilocks is a fox!, 2002
J.J. Guest · 2002
Cited alongside, same era.
A survey of exploration strategies in reinforcement learning
Roger McFarlane · 2003
Cited alongside, same era.
General game playing: Overview of the aaai competition
Michael Genesereth, Nathaniel Love, and Barney Pell · 2005
Cited alongside, same era.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston · 2009
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Later among the works it cites.
Vizdoom: A doom-based ai research platform for visual reinforcement learning
Michał Kempka, Marek Wydmuch, Grzegorz Runc, Jakub Toczek, and Wojciech Jaśkowski · 2016
Later among the works it cites.
Artificial intelligence: a modern approach
Stuart J Russell and Peter Norvig · 2016
Later among the works it cites.
Expressionist: An Authoring Tool for In-Game Text Generation , pages 221–233
James Ryan, Ethan Seither, Michael Mateas, and Noah Wardrip-Fruin · 2016
Later among the works it cites.
Commai: Evaluating the first steps towards a useful general ai
Marco Baroni, Armand Joulin, Allan Jabri, Germàn Kruszewski, Angeliki Lazaridou, Klemen Simonic, and Tomas Mikolov · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
The second dialog state tracking challenge
Matthew Henderson, Blaise Thomson, and Jason D Williams · 2014
Cited alongside, same era.
Deep reinforcement learning with a natural language action space
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf · 2015
Cited alongside, same era.
Ceptre: A language for modeling generative interactive systems
Chris Martens · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Gated-attention architectures for task-oriented language grounding
Devendra Singh Chaplot, Kanthashree Mysore Sathyendra, Rama Kumar Pasumarthi, Dheeraj Rajagopal, and Ruslan Salakhutdinov · 2017
Later among the works it cites.
What can you do with a rock? affordance extraction via word embeddings
Nancy Fulda, Daniel Ricks, Ben Murdoch, and David Wingate · 2017
Later among the works it cites.
Grounded language learning in a simulated 3d world
Karl Moritz Hermann, Felix Hill, Simon Green, Fumin Wang, Ryan Faulkner, Hubert Soyer, David Szepesvari, Wojtek Czarnecki, Max Jaderberg, Denis Teplyashin, et al · 2017
Later among the works it cites.
Text-based adventures of the golovin ai agent
Bartosz Kostka, Jaroslaw Kwiecieli, Jakub Kowalski, and Pawel Rychlikowski · 2017
Later among the works it cites.
Marlos C Machado, Marc G Bellemare, Erik Talvitie, Joel Veness, Matthew Hausknecht, and Michael Bowling · 2017
Later among the works it cites.
Neural map: Structured memory for deep reinforcement learning
Emilio Parisotto and Ruslan Salakhutdinov · 2017
Later among the works it cites.
The text-based adventure ai competition
Timothy Atkinson, Hendrik Baier, Tara Copplestone, Sam Devlin, and Jerry Swan · 2018
Closest in time.
Learning how not to act in text-based games
Matan Haroush, Tom Zahavy, Daniel J Mankowitz, and Shie Mannor · 2018
Closest in time.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Closest in time.