Fetching the paper…
Reading the bibliography…
We explore using latent natural language instructions as an expressive and compositional representation of complex actions for hierarchical decision making.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Sutton Richard, Precup Doina, and Singh Satinder · 1999
Earlier work this paper cites.
Online learning of relaxed ccg grammars for parsing to logical form
Luke Zettlemoyer and Michael Collins · 2007
Earlier work this paper cites.
Build order optimization in starcraft
David Churchill and Michael Buro · 2011
Earlier work this paper cites.
Fast heuristic search for rts game combat scenarios
David Churchill, Abdallah Saffidine, and Michael Buro · 2012
Earlier work this paper cites.
Weakly supervised learning of semantic parsers for mapping instructions to actions
Yoav Artzi and Luke Zettlemoyer · 2013
Earlier work this paper cites.
Implementing a wall-in building placement in starcraft with declarative programming
Michal Certicky · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
A survey of real-time strategy game ai research and competition in starcraft
Santiago Ontanón, Gabriel Synnaeve, Alberto Uriarte, Florian Richoux, David Churchill, and Mike Preuss · 2013
Earlier work this paper cites.
Mazebase: A sandbox for learning from games
Sainbayar Sukhbaatar, Arthur Szlam, Gabriel Synnaeve, Soumith Chintala, and Rob Fergus · 2015
Earlier work this paper cites.
The option-critic architecture
Pierre-Luc Bacon, Jean Harb, and Doina Precup · 2016
Earlier work this paper cites.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
Hongyuan Mei, Mohit Bansal, and Matthew R Walter · 2016
Cited alongside, same era.
Torchcraft: a library for machine learning research on real-time strategy games
Gabriel Synnaeve, Nantas Nardelli, Alex Auvolat, Soumith Chintala, Timothée Lacroix, Zeming Lin, Florian Richoux, and Nicolas Usunier · 2016
Cited alongside, same era.
A deep hierarchical approach to lifelong learning in minecraft
Chen Tessler, Shahar Givony, Tom Zahavy, Daniel J. Mankowitz, and Shie Mannor · 2016
Cited alongside, same era.
Modular multitask reinforcement learning with policy sketches
Jacob Andreas, Dan Klein, and Sergey Levine · 2017
Cited alongside, same era.
Navigational instruction generation as inverse reinforcement learning with neural machine translation
Andrea F Daniele, Mohit Bansal, and Matthew R Walter · 2017
Elf: An extensive, lightweight and flexible research platform for real-time strategy games
Yuandong Tian, Qucheng Gong, Wenling Shang, Yuxin Wu, and C Lawrence Zitnick · 2017
Later among the works it cites.
Episodic exploration for deep deterministic policies: An application to starcraft micromanagement tasks
Nicolas Usunier, Gabriel Synnaeve, Zeming Lin, and Soumith Chintala · 2017
Later among the works it cites.
Starcraft II: A new challenge for reinforcement learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, John Quan, Stephen Gaffney, Stig Petersen, Karen Simonyan, Tom Schaul, Hado van Hasselt, David Silver, Timothy P. Lillicrap, Kevin Calderone, Paul Keet, Anthony Brunasso, David Lawrence, Anders Ekermo, Jacob Repp, and Rodney Tsing · 2017
Later among the works it cites.
Speaker-follower models for vision-and-language navigation
Daniel Fried, Ronghang Hu, Volkan Cirik, Anna Rohrbach, Jacob Andreas, Louis-Philippe Morency, Taylor Berg-Kirkpatrick, Kate Saenko, Dan Klein, and Trevor Darrell · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Unified pragmatic models for generating and following instructions
Daniel Fried, Jacob Andreas, and Dan Klein · 2017
Cited alongside, same era.
Grounded language learning in a simulated 3d world
Karl Moritz Hermann, Felix Hill, Simon Green, Fumin Wang, Ryan Faulkner, Hubert Soyer, David Szepesvari, Wojciech Marian Czarnecki, Max Jaderberg, Denis Teplyashin, et al · 2017
Cited alongside, same era.
Parlai: A dialog research software platform
A. H. Miller, W. Feng, A. Fisch, J. Lu, D. Batra, A. Bordes, D. Parikh, and J. Weston · 2017
Cited alongside, same era.
Zero-shot task generalization with multi-task deep reinforcement learning
Junhyuk Oh, Satinder Singh, Honglak Lee, and Pushmeet Kohli · 2017
Cited alongside, same era.
Multiagent bidirectionally-coordinated nets for learning to play starcraft combat games
Peng Peng, Quan Yuan, Ying Wen, Yaodong Yang, Zhenkun Tang, Haitao Long, and Jun Wang · 2017
Cited alongside, same era.
Openai five
OpenAI · 2018
Later among the works it cites.
Tstarbots: Defeating the cheating level builtin ai in starcraft ii in the full game
Peng Sun, Xinghai Sun, Lei Han, Jiechao Xiong, Qing Wang, Bo Li, Yang Zheng, Ji Liu, Yongsheng Liu, Han Liu, et al · 2018
Later among the works it cites.
Hierarchical macro strategy model for moba game ai
Bin Wu, Qiang Fu, Jing Liang, Peng Qu, Xiaoqian Li, Liang Wang, Wei Liu, Wei Yang, and Yongsheng Liu · 2018
Later among the works it cites.
Hierarchical text generation and planning for strategic dialogue
Denis Yarats and Mike Lewis · 2018
Later among the works it cites.
Relational deep reinforcement learning
Vinicius Zambaldi, David Raposo, Adam Santoro, Victor Bapst, Yujia Li, Igor Babuschkin, Karl Tuyls, David Reichert, Timothy Lillicrap, Edward Lockhart, et al · 2018
Later among the works it cites.