Mapping instructions and visual observations to actions with reinforcement learning
Dipendra Misra, John Langford, and Yoav Artzi. 2017 · 2017
Later among the works it cites.
Colors in context: A pragmatic neural model for grounded language understanding
Will Monroe, Robert X.D. Hawkins, Noah D. Goodman, and Christopher Potts. 2017 · 2017
Later among the works it cites.
Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
Peter Anderson, Qi Wu, Damien Teney, Jake Bruce, Mark Johnson, Niko Sünderhauf, Ian D. Reid, Stephen Gould, and Anton van den Hengel. 2018 · 2018
Later among the works it cites.
Pyro: Deep universal probabilistic programming
Original
Eli Bingham, Jonathan P. Chen, Martin Jankowiak, Fritz Obermeyer, Neeraj Pradhan, Theofanis Karaletsos, Rohit Singh, Paul A. Szerlip, Paul Horsfall, and Noah D. Goodman. 2018 · 2018
Later among the works it cites.
Speaker-follower models for vision-and-language navigation
Daniel Fried, Ronghang Hu, Volkan Cirik, Anna Rohrbach, Jacob Andreas, Louis-Philippe Morency, Taylor Berg-Kirkpatrick, Kate Saenko, Dan Klein, and Trevor Darrell. 2018b · 2018
Later among the works it cites.
Learning to map context-dependent sentences to executable formal queries
Alane Suhr, Srinivasan Iyer, and Yoav Artzi. 2018 · 2018
Later among the works it cites.
Learning to understand goal specifications by modelling reward
Dzmitry Bahdanau, Felix Hill, Jan Leike, Edward Hughes, Seyed Arian Hosseini, Pushmeet Kohli, and Edward Grefenstette. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
From language to goals: Inverse reinforcement learning for vision-based instruction following
Justin Fu, Anoop Korattikara, Sergey Levine, and Sergio Guadarrama. 2019 · 2019
Later among the works it cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Later among the works it cites.
Learning from omission
Bill McDowell and Noah Goodman. 2019 · 2019
Later among the works it cites.
Reward-rational (implicit) choice: A unifying formalism for reward learning
Hong Jun Jeon, Smitha Milli, and Anca D. Dragan. 2020 · 2020
Later among the works it cites.
ALFRED: A benchmark for interpreting grounded instructions for everyday tasks
Mohit Shridhar, Jesse Thomason, Daniel Gordon, Yonatan Bisk, Winson Han, Roozbeh Mottaghi, Luke Zettlemoyer, and Dieter Fox. 2020 · 2020
Later among the works it cites.
Extending rational models of communication from beliefs to actions
Theodore R Sumers, Robert D Hawkins, Mark K Ho, and Thomas L Griffiths. 2021 · 2021
Later among the works it cites.
Situated mapping of sequential instructions to actions with single-step reward observation
Alane Suhr and Yoav Artzi. 2018 · 2082
Closest in time.