Fetching the paper…
Reading the bibliography…
We present VisualHints, a novel environment for multimodal reinforcement learning (RL) involving text-based interactions along with visual hints (obtained from the environment).
Yin, X.; and May, J. 2019 · 1908
Earlier work this paper cites.
Ledeepchef: Deep reinforcement learning agent for families of text-based games
Adolphs, L.; and Hofmann, T. 2019 · 1909
Earlier work this paper cites.
ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks
Shridhar, M.; Thomason, J.; Gordon, D.; Bisk, Y.; Han, W.; Mottaghi, R.; Zettlemoyer, L.; and Fox, D. 2020 · 1912
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Mnih, V.; Badia, A. P.; Mirza, M.; Graves, A.; Lillicrap, T.; Harley, T.; Silver, D.; and Kavukcuoglu, K. 2016 · 1937
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
Sutton, R. S.; and Barto, A. G. 1998 · 1998
Earlier work this paper cites.
Graph constrained reinforcement learning for natural language action spaces
Ammanabrolu, P.; and Hausknecht, M. 2020 · 2001
Earlier work this paper cites.
Learning dynamic knowledge graphs to generalize on text-based games
Adhikari, A.; Yuan, X.; Côté, M.-A.; Zelinka, M.; Rondeau, M.-A.; Laroche, R.; Poupart, P.; Tang, J.; Trischler, A.; and Hamilton, W. L. 2020 · 2002
Earlier work this paper cites.
Probabilistic planning vs. replanning
Little, I.; and Thiebaux, S. 2007 · 2007
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, E.; Erez, T.; and Tassa, Y. 2012 · 2012
Earlier work this paper cites.
The Arcade Learning Environment: An Evaluation Platform for General Agents
Bellemare, M. G.; Naddaf, Y.; Veness, J.; and Bowling, M. 2013 · 2013
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; Petersen, S.; Beattie, C.; Sadik, A.; Antonoglou, I.; King, H.; Kumaran, D.; Wierstra, D.; Legg, S.; and Hassabis, D. 2015 · 2015
Cited alongside, same era.
Language Understanding for Text-based Games using Deep Reinforcement Learning
Narasimhan, K.; Kulkarni, T.; and Barzilay, R. 2015 · 2015
Cited alongside, same era.
Trust region policy optimization
Schulman, J.; Levine, S.; Abbeel, P.; Jordan, M.; and Moritz, P. 2015 · 2015
Cited alongside, same era.
OpenAI Gym
Brockman, G.; Cheung, V.; Pettersson, L.; Schneider, J.; Schulman, J.; Tang, J.; and Zaremba, W. 2016 · 2016
Cited alongside, same era.
Deep Reinforcement Learning with a Natural Language Action Space
Ai2-thor: An interactive 3d environment for visual ai
Kolve, E.; Mottaghi, R.; Han, W.; VanderBilt, E.; Weihs, L.; Herrasti, A.; Gordon, D.; Zhu, Y.; Gupta, A.; and Farhadi, A. 2017 · 2017
Later among the works it cites.
A deep reinforcement learning chatbot
Serban, I. V.; Sankar, C.; Germain, M.; Zhang, S.; Lin, Z.; Subramanian, S.; Kim, T.; Pieper, M.; Chandar, S.; Ke, N. R.; et al. 2017 · 2017
Later among the works it cites.
Playing text-adventure games with graph-based deep reinforcement learning
Ammanabrolu, P.; and Riedl, M. O. 2018 · 2018
Later among the works it cites.
Textworld: A learning environment for text-based games
Côté, M.-A.; Kádár, Á.; Yuan, X.; Kybartas, B.; Barnes, T.; Fine, E.; Moore, J.; Hausknecht, M.; Asri, L. E.; Adada, M.; et al. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
He, J.; Chen, J.; He, X.; Gao, J.; Li, L.; Deng, L.; and Ostendorf, M. 2016 · 2016
Cited alongside, same era.
Towards End-to-End Reinforcement Learning of Dialogue Agents for Information Access
Dhingra, B.; Li, L.; Li, X.; Gao, J.; Chen, Y.-N.; Ahmed, F.; and Deng, L. 2017 · 2017
Cited alongside, same era.
What can you do with a rock? affordance extraction via word embeddings
Fulda, N.; Ricks, D.; Murdoch, B.; and Wingate, D. 2017 · 2017
Cited alongside, same era.
Yuan, X.; Côté, M.-A.; Sordoni, A.; Laroche, R.; Combes, R. T. d.; Hausknecht, M.; and Trischler, A. 2018 · 2018
Later among the works it cites.
Learn what not to learn: Action elimination with deep reinforcement learning
Zahavy, T.; Haroush, M.; Merlis, N.; Mankowitz, D. J.; and Mannor, S. 2018 · 2018
Later among the works it cites.
Habitat: A platform for embodied ai research
Savva, M.; Kadian, A.; Maksymets, O.; Zhao, Y.; Wijmans, E.; Jain, B.; Straub, J.; Liu, J.; Koltun, V.; Malik, J.; et al. 2019 · 2019
Later among the works it cites.
AllenAct: A Framework for Embodied AI Research
Weihs, L.; Salvador, J.; Kotar, K.; Jain, U.; Zeng, K.-H.; Mottaghi, R.; and Kembhavi, A. 2020 · 2020
Closest in time.