TorchBeast: A PyTorch platform for distributed RL
Original
H. Küttler, N. Nardelli, T. Lavril, M. Selvatici, V. Sivakumar, T. Rocktäschel, and E. Grefenstette · 2019
Later among the works it cites.
Language is power: Representing states using natural language in reinforcement learning
Original
E. Schwartz, G. Tennenholtz, C. Tessler, and S. Mannor · 2019
Later among the works it cites.
Learning to speak and act in a fantasy text adventure game
J. Urbanek, A. Fan, S. Karamcheti, S. Jain, S. Humeau, E. Dinan, T. Rocktäschel, D. Kiela, A. Szlam, and J. Weston · 2019
Later among the works it cites.
A narration-based reward shaping approach using grounded natural language commands
Original
N. Waytowich, S. L. Barton, V. Lawhern, and G. Warnell · 2019
Later among the works it cites.
Pixl2r: Guiding reinforcement learning using natural language by mapping pixels to rewards
P. Goyal, S. Niekum, and R. J. Mooney · 2020
Later among the works it cites.
The NetHack learning environment
H. Küttler, N. Nardelli, A. H. Miller, R. Raileanu, M. Selvatici, E. Grefenstette, and T. Rocktäschel · 2020
Later among the works it cites.
Count-based exploration with the successor representation
M. C. Machado, M. G. Bellemare, and M. Bowling · 2020
Later among the works it cites.
Automated curricula through setter-solver interactions
S. Racaniere, A. K. Lampinen, A. Santoro, D. P. Reichert, V. Firoiu, and T. P. Lillicrap · 2020
Later among the works it cites.
RIDE: Rewarding impact-driven exploration for procedurally-generated environments
R. Raileanu and T. Rocktäschel · 2020
Later among the works it cites.
Deep reinforcement learning at the edge of the statistical precipice
R. Agarwal, M. Schwarzer, P. S. Castro, A. C. Courville, and M. Bellemare · 2021
Later among the works it cites.
Learning with AMIGo: Adversarially motivated intrinsic goals
A. Campero, R. Raileanu, H. Kuttler, J. B. Tenenbaum, T. Rocktäschel, and E. Grefenstette · 2021
Later among the works it cites.
ELLA: Exploration through learned language abstraction
S. Mirchandani, S. Karamcheti, and D. Sadigh · 2021
Later among the works it cites.
MiniHack the planet: A sandbox for open-ended reinforcement learning research
M. Samvelyan, R. Kirk, V. Kurin, J. Parker-Holder, M. Jiang, E. Hambro, F. Petroni, H. Kuttler, E. Grefenstette, and T. Rocktäschel · 2021
Later among the works it cites.
ALFWorld: Aligning text and embodied environments for interactive learning
M. Shridhar, X. Yuan, M.-A. Côté, Y. Bisk, A. Trischler, and M. Hausknecht · 2021
Later among the works it cites.
Influencing reinforcement learning through natural language guidance
Original
T. Tasrin, M. S. A. Nahian, H. Perera, and B. Harrison · 2021
Later among the works it cites.
NovelD: A simple yet effective exploration criterion
T. Zhang, H. Xu, X. Wang, Y. Wu, K. Keutzer, J. E. Gonzalez, and Y. Tian · 2021
Later among the works it cites.
Semantic exploration from language abstractions and pretrained representations
Original
A. C. Tam, N. C. Rabinowitz, A. K. Lampinen, N. A. Roy, S. C. Chan, D. Strouse, J. X. Wang, A. Banino, and F. Hill · 2022
Closest in time.