Unsupervised control through non-parametric discriminative rewards
Original
Warde-Farley, D., Van de Wiele, T., Kulkarni, T., Ionescu, C., Hansen, S., and Mnih, V · 2018
Later among the works it cites.
Actrce: Augmenting experience via teacher’s advice for multi-goal reinforcement learning
Original
Chan, H., Wu, Y., Kiros, J., Fidler, S., and Ba, J · 2019
Later among the works it cites.
Self-educated language agent with hindsight experience replay for instruction following
Original
Cideron, G., Seurin, M., Strub, F., and Pietquin, O · 2019
Later among the works it cites.
CURIOUS: Intrinsically motivated multi-task, multi-goal reinforcement learning
Colas, C., Oudeyer, P.-Y., Sigaud, O., Fournier, P., and Chetouani, M · 2019
Later among the works it cites.
Clic: Curriculum learning and imitation for object control in non-rewarding environments
Fournier, P., Colas, C., Chetouani, M., and Sigaud, O · 2019
Later among the works it cites.
Emergent systematic generalization in a situated agent, 2019
Hill, F., Lampinen, A., Schneider, R., Clark, S., Botvinick, M., McClelland, J. L., and Santoro, A · 2019
Later among the works it cites.
Language as an abstraction for hierarchical deep reinforcement learning
Jiang, Y., Gu, S. S., Murphy, K. P., and Finn, C · 2019
Later among the works it cites.
A survey of reinforcement learning informed by natural language
Original
Luketina, J., Nardelli, N., Farquhar, G., Foerster, J., Andreas, J., Grefenstette, E., Whiteson, S., and Rocktäschel, T · 2019
Later among the works it cites.
Contextual imagined goals for self-supervised robotic learning
Original
Nair, A., Bahl, S., Khazatsky, A., Pong, V., Berseth, G., and Levine, S · 2019
Later among the works it cites.
Skew-fit: State-covering self-supervised reinforcement learning
Original
Pong, V. H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S · 2019
Later among the works it cites.
Automated curricula through setter-solver interactions
Original
Racaniere, S., Lampinen, A. K., Santoro, A., Reichert, D. P., Firoiu, V., and Lillicrap, T. P · 2019
Later among the works it cites.
Grounding natural language commands to starcraft ii game states for narration-guided reinforcement learning
Waytowich, N., Barton, S. L., Lawhern, V., Stump, E., and Warnell, G · 2019
Later among the works it cites.
Rtfm: Generalising to novel environment dynamics via reading
Original
Zhong, V., Rocktäschel, T., and Grefenstette, E · 2019
Later among the works it cites.