Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Original
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., and others (2016) · 2016
Later among the works it cites.
Deep reinforcement learning from human preferences
Christiano, P., Leike, J., Brown, T. B., Martic, M., Legg, S., and Amodei, D. (2017) · 2017
Later among the works it cites.
Grounded Language Learning in a Simulated 3d World
Original
Hermann, K. M., Hill, F., Green, S., Wang, F., Faulkner, R., Soyer, H., Szepesvari, D., Czarnecki, W. M., Jaderberg, M., Teplyashin, D., Wainwright, M., Apps, C., Hassabis, D., and Blunsom, P. (2017) · 2017
Later among the works it cites.
AI2-THOR: An Interactive 3d Environment for Visual AI
Original
Kolve, E., Mottaghi, R., Gordon, D., Zhu, Y., Gupta, A., and Farhadi, A. (2017) · 2017
Later among the works it cites.
FiLM: Visual Reasoning with a General Conditioning Layer
Perez, E., Strub, F., de Vries, H., Dumoulin, V., and Courville, A. (2017) · 2017
Later among the works it cites.
Proximal Policy Optimization Algorithms
Original
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O. (2017) · 2017
Later among the works it cites.
Deep TAMER: Interactive Agent Shaping in High-Dimensional State Spaces
Warnell, G., Waytowich, N., Lawhern, V., and Stone, P. (2017) · 2017
Later among the works it cites.
Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments
Anderson, P., Wu, Q., Teney, D., Bruce, J., Johnson, M., Sünderhauf, N., Reid, I., Gould, S., and Hengel, A. v. d. (2018) · 2018
Closest in time.
Learning to Understand Goal Specifications by Modelling Reward
Bahdanau, D., Hill, F., Leike, J., Hughes, E., Hosseini, A., Kohli, P., and Grefenstette, E. (2018) · 2018
Closest in time.
Gated-Attention Architectures for Task-Oriented Language Grounding
Chaplot, D. S., Sathyendra, K. M., Pasumarthi, R. K., Rajagopal, D., and Salakhutdinov, R. (2018) · 2018
Closest in time.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., Legg, S., and Kavukcuoglu, K. (2018) · 2018
Closest in time.
Learning to Parse Natural Language to Grounded Reward Functions with Weak Supervision
Williams, E. C., Gopalan, N., Rhee, M., and Tellex, S. (2018) · 2018
Closest in time.
Building Generalizable Agents with a Realistic and Rich 3d Environment
Original
Wu, Y., Wu, Y., Gkioxari, G., and Tian, Y. (2018) · 2018
Closest in time.
Interactive Grounded Language Acquisition and Generalization in 2d Environment
Yu, H., Zhang, H., and Xu, W. (2018) · 2018
Closest in time.