Watch-and-help: A challenge for social perception and human-ai collaboration
Original
X. Puig, T. Shu, S. Li, Z. Wang, J. B. Tenenbaum, S. Fidler, and A. Torralba · 2020
Later among the works it cites.
igibson, a simulation environment for interactive tasks in large realisticscenes
Original
B. Shen, F. Xia, C. Li, R. Martín-Martín, L. Fan, G. Wang, S. Buch, C. D’Arpino, S. Srivastava, L. P. Tchapmi, et al · 2020
Later among the works it cites.
Alfred: A benchmark for interpreting grounded instructions for everyday tasks
M. Shridhar, J. Thomason, D. Gordon, Y. Bisk, W. Han, R. Mottaghi, L. Zettlemoyer, and D. Fox · 2020
Later among the works it cites.
Bert representations for video question answering
Z. Yang, N. Garcia, C. Chu, M. Otani, Y. Nakashima, and H. Takemura · 2020
Later among the works it cites.
Incorporating bert into neural machine translation
Original
J. Zhu, Y. Xia, L. Wu, D. He, T. Qin, W. Zhou, H. Li, and T.-Y. Liu · 2020
Later among the works it cites.
Multitasking inhibits semantic drift
M. L. Athul Paul Jacob and J. Andreas · 2021
Later among the works it cites.
Decision transformer: Reinforcement learning via sequence modeling
Original
L. Chen, K. Lu, A. Rajeswaran, K. Lee, A. Grover, M. Laskin, P. Abbeel, A. Srinivas, and I. Mordatch · 2021
Later among the works it cites.
Pretrained transformers as universal computation engines
Original
K. Lu, A. Grover, P. Abbeel, and I. Mordatch · 2021
Later among the works it cites.
Value-agnostic conversational semantic parsing
E. A. Platanios, A. Pauls, S. Roy, Y. Zhang, A. Kyte, A. Guo, S. Thomson, J. Krishnamurthy, J. Wolfe, J. Andreas, et al · 2021
Later among the works it cites.
Stable-baselines3: Reliable reinforcement learning implementations
A. Raffin, A. Hill, A. Gleave, A. Kanervisto, M. Ernestus, and N. Dormann · 2021
Later among the works it cites.
Multimodal few-shot learning with frozen language models
Original
M. Tsimpoukelli, J. Menick, S. Cabi, S. Eslami, O. Vinyals, and F. Hill · 2021
Later among the works it cites.
Learning with latent language
J. Andreas and D. Klein · 2022
Closest in time.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
Original
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch · 2022
Closest in time.
Bc-z: Zero-shot task generalization with robotic imitation learning
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn · 2022
Closest in time.
Can wikipedia help offline reinforcement learning?
Original
M. Reid, Y. Yamada, and S. S. Gu · 2022
Closest in time.
Skill induction and planning with latent language
P. Sharma, A. Torralba, and J. Andreas · 2022
Closest in time.