Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Later among the works it cites.
Flamingo: a visual language model for few-shot learning
Original
J.-B. Alayrac, J. Donahue, P. Luc, A. Miech, I. Barr, Y. Hasson, K. Lenc, A. Mensch, K. Millican, M. Reynolds, et al · 2022
Later among the works it cites.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Original
Y. Bai, A. Jones, K. Ndousse, A. Askell, A. Chen, N. DasSarma, D. Drain, S. Fort, D. Ganguli, T. Henighan, et al · 2022
Later among the works it cites.
Can foundation models perform zero-shot task specification for robot manipulation?
Y. Cui, S. Niekum, A. Gupta, V. Kumar, and A. Rajeswaran · 2022
Later among the works it cites.
Enabling multimodal generation on clip via vision-language knowledge distillation
Original
W. Dai, L. Hou, L. Shang, X. Jiang, Q. Liu, and P. Fung · 2022
Later among the works it cites.
Minedojo: Building open-ended embodied agents with internet-scale knowledge
Original
L. Fan, G. Wang, Y. Jiang, A. Mandlekar, Y. Yang, H. Zhu, A. Tang, D.-A. Huang, Y. Zhu, and A. Anandkumar · 2022
Later among the works it cites.
Improving alignment of dialogue agents via targeted human judgements
Original
A. Glaese, N. McAleese, M. Trebacz, J. Aslanides, V. Firoiu, T. Ewalds, M. Rauh, L. Weidinger, M. Chadwick, P. Thacker, L. Campbell-Gillingham, J. Uesato, P.-S. Huang, R. Comanescu, F. Yang, A. See, S. Dathathri, R. Greig, C. Chen, D. Fritz, J. S. Elias, R. Green, S. Mokrá, N. Fernando, B. Wu, R. Foley, S. Young, I. Gabriel, W. Isaac, J. Mellor, D. Hassabis, K. Kavukcuoglu, L. A. Hendricks, and G. Irving · 2022
Later among the works it cites.
Ego4d: Around the world in 3,000 hours of egocentric video
K. Grauman, A. Westbury, E. Byrne, Z. Chavis, A. Furnari, R. Girdhar, J. Hamburger, H. Jiang, M. Liu, X. Liu, et al · 2022
Later among the works it cites.
Vip: Towards universal visual reward and representation via value-implicit pre-training
Original
Y. J. Ma, S. Sodhani, D. Jayaraman, O. Bastani, V. Kumar, and A. Zhang · 2022
Later among the works it cites.
Zero-shot reward specification via grounded natural language
P. Mahmoudieh, D. Pathak, and T. Darrell · 2022
Later among the works it cites.
Teaching language models to support answers with verified quotes
Original
J. Menick, M. Trebacz, V. Mikulik, J. Aslanides, F. Song, M. Chadwick, M. Glaese, S. Young, L. Campbell-Gillingam, G. Irving, and N. McAleese · 2022
Later among the works it cites.
Webgpt: Browser-assisted question-answering with human feedback
Original
R. Nakano, J. Hilton, S. Balaji, J. Wu, L. Ouyang, C. Kim, C. Hesse, S. Jain, V. Kosaraju, W. Saunders, X. Jiang, K. Cobbe, T. Eloundou, G. Krueger, K. Button, M. Knight, B. Chess, and J. Schulman · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Original
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al · 2022
Later among the works it cites.
A generalist agent
Original
S. Reed, K. Zolna, E. Parisotto, S. G. Colmenarejo, A. Novikov, G. Barth-Maron, M. Gimenez, Y. Sulsky, J. Kay, J. T. Springenberg, T. Eccles, J. Bruce, A. Razavi, A. Edwards, N. Heess, Y. Chen, R. Hadsell, O. Vinyals, M. Bordbar, and N. de Freitas · 2022
Later among the works it cites.
Plug-and-play vqa: Zero-shot vqa by conjoining large pretrained models with zero training
Original
A. M. H. Tiong, J. Li, B. Li, S. Savarese, and S. C. Hoi · 2022
Later among the works it cites.
Mastering diverse domains through world models
Original
D. Hafner, J. Pasukonis, J. Ba, and T. Lillicrap · 2023
Closest in time.
Grounding language models to images for multimodal generation
Original
J. Y. Koh, R. Salakhutdinov, and D. Fried · 2023
Closest in time.
Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Original
J. Li, D. Li, S. Savarese, and S. Hoi · 2023
Closest in time.