Fetching the paper…
Reading the bibliography…
Visual language tasks require AI models to comprehend and reason with both visual and textual content.
Lee, K.; Palangi, H.; Chen, X.; Hu, H.; and Gao, J. 2019 · 1909
Earlier work this paper cites.
Fine-tuning language models from human preferences
Ziegler, D. M.; Stiennon, N.; Wu, J.; Brown, T. B.; Radford, A.; Amodei, D.; Christiano, P.; and Irving, G. 2019 · 1909
Earlier work this paper cites.
Language in mind: Advances in the study of language and thought
Goldin-Meadow, S.; and Gentner, D. 2003 · 2003
Earlier work this paper cites.
Unifiedqa: Crossing format boundaries with a single qa system
Khashabi, D.; Min, S.; Khot, T.; Sabharwal, A.; Tafjord, O.; Clark, P.; and Hajishirzi, H. 2020 · 2005
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Bowman, S. R.; Angeli, G.; Potts, C.; and Manning, C. D. 2015 · 2015
Earlier work this paper cites.
Flickr30k entities: Collecting region-to-phrase correspondences for richer image-to-sentence models
Plummer, B. A.; Wang, L.; Cervantes, C. M.; Caicedo, J. C.; Hockenmaier, J.; and Lazebnik, S. 2015 · 2015
Earlier work this paper cites.
Learning to Communicate with Deep Multi-Agent Reinforcement Learning
Foerster, J. N.; Assael, Y. M.; de Freitas, N.; and Whiteson, S. 2016 · 2016
Earlier work this paper cites.
Pre-training neural networks with human demonstrations for deep reinforcement learning
Cruz Jr, G. V.; Du, Y.; and Taylor, M. E. 2017 · 2017
Earlier work this paper cites.
Representation and Computation in Cognitive Models
Forbus, K.; Liang, C.; and Rabkina, I. 2017 · 2017
Earlier work this paper cites.
Sequence tutor: Conservative fine-tuning of sequence generation models with kl-control
Jaques, N.; Gu, S.; Bahdanau, D.; Hernández-Lobato, J. M.; Turner, R. E.; and Eck, D. 2017 · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J.; Wolski, F.; Dhariwal, P.; Radford, A.; and Klimov, O. 2017 · 2017
Earlier work this paper cites.
Learning from Unannotated QA Pairs to Analogically Disanbiguate and Answer Questions
Crouse, M.; McFate, C.; and Forbus, K. D. 2018 · 2018
Earlier work this paper cites.
Visual entailment task for visually-grounded language learning
Xie, N.; Lai, F.; Doran, D.; and Kadav, A. 2018 · 2018
Earlier work this paper cites.
Mapping Natural-language Problems to Formal-language Solutions Using Structured Neural Representations
Chen, K.; Huang, Q.; Palangi, H.; Smolensky, P.; Forbus, K.; and Gao, J. 2020 · 2020
Cited alongside, same era.
Learning to summarize with human feedback
Stiennon, N.; Ouyang, L.; Wu, J.; Ziegler, D.; Lowe, R.; Voss, C.; Radford, A.; Amodei, D.; and Christiano, P. F. 2020 · 2020
Cited alongside, same era.
TRL: Transformer Reinforcement Learning
von Werra, L.; Belkada, Y.; Tunstall, L.; Beeching, E.; Thrush, T.; and Lambert, N. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-Art Natural Language Processing
Wolf, T.; Debut, L.; Sanh, V.; Chaumond, J.; Delangue, C.; Moi, A.; Cistac, P.; Rault, T.; Louf, R.; Funtowicz, M.; Davison, J.; Shleifer, S.; von Platen, P.; Ma, C.; Jernite, Y.; Plu, J.; Xu, C.; Scao, T. L.; Gugger, S.; Drame, M.; Lhoest, Q.; and Rush, A. M. 2020 · 2020
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Hu, E. J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W. 2021 · 2021
Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework
Wang, P.; Yang, A.; Men, R.; Lin, J.; Bai, S.; Li, Z.; Ma, J.; Zhou, C.; Zhou, J.; and Yang, H. 2022 · 2022
Later among the works it cites.
An empirical study of gpt-3 for few-shot knowledge-based vqa
Yang, Z.; Gan, Z.; Wang, J.; Hu, X.; Lu, Y.; Liu, Z.; and Wang, L. 2022 · 2022
Later among the works it cites.
Chen, F.; Han, M.; Zhao, H.; Zhang, Q.; Shi, J.; Xu, S.; and Xu, B. 2023 · 2023
Closest in time.
Everything to Know About Your Internal Monologue
Cherney, K. 2023 · 2023
Closest in time.
Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality
Chiang, W.-L.; Li, Z.; Lin, Z.; Sheng, Y.; Wu, Z.; Zhang, H.; Zheng, L.; Zhuang, S.; Zhuang, Y.; Gonzalez, J. E.; Stoica, I.; and Xing, E. P. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
Zhang, K.; Yang, Z.; and Başar, T. 2021 · 2021
Cited alongside, same era.
PaLM: Scaling Language Modeling with Pathways
Chowdhery, A.; Narang, S.; Devlin, J.; Bosma, M.; Mishra, G.; Roberts, A.; Barham, P.; et al. 2022 · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models
Chung, H. W.; Hou, L.; Longpre, S.; Zoph, B.; Tay, Y.; Fedus, W.; Li, E.; Wang, X.; Dehghani, M.; Brahma, S.; et al. 2022 · 2022
Cited alongside, same era.
Transform-Retrieve-Generate: Natural Language-Centric Outside-Knowledge Visual Question Answering
Gao, F.; Ping, Q.; Thattai, G.; Reganti, A.; Wu, Y. N.; and Natarajan, P. 2022 · 2022
Cited alongside, same era.
Inner Monologue: Embodied Reasoning through Planning with Language Models
Huang, W.; Xia, F.; Xiao, T.; Chan, H.; Liang, J.; Florence, P.; et al. 2022 · 2022
Cited alongside, same era.
Learn to explain: Multimodal reasoning via thought chains for science question answering
Lu, P.; Mishra, S.; Xia, T.; Qiu, L.; Chang, K.-W.; Zhu, S.-C.; Tafjord, O.; Clark, P.; and Kalyan, A. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Ouyang, L.; Wu, J.; Jiang, X.; Almeida, D.; Wainwright, C.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al. 2022 · 2022
Cited alongside, same era.
Instructblip: Towards general-purpose vision-language models with instruction tuning
Dai, W.; Li, J.; Li, D.; Tiong, A. M. H.; Zhao, J.; Wang, W.; Li, B.; Fung, P.; and Hoi, S. 2023 · 2023
Closest in time.
Liu, H.; Li, C.; Wu, Q.; and Lee, Y. J. 2023 · 2023
Closest in time.
Chameleon: Plug-and-play compositional reasoning with large language models
Lu, P.; Peng, B.; Cheng, H.; Galley, M.; Chang, K.-W.; Wu, Y. N.; Zhu, S.-C.; and Gao, J. 2023 · 2023
Closest in time.
Image captioning for effective use of language models in knowledge-based visual question answering
Salaberria, A.; Azkune, G.; de Lacalle, O. L.; Soroa, A.; and Agirre, E. 2023 · 2023
Closest in time.
Stanford Alpaca: An Instruction-following LLaMA model
Taori, R.; Gulrajani, I.; Zhang, T.; Dubois, Y.; Li, X.; Guestrin, C.; Liang, P.; and Hashimoto, T. B. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H.; Lavril, T.; Izacard, G.; Martinet, X.; Lachaux, M.-A.; Lacroix, T.; Rozière, B.; Goyal, N.; Hambro, E.; Azhar, F.; et al. 2023 · 2023
Closest in time.
IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models
You, H.; Sun, R.; Wang, Z.; Chen, L.; Wang, G.; Ayyubi, H.; Chang, K.-W.; and Chang, S.-F. 2023 · 2023
Closest in time.
Multimodal chain-of-thought reasoning in language models
Zhang, Z.; Zhang, A.; Li, M.; Zhao, H.; Karypis, G.; and Smola, A. 2023 · 2023
Closest in time.