Fetching the paper…
Reading the bibliography…
Large vision-language models (VLMs) have shown significant performance boost in various application domains.
Catastrophic forgetting in connectionist networks
Robert M French · 1999
Earlier work this paper cites.
Vqa: Visual question answering, 2016
Aishwarya Agrawal, Jiasen Lu, Stanislaw Antol, Margaret Mitchell, C. Lawrence Zitnick, Dhruv Batra, and Devi Parikh · 2016
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning, 2016
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C. Lawrence Zitnick, and Ross Girshick · 2016
Earlier work this paper cites.
Making the v in vqa matter: Elevating the role of image understanding in visual question answering, 2017
Yash Goyal, Tejas Khot, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell · 2017
Earlier work this paper cites.
Learning without forgetting, 2017
Zhizhong Li and Derek Hoiem · 2017
Earlier work this paper cites.
icarl: Incremental classifier and representation learning, 2017
Sylvestre-Alvise Rebuffi, Alexander Kolesnikov, Georg Sperl, and Christoph H. Lampert · 2017
Earlier work this paper cites.
Visualbert: A simple and performant baseline for vision and language, 2019
Liunian Harold Li, Mark Yatskar, Da Yin, Cho-Jui Hsieh, and Kai-Wei Chang · 2019
Earlier work this paper cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks, 2019
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee · 2019
Earlier work this paper cites.
Ok-vqa: A visual question answering benchmark requiring external knowledge, 2019
Kenneth Marino, Mohammad Rastegari, Ali Farhadi, and Roozbeh Mottaghi · 2019
Earlier work this paper cites.
Experience replay for continual learning, 2019
David Rolnick, Arun Ahuja, Jonathan Schwarz, Timothy P. Lillicrap, and Greg Wayne · 2019
Earlier work this paper cites.
Three scenarios for continual learning
Gido M Van de Ven and Andreas S Tolias · 2019
Cited alongside, same era.
Language models are few-shot learners, 2020
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Cited alongside, same era.
Unsupervised machine learning via transfer learning and k-means clustering to classify materials image data
Ryan Cohn and Elizabeth Holm · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models, 2021
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Cited alongside, same era.
Learn continually, generalize rapidly: Lifelong knowledge accumulation for few-shot learning
Lifelong domain adaptation via consolidated internal distribution
Mohammad Rostami · 2021
Later among the works it cites.
Gem: A general evaluation benchmark for multimodal tasks, 2021
Lin Su, Nan Duan, Edward Cui, Lei Ji, Chenfei Wu, Huaishao Luo, Yongfei Liu, Ming Zhong, Taroon Bharti, and Arun Sacheti · 2021
Later among the works it cites.
Dytox: Transformers for continual learning with dynamic token expansion, 2022
Arthur Douillard, Alexandre Ramé, Guillaume Couairon, and Matthieu Cord · 2022
Later among the works it cites.
Symbolic replay: Scene graph as prompt for continual learning on vqa task, 2022
Stan Weixian Lei, Difei Gao, Jay Zhangjie Wu, Yuxuan Wang, Wei Liu, Mengmi Zhang, and Mike Zheng Shou · 2022
Later among the works it cites.
Gradient episodic memory for continual learning, 2022
David Lopez-Paz and Marc’Aurelio Ranzato · 2022
Later among the works it cites.
Climb: A continual learning benchmark for vision-and-language tasks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xisen Jin, Bill Yuchen Lin, Mohammad Rostami, and Xiang Ren · 2021
Cited alongside, same era.
Vilt: Vision-and-language transformer without convolution or region supervision, 2021
Wonjae Kim, Bokyung Son, and Ildoo Kim · 2021
Cited alongside, same era.
The power of scale for parameter-efficient prompt tuning, 2021
Brian Lester, Rami Al-Rfou, and Noah Constant · 2021
Cited alongside, same era.
Align before fuse: Vision and language representation learning with momentum distillation, 2021
Junnan Li, Ramprasaath R. Selvaraju, Akhilesh Deepak Gotmare, Shafiq Joty, Caiming Xiong, and Steven Hoi · 2021
Cited alongside, same era.
Prefix-tuning: Optimizing continuous prompts for generation, 2021
Xiang Lisa Li and Percy Liang · 2021
Cited alongside, same era.
Adapterfusion: Non-destructive task composition for transfer learning, 2021
Jonas Pfeiffer, Aishwarya Kamath, Andreas Rücklé, Kyunghyun Cho, and Iryna Gurevych · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision, 2021
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Cited alongside, same era.
Task-attentive transformer architecture for continual learning of vision-and-language tasks using knowledge distillation, 2023b
Yuliang Cai, Jesse Thomason, and Mohammad Rostami
Cited in the paper.
Tejas Srinivasan, Ting-Yun Chang, Leticia Pinto Alva, Georgios Chochlakis, Mohammad Rostami, and Jesse Thomason · 2022
Later among the works it cites.
Task-attentive transformer architecture for continual learning of vision-and-language tasks using knowledge distillation
Yuliang Cai, Jesse Thomason, and Mohammad Rostami · 2023
Later among the works it cites.
Promptcap: Prompt-guided image captioning for vqa with gpt-3
Yushi Hu, Hang Hua, Zhengyuan Yang, Weijia Shi, Noah A Smith, and Jiebo Luo · 2023
Later among the works it cites.
Cognitively inspired learning of incremental drifting concepts
Mohammad Rostami and Aram Galstyan · 2023
Later among the works it cites.
S-prompts learning with pre-trained transformers: An occam’s razor for domain incremental learning, 2023
Yabin Wang, Zhiwu Huang, and Xiaopeng Hong · 2023
Later among the works it cites.