Fetching the paper…
Reading the bibliography…
Recently, foundation language models (LMs) have marked significant achievements in the domains of natural language processing (NLP) and computer vision (CV).
An analysis of visual question answering algorithms. In ICCV . 1965–1973
Kushal Kafle and Christopher Kanan. 2017 · 1973
Earlier work this paper cites.
Newsweeder: Learning to filter netnews
Ken Lang. 1995 · 1995
Earlier work this paper cites.
Secure multi-party computation
Oded Goldreich. 1998 · 1998
Earlier work this paper cites.
Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition
Erik F Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
Mining and summarizing customer reviews. In SIGKDD . 168–177
Minqing Hu and Bing Liu. 2004 · 2004
Earlier work this paper cites.
Differential privacy. In International colloquium on automata, languages, and programming . 1–12
Cynthia Dwork. 2006 · 2006
Earlier work this paper cites.
The iapr tc-12 benchmark: A new evaluation resource for visual information systems. In International workshop ontoImage , Vol. 2
Michael Grubinger, Paul Clough, Henning Müller, and Thomas Deselaers. 2006 · 2006
Earlier work this paper cites.
OntoNotes: The 90% Solution. In NAACL . 57–60
Eduard Hovy, Mitchell Marcus, Martha Palmer, Lance Ramshaw, and Ralph Weischedel. 2006 · 2006
Earlier work this paper cites.
A holistic lexicon-based approach to opinion mining. In WSDM . 231–240
Xiaowen Ding, Bing Liu, and Philip S Yu. 2008 · 2008
Earlier work this paper cites.
Learning Word Vectors for Sentiment Analysis. In ACL . 142–150
Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011 · 2011
Earlier work this paper cites.
Microsoft coco: Common objects in context. In ECCV . 740–755
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Peter Young, Alice Lai, Micah Hodosh, and Julia Hockenmaier. 2014 · 2014
Earlier work this paper cites.
Large-scale simple question answering with memory networks
Antoine Bordes, Nicolas Usunier, Sumit Chopra, and Jason Weston. 2015 · 2015
Earlier work this paper cites.
Automated rule selection for aspect extraction in opinion mining. In IJCAI
Qian Liu, Zhiqiang Gao, Bing Liu, and Yuanlin Zhang. 2015 · 2015
Earlier work this paper cites.
Semantically conditioned lstm-based natural language generation for spoken dialogue systems
Tsung-Hsien Wen, Milica Gasic, Nikola Mrksic, Pei-Hao Su, David Vandyke, and Steve Young. 2015 · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Earlier work this paper cites.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books. In ICCV . 19–27
Yukun Zhu, Ryan Kiros, Rich Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering. In WWW . 507–517
Ruining He and Julian McAuley. 2016 · 2016
Earlier work this paper cites.
SQuAD: 100,000+ Questions for Machine Comprehension of Text. In EMNLP . 2383–2392
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Earlier work this paper cites.
Tweets Dataset - Top 20 most followed users in Twitter social platform
Raad Bin Tareaf. 2017 · 2017
Earlier work this paper cites.
Making the v in vqa matter: Elevating the role of image understanding in visual question answering. In CVPR . 6904–6913
Yash Goyal, Tejas Khot, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2017 · 2017
Earlier work this paper cites.
TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension. In ACL . 1601–1611
Mandar Joshi, Eunsol Choi, Daniel Weld, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al · 2017
Earlier work this paper cites.
Hdltex: Hierarchical deep learning for text classification. In ICMLA . 364–371
Kamran Kowsari, Donald E Brown, Mojtaba Heidarysafa, Kiana Jafari Meimandi, Matthew S Gerber, and Laura E Barnes. 2017 · 2017
Earlier work this paper cites.
Toward continual learning for conversational agents
Sungjin Lee. 2017 · 2017
Earlier work this paper cites.
Gradient episodic memory for continual learning
David Lopez-Paz and Marc’Aurelio Ranzato. 2017 · 2017
Earlier work this paper cites.
Exploring models and data for remote sensing image caption generation
Xiaoqiang Lu, Binqiang Wang, Xiangtao Zheng, and Xuelong Li. 2017 · 2017
Earlier work this paper cites.
The E2E dataset: New challenges for end-to-end generation
Jekaterina Novikova, Ondřej Dušek, and Verena Rieser. 2017 · 2017
Earlier work this paper cites.
iCaRL: Incremental Classifier and Representation Learning. In CVPR
Sylvestre-Alvise Rebuffi, Alexander Kolesnikov, Georg Sperl, and Christoph H. Lampert. 2017 · 2017
Earlier work this paper cites.
Get to the point: Summarization with pointer-generator networks
Abigail See, Peter J Liu, and Christopher D Manning. 2017 · 2017
Earlier work this paper cites.
TL;DR: Mining Reddit to Learn Automatic Summarization. In the Workshop on New Frontiers in Summarization . 59–63
Michael Volske, Martin Potthast, Shahbaz Syed, and Benno Stein. 2017 · 2017
Earlier work this paper cites.
Self-taught convolutional neural networks for short text clustering
Jiaming Xu, Bo Xu, Peng Wang, Suncong Zheng, Guanhua Tian, and Jun Zhao. 2017 · 2017
Earlier work this paper cites.
Seq2sql: Generating structured queries from natural language using reinforcement learning
Victor Zhong, Caiming Xiong, and Richard Socher. 2017 · 2017
Earlier work this paper cites.
Memory aware synapses: Learning what (not) to forget. In ECCV . 139–154
Rahaf Aljundi, Francesca Babiloni, Mohamed Elhoseiny, Marcus Rohrbach, and Tinne Tuytelaars. 2018 · 2018
Earlier work this paper cites.
The description length of deep learning models
Léonard Blier and Yann Ollivier. 2018 · 2018
Earlier work this paper cites.
MultiWOZ–a large-scale multi-domain wizard-of-oz dataset for task-oriented dialogue modelling
Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng, Inigo Casanueva, Stefan Ultes, Osman Ramadan, and Milica Gašić. 2018 · 2018
Earlier work this paper cites.
e-snli: Natural language inference with natural language explanations
Oana-Maria Camburu, Tim Rocktäschel, Thomas Lukasiewicz, and Phil Blunsom. 2018 · 2018
Earlier work this paper cites.
Riemannian walk for incremental learning: Understanding forgetting and intransigence. In ECCV . 532–547
Arslan Chaudhry, Puneet K Dokania, Thalaiyasingam Ajanthan, and Philip HS Torr. 2018 · 2018
Earlier work this paper cites.
QuAC: Question Answering in Context. In EMNLP . 2174–2184
Eunsol Choi, He He, Mohit Iyyer, Mark Yatskar, Wen-tau Yih, Yejin Choi, Percy Liang, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
Alice Coucke, Alaa Saade, Adrien Ball, Théodore Bluche, Alexandre Caulier, David Leroy, Clément Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, et al · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin. 2018 · 2018
Earlier work this paper cites.
Semantic parsing for task oriented dialog using hierarchical representations
Sonal Gupta, Rushin Shah, Mrinal Mohit, Anuj Kumar, and Mike Lewis. 2018 · 2018
Earlier work this paper cites.
Xu Han, Hao Zhu, Pengfei Yu, Ziyun Wang, Yuan Yao, Zhiyuan Liu, and Maosong Sun. 2018 · 2018
Earlier work this paper cites.
Mapping language to code in programmatic context
Srinivasan Iyer, Ioannis Konstas, Alvin Cheung, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
Measuring catastrophic forgetting in neural networks. In AAAI , Vol. 32
Ronald Kemker, Marc McClure, Angelina Abitino, Tyler Hayes, and Christopher Kanan. 2018 · 2018
Earlier work this paper cites.
The natural language decathlon: Multitask learning as question answering. arXiv 2018
Bryan McCann, Nitish Shirish Keskar, Caiming Xiong, and Richard Socher. [n. d.] · 2018
Earlier work this paper cites.
Improving language understanding with unsupervised learning
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever. 2018 · 2018
Earlier work this paper cites.
Cross-lingual transfer learning for multilingual task oriented dialog
Sebastian Schuster, Sonal Gupta, Rushin Shah, and Mike Lewis. 2018 · 2018
Earlier work this paper cites.
Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning. In ACL . 2556–2565
Piyush Sharma, Nan Ding, Sebastian Goodman, and Radu Soricut. 2018 · 2018
Earlier work this paper cites.
GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding. In EMNLP . 353–355
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
Taskmaster-1: Toward a realistic and diverse dialog dataset
Bill Byrne, Karthik Krishnamoorthi, Chinnadhurai Sankar, Arvind Neelakantan, Daniel Duckworth, Semih Yavuz, Ben Goodrich, Amit Dubey, Andy Cedilnik, and Kyu-Young Kim. 2019 · 2019
Earlier work this paper cites.
Efficient Lifelong Learning with A-GEM. In ICLR
Arslan Chaudhry, Marc’Aurelio Ranzato, Marcus Rohrbach, and Mohamed Elhoseiny. 2019 · 2019
Earlier work this paper cites.
Episodic memory in lifelong language learning
Cyprien de Masson D’Autume, Sebastian Ruder, Lingpeng Kong, and Dani Yogatama. 2019 · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for NLP. In ICML . 2790–2799
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019 · 2019
Earlier work this paper cites.
Codesearchnet challenge: Evaluating the state of semantic code search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt. 2019 · 2019
Earlier work this paper cites.
Meta-learning representations for continual learning
Khurram Javed and Martha White. 2019 · 2019
Earlier work this paper cites.
An evaluation dataset for intent classification and out-of-scope prediction
Stefan Larson, Anish Mahendran, Joseph J Peper, Christopher Clarke, Andrew Lee, Parker Hill, Jonathan K Kummerfeld, Kevin Leach, Michael A Laurenzano, Lingjia Tang, et al · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019a · 2019
Earlier work this paper cites.
Continual lifelong learning with neural networks: A review
German I Parisi, Ronald Kemker, Jose L Part, Christopher Kanan, and Stefan Wermter. 2019 · 2019
Earlier work this paper cites.
A progressive model to enable continual learning for semantic slot filling. In EMNLP . 1279–1284
Yilin Shen, Xiangyu Zeng, and Hongxia Jin. 2019 · 2019
Earlier work this paper cites.
An empirical study on learning bug-fixing patches in the wild via neural machine translation
Michele Tufano, Cody Watson, Gabriele Bavota, Massimiliano Di Penta, Martin White, and Denys Poshyvanyk. 2019 · 2019
Earlier work this paper cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2019a · 2019
Earlier work this paper cites.
BERT post-training for review reading comprehension and aspect-based sentiment analysis
Hu Xu, Bing Liu, Lei Shu, and Philip S Yu. 2019 · 2019
Earlier work this paper cites.
Learning and evaluating general linguistic intelligence
Dani Yogatama, Cyprien de Masson d’Autume, Jerome Connor, Tomas Kocisky, Mike Chrzanowski, Lingpeng Kong, Angeliki Lazaridou, Wang Ling, Lei Yu, Chris Dyer, et al · 2019
Earlier work this paper cites.
Defending against neural fake news
Rowan Zellers, Ari Holtzman, Hannah Rashkin, Yonatan Bisk, Ali Farhadi, Franziska Roesner, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
Shawn Beaulieu, Lapo Frati, Thomas Miconi, Joel Lehman, Kenneth O Stanley, Jeff Clune, and Nick Cheney. 2020 · 2020
Earlier work this paper cites.
Continual Lifelong Learning in Natural Language Processing: A Survey. In COLING . 6523–6541
Magdalena Biesialska, Katarzyna Biesialska, and Marta R Costa-jussà. 2020 · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom B Brown. 2020 · 2020
Earlier work this paper cites.
Efficient intent detection with dual sentence encoders
Iñigo Casanueva, Tadas Temčinas, Daniela Gerz, Matthew Henderson, and Ivan Vulić. 2020 · 2020
Earlier work this paper cites.
Recall and learn: Fine-tuning deep pretrained language models with less forgetting
Sanyuan Chen, Yutai Hou, Yiming Cui, Wanxiang Che, Ting Liu, and Xiangzhan Yu. 2020 · 2020
Cited alongside, same era.
Podnet: Pooled outputs distillation for small-tasks incremental learning. In ECCV . 86–102
Arthur Douillard, Matthieu Cord, Charles Ollion, Thomas Robert, and Eduardo Valle. 2020 · 2020
Cited alongside, same era.
Meta-learning with sparse experience replay for lifelong language learning
Nithin Holla, Pushkar Mishra, Helen Yannakoudakis, and Ekaterina Shutova. 2020 · 2020
Cited alongside, same era.
Continual learning for robotics: Definition, framework, learning strategies, opportunities and challenges
Timothée Lesort, Vincenzo Lomonaco, Andrei Stoian, Davide Maltoni, David Filliat, and Natalia Díaz-Rodríguez. 2020 · 2020
Cited alongside, same era.
Learning on the job: Online lifelong and continual learning. In AAAI , Vol. 34. 13544–13549
Lifelong language pretraining with distribution-specialized experts. In ICML . PMLR, 5383–5395
Wuyang Chen, Yanqi Zhou, Nan Du, Yanping Huang, James Laudon, Zhifeng Chen, and Claire Cui. 2023 · 2023
Later among the works it cites.
Adapting large language models via reading comprehension
Daixuan Cheng, Shaohan Huang, and Furu Wei. 2023a · 2023
Later among the works it cites.
From Static to Dynamic: A Continual Learning Framework for Large Language Models
Mingzhe Du, Anh Tuan Luu, Bin Ji, and See-kiong Ng. 2023 · 2023
Later among the works it cites.
A study of continual learning under language shift
Evangelia Gogoulou, Timothée Lesort, Magnus Boman, and Joakim Nivre. 2023 · 2023
Later among the works it cites.
Exploring the benefits of training expert language models over instruction tuning. In ICML . 14702–14729
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bing Liu. 2020 · 2020
Cited alongside, same era.
S2ORC: The Semantic Scholar Open Research Corpus. In ACL . 4969–4983
Kyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney, and Daniel Weld. 2020 · 2020
Cited alongside, same era.
Continual learning in task-oriented dialogue systems
Andrea Madotto, Zhaojiang Lin, Zhenpeng Zhou, Seungwhan Moon, Paul Crook, Bing Liu, Zhou Yu, Eunjoon Cho, and Zhiguang Wang. 2020 · 2020
Cited alongside, same era.
Towards scalable multi-domain conversational agents: The schema-guided dialogue dataset. In AAAI , Vol. 34. 8689–8696
Abhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta, and Pranav Khaitan. 2020 · 2020
Cited alongside, same era.
Efficient Meta Lifelong-Learning with Limited Memory. In EMNLP . 535–548
Zirui Wang, Sanket Vaibhav Mehta, Barnabas Poczos, and Jaime Carbonell. 2020 · 2020
Cited alongside, same era.
Parameter-efficient transfer from sequential behaviors for user modeling and recommendation. In SIGIR . 1469–1478
Fajie Yuan, Xiangnan He, Alexandros Karatzoglou, and Liguang Zhang. 2020 · 2020
Cited alongside, same era.
A comprehensive study of class incremental learning algorithms for visual tasks
Eden Belouadah, Adrian Popescu, and Ioannis Kanellos. 2021 · 2021
Cited alongside, same era.
Learning to Solve NLP Tasks in an Incremental Number of Languages. In ACL . 837–847
Giuseppe Castellucci, Simone Filice, Danilo Croce, and Roberto Basili. 2021 · 2021
Cited alongside, same era.
Joel Jang, Seungone Kim, Seonghyeon Ye, Doyoung Kim, Lajanugen Logeswaran, Moontae Lee, Kyungjae Lee, and Minjoon Seo. 2023 · 2023
Later among the works it cites.
Continual Pre-training of Language Models. In ICLR
Zixuan Ke, Yijia Shao, Haowei Lin, Tatsuya Konishi, Gyuhak Kim, and Bing Liu. 2023 · 2023
Later among the works it cites.
Introducing language guidance in prompt-based continual learning. In ICCV . 11463–11473
Muhammad Gul Zain Ali Khan, Muhammad Ferjad Naeem, Luc Van Gool, Didier Stricker, Federico Tombari, and Muhammad Zeshan Afzal. 2023 · 2023
Later among the works it cites.
Task Relation-aware Continual User Representation Learning. In SIGKDD . 1107–1119
Sein Kim, Namkyeong Lee, Donghyun Kim, Minchul Yang, and Chanyoung Park. 2023 · 2023
Later among the works it cites.
Do pre-trained models benefit equally in continual learning?. In WCACV . 6485–6493
Kuan-Ying Lee, Yuanyi Zhong, and Yu-Xiong Wang. 2023 · 2023
Later among the works it cites.
Online noisy continual relation learning. In AAAI , Vol. 37. 13059–13066
Guozheng Li, Peng Wang, Qiqing Luo, Yanhe Liu, and Wenjun Ke. 2023 · 2023
Later among the works it cites.
AI Autonomy: Self-initiated Open-world Continual Learning and Adaptation
Bing Liu, Sahisnu Mazumder, Eric Robertson, and Scott Grigsby. 2023b · 2023
Later among the works it cites.
Class Incremental Learning with Pre-trained Vision-Language Models
Xialei Liu, Xusheng Cao, Haori Lu, Jia-wen Xiao, Andrew D Bagdanov, and Ming-Ming Cheng. 2023a · 2023
Later among the works it cites.
An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
Yun Luo, Zhen Yang, Fandong Meng, Yafu Li, Jie Zhou, and Yue Zhang. 2023 · 2023
Later among the works it cites.
EcomGPT-CT: Continual pre-training of e-commerce large language models with semi-structured data
Shirong Ma, Shen Huang, Shulin Huang, Xiaobin Wang, Yangning Li, Hai-Tao Zheng, Pengjun Xie, Fei Huang, and Yong Jiang. 2023 · 2023
Later among the works it cites.
Generative Replay Inspired by Hippocampal Memory Indexing for Continual Language Learning. In EACL . 930–942
Aru Maekawa, Hidetaka Kamigaito, Kotaro Funakoshi, and Manabu Okumura. 2023 · 2023
Later among the works it cites.
Umberto Michieli, Pablo Peso Parada, and Mete Ozay. 2023 · 2023
Later among the works it cites.
Recent advances in natural language processing via large pre-trained language models: A survey
Bonan Min, Hayley Ross, Elior Sulem, Amir Pouran Ben Veyseh, Thien Huu Nguyen, Oscar Sainz, Eneko Agirre, Ilana Heintz, and Dan Roth. 2023 · 2023
Later among the works it cites.
Large-scale Lifelong Learning of In-context Instructions and How to Tackle It. In ACL . 12573–12589
Jisoo Mok, Jaeyoung Do, Sungjin Lee, Tara Taghavi, Seunghak Yu, and Sungroh Yoon. 2023 · 2023
Later among the works it cites.
Continual vision-language representation learning with off-diagonal information. In ICML . 26129–26149
Zixuan Ni, Longhui Wei, Siliang Tang, Yueting Zhuang, and Qi Tian. 2023 · 2023
Later among the works it cites.
Adapters: A Unified Library for Parameter-Efficient and Modular Transfer Learning. In EMNLP . 149–160
Clifton Poth, Hannah Sterz, Indraneil Paul, Sukannya Purkayastha, Leon Engländer, Timo Imhof, Ivan Vulić, Sebastian Ruder, Iryna Gurevych, and Jonas Pfeiffer. 2023 · 2023
Later among the works it cites.
Decouple before interact: Multi-modal prompt learning for continual visual question answering. In ICCV . 2953–2962
Zi Qian, Xin Wang, Xuguang Duan, Pengda Qin, Yuhong Li, and Wenwu Zhu. 2023 · 2023
Later among the works it cites.
Recyclable Tuning for Continual Pre-training. In ACL . 11403–11426
Yujia Qin, Cheng Qian, Xu Han, Yankai Lin, Huadong Wang, Ruobing Xie, Zhiyuan Liu, Maosong Sun, and Jie Zhou. 2023 · 2023
Later among the works it cites.
Progressive Prompts: Continual Learning for Language Models. In ICLR
Anastasia Razdaibiedina, Yuning Mao, Rui Hou, Madian Khabsa, Mike Lewis, and Amjad Almahairi. 2023 · 2023
Later among the works it cites.
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
Chenyang Song, Xu Han, Zheni Zeng, Kuai Li, Chen Chen, Zhiyuan Liu, Maosong Sun, and Tao Yang. 2023 · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Later among the works it cites.
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
Steven Vander Eeckt et al · 2023
Later among the works it cites.
Large-scale multi-modal pre-trained models: A comprehensive survey
Xiao Wang, Guangyao Chen, Guangwu Qian, Pengcheng Gao, Xiao-Yong Wei, Yaowei Wang, Yonghong Tian, and Wen Gao. 2023b · 2023
Later among the works it cites.
Domain-Agnostic Neural Architecture for Class Incremental Continual Learning in Document Processing Platform. In ACL . 527–537
Mateusz Wójcik, Witold Kościukiewicz, Mateusz Baran, Tomasz Kajdanowicz, and Adam Gonczarek. 2023 · 2023
Later among the works it cites.
Online Continual Knowledge Learning for Language Models
Yuhao Wu, Tongjun Shi, Karthick Sharma, Chun Wei Seah, and Shuhao Zhang. 2023 · 2023
Later among the works it cites.
Efficient continual pre-training for building domain specific large language models
Yong Xie, Karan Aggarwal, and Aitzaz Ahmad. 2023 · 2023
Later among the works it cites.
Exploring continual learning for code generation models
Prateek Yadav, Qing Sun, Hantian Ding, Xiaopeng Li, Dejiao Zhang, Ming Tan, Xiaofei Ma, Parminder Bhatia, Ramesh Nallapati, Murali Krishna Ramanathan, et al · 2023
Later among the works it cites.
Towards General Purpose Medical AI: Continual Learning Medical Foundation Model
Huahui Yi, Ziyuan Qin, Qicheng Lao, Wei Xu, Zekun Jiang, Dequan Wang, Shaoting Zhang, and Kang Li. 2023 · 2023
Later among the works it cites.
Copf: Continual learning human preference through optimal policy fitting
Han Zhang, Lin Gui, Yuanzhao Zhai, Hui Wang, Yu Lei, and Ruifeng Xu. 2023b · 2023
Later among the works it cites.
A Survey on Incremental Update for Neural Recommender Systems
Peiyan Zhang and Sunghun Kim. 2023 · 2023
Later among the works it cites.
Reformulating Domain Adaptation of Large Language Models as Adapt-Retrieve-Revise
Yating Zhang, Yexiang Wang, Fei Cheng, Sadao Kurohashi, et al · 2023
Later among the works it cites.
Citb: A benchmark for continual instruction tuning
Zihan Zhang, Meng Fang, Ling Chen, and Mohammad-Reza Namazi-Rad. 2023a · 2023
Later among the works it cites.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Later among the works it cites.
Preventing zero-shot transfer degradation in continual learning of vision-language models. In ICCV . 19125–19136
Zangwei Zheng, Mingyuan Ma, Kai Wang, Ziheng Qin, Xiangyu Yue, and Yang You. 2023 · 2023
Later among the works it cites.
Deep class-incremental learning: A survey
Da-Wei Zhou, Qi-Wei Wang, Zhi-Hong Qi, Han-Jia Ye, De-Chuan Zhan, and Ziwei Liu. 2023b · 2023
Later among the works it cites.
Learning without forgetting for vision-language models
Da-Wei Zhou, Yuanhan Zhang, Jingyi Ning, Han-Jia Ye, De-Chuan Zhan, and Ziwei Liu. 2023c · 2023
Later among the works it cites.
ChatGPT: potential, prospects, and limitations
Jie Zhou, Pei Ke, Xipeng Qiu, Minlie Huang, and Junping Zhang. 2023a · 2023
Later among the works it cites.
Ctp: Towards vision-language continual pretraining via compatible momentum contrast and topology preservation. In ICCV . 22257–22267
Hongguang Zhu, Yunchao Wei, Xiaodan Liang, Chunjie Zhang, and Yao Zhao. 2023 · 2023
Later among the works it cites.
Generative Multi-modal Models are Good Class-Incremental Learners
Xusheng Cao, Haori Lu, Linlan Huang, Xialei Liu, and Ming-Ming Cheng. 2024 · 2024
Closest in time.
Continual Vision-Language Retrieval via Dynamic Knowledge Rectification. In AAAI , Vol. 38. 11704–11712
Zhenyu Cui, Yuxin Peng, Xun Wang, Manyu Zhu, and Jiahuan Zhou. 2024 · 2024
Closest in time.
Boosting Large Language Models with Continual Learning for Aspect-based Sentiment Analysis. In EMNLP
Xuanwen Ding, Jie Zhou, Liang Dou, Qin Chen, Yuanbin Wu, Chengcai Chen, and Liang He. 2024 · 2024
Closest in time.
Datacomp: In search of the next generation of multimodal datasets
Samir Yitzhak Gadre, Gabriel Ilharco, Alex Fang, Jonathan Hayase, Georgios Smyrnis, Thao Nguyen, Ryan Marten, Mitchell Wortsman, Dhruba Ghosh, Jieyu Zhang, et al · 2024
Closest in time.
Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey
Zeyu Han, Chao Gao, Jinyang Liu, Sai Qian Zhang, et al · 2024
Closest in time.
Class-Incremental Learning with CLIP: Adaptive Representation Adjustment and Parameter Fusion
Linlan Huang, Xusheng Cao, Haori Lu, and Xialei Liu. 2024 · 2024
Closest in time.
CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models
Saurav Jha, Dong Gong, and Lina Yao. 2024 · 2024
Closest in time.
Junsu Kim, Yunhoe Ku, Jihyeon Kim, Junuk Cha, and Seungryul Baek. 2024 · 2024
Closest in time.
RoboCoder: Robotic Learning from Basic Skills to General Tasks with Large Language Models
Jingyao Li, Pengguang Chen, Sitong Wu, Chuanyang Zheng, Hong Xu, and Jiaya Jia. 2024 · 2024
Closest in time.
Mitigating the Alignment Tax of RLHF
Yong Lin, Hangyu Lin, Wei Xiong, Shizhe Diao, Jianmeng Liu, Jipeng Zhang, Rui Pan, Haoxiang Wang, Wenbin Hu, Hanning Zhang, Hanze Dong, Renjie Pi, Han Zhao, Nan Jiang, Heng Ji, Yuan Yao, and Tong Zhang. 2024 · 2024
Closest in time.
Self-Corrected Multimodal Large Language Model for End-to-End Robot Manipulation
Jiaming Liu, Chenxuan Li, Guanqun Wang, Lily Lee, Kaichen Zhou, Sixiang Chen, Chuyan Xiong, Jiaxin Ge, Renrui Zhang, and Shanghang Zhang. 2024 · 2024
Closest in time.
Eureka: Human-Level Reward Design via Coding Large Language Models. In ICLR
Yecheng Jason Ma, William Liang, Guanzhi Wang, De-An Huang, Osbert Bastani, Dinesh Jayaraman, Yuke Zhu, Linxi Fan, and Anima Anandkumar. 2024 · 2024
Closest in time.
Lifelong and Continual Learning Dialogue Systems
Sahisnu Mazumder and Bing Liu. 2024 · 2024
Closest in time.
Semantic Residual Prompts for Continual Learning
Martin Menabue, Emanuele Frascaroli, Matteo Boschini, Enver Sangineto, Lorenzo Bonicelli, Angelo Porrello, and Simone Calderara. 2024 · 2024
Closest in time.
A Comprehensive Analysis of Adapter Efficiency. In IKDD . 136–154
Nandini Mundra, Sumanth Doddapaneni, Raj Dabre, Anoop Kunchukuttan, Ratish Puduppully, and Mitesh M Khapra. 2024 · 2024
Closest in time.
OLiVia-Nav: An Online Lifelong Vision Language Approach for Mobile Robot Social Navigation
Siddarth Narasimhan, Aaron Hao Tan, Daniel Choi, and Goldie Nejat. 2024 · 2024
Closest in time.
Scalable Language Model with Generalized Continual Learning. In ICLR
Bohao PENG, Zhuotao Tian, Shu Liu, Ming-Chang Yang, and Jiaya Jia. 2024 · 2024
Closest in time.
Gradient Projection For Parameter-Efficient Continual Learning
Jingyang Qiao, Zhizhong Zhang, Xin Tan, Yanyun Qu, Wensheng Zhang, and Yuan Xie. 2024 · 2024
Closest in time.
Just Say the Name: Online Continual Learning with Category Names Only via Data Generation
Minhyuk Seo, Diganta Misra, Seongwon Cho, Minjae Lee, and Jonghyun Choi. 2024 · 2024
Closest in time.
Continual learning of large language models: A comprehensive survey
Haizhou Shi, Zihao Xu, Hengyi Wang, Weiyi Qin, Wenyuan Wang, Yibin Wang, and Hao Wang. 2024 · 2024
Closest in time.
Longxiang Tang, Zhuotao Tian, Kai Li, Chunming He, Hantao Zhou, Hengshuang Zhao, Xiu Li, and Jiaya Jia. 2024 · 2024
Closest in time.
CLIP model is an Efficient Online Lifelong Learner
Leyuan Wang, Liuyu Xiang, Yujie Wei, Yunlong Wang, and Zhaofeng He. 2024a · 2024
Closest in time.
A comprehensive survey of continual learning: Theory, method and application
Liyuan Wang, Xingxing Zhang, Hang Su, and Jun Zhu. 2024c · 2024
Closest in time.
Continual learning for large language models: A survey
Tongtong Wu, Linhao Luo, Yuan-Fang Li, Shirui Pan, Thuy-Trang Vu, and Gholamreza Haffari. 2024 · 2024
Closest in time.
Yu-Chu Yu, Chi-Pin Huang, Jr-Jen Chen, Kai-Po Chang, Yung-Hsuan Lai, Fu-En Yang, and Yu-Chiang Frank Wang. 2024a · 2024
Closest in time.
COPR: Continual Human Preference Learning via Optimal Policy Regularization
Han Zhang, Lin Gui, Yu Lei, Yuanzhao Zhai, Yehong Zhang, Yulan He, Hui Wang, Yue Yu, Kam-Fai Wong, Bin Liang, et al · 2024
Closest in time.
Towards Lifelong Learning of Large Language Models: A Survey
Junhao Zheng, Shengjie Qiu, Chengming Shi, and Qianli Ma. 2024 · 2024
Closest in time.