Fetching the paper…
Reading the bibliography…
Unified models capable of solving a wide variety of tasks have gained traction in vision and NLP due to their ability to share regularities and structures across tasks, which improves individual task performance and reduces computational footprint.
N. Reimers and I. Gurevych, “Sentence-bert: Sentence embeddings using siamese bert-networks,” in
1908
Earlier work this paper cites.
1910
Earlier work this paper cites.
J. S. Vitter, “Random sampling with a reservoir,”
1985
Earlier work this paper cites.
A. Robins, “Catastrophic forgetting in neural networks: the role of rehearsal mechanisms,” in
1993
Earlier work this paper cites.
R. Caruana, “Multitask learning,”
1997
Earlier work this paper cites.
S. Thrun, “Lifelong learning algorithms.”
1998
Earlier work this paper cites.
S. M. Barnett and S. J. Ceci, “When and where do we apply what we learn?: A taxonomy for far transfer.”
2002
Earlier work this paper cites.
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert, “icarl: Incremental classifier and representation learning,” in
2010
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in
2012
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
Y. Teh, V. Bapst, W. M. Czarnecki, J. Quan, J. Kirkpatrick, R. Hadsell, N. Heess, and R. Pascanu, “Distral: Robust multitask reinforcement learning,”
2017
Earlier work this paper cites.
G. I. Parisi, J. Tani, C. Weber, and S. Wermter, “Lifelong learning of human actions with deep neural network self-organization,”
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,”
2017
Earlier work this paper cites.
J. D. Power and B. L. Schlaggar, “Neural plasticity across the lifespan,”
2017
Earlier work this paper cites.
Z. Li and D. Hoiem, “Learning without forgetting,”
2017
Earlier work this paper cites.
W. Ying, Y. Zhang, J. Huang, and Q. Yang, “Transfer learning via learning to transfer,” in
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Z. Chen, V. Badrinarayanan, C.-Y. Lee, and A. Rabinovich, “Gradnorm: Gradient normalization for adaptive loss balancing in deep multitask networks,” in
2018
Earlier work this paper cites.
J. Mendez, S. Shivkumar, and E. Eaton, “Lifelong inverse reinforcement learning,”
2018
Earlier work this paper cites.
R. T. Chen, Y. Rubanova, J. Bettencourt, and D. K. Duvenaud, “Neural ordinary differential equations,”
2018
Cited alongside, same era.
L. Dong, N. Yang, W. Wang, F. Wei, X. Liu, Y. Wang, J. Gao, M. Zhou, and H.-W. Hon, “Unified language model pre-training for natural language understanding and generation,”
2019
Cited alongside, same era.
W. M. Czarnecki, R. Pascanu, S. Osindero, S. Jayakumar, G. Swirszcz, and M. Jaderberg, “Distilling policy distillation,” in
2019
Cited alongside, same era.
2019
Cited alongside, same era.
S. Sodhani, L. Denoyer, P.-A. Kamienny, and O. Delalleau, “Mtenv - environment interface for mulit-task reinforcement learning,” Github, 2021. [Online]. Available:
2021
Later among the works it cites.
2021
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
G. I. Parisi, R. Kemker, J. L. Part, C. Kanan, and S. Wermter, “Continual lifelong learning with neural networks: A review,”
2019
Cited alongside, same era.
H. Bao, L. Dong, F. Wei, W. Wang, N. Yang, X. Liu, Y. Wang, J. Gao, S. Piao, M. Zhou,
2020
Cited alongside, same era.
R. Jena, C. Liu, and K. Sycara, “Augmenting gail with bc for sample efficient imitation learning,”
2020
Cited alongside, same era.
T. Yu, S. Kumar, A. Gupta, S. Levine, K. Hausman, and C. Finn, “Gradient surgery for multi-task learning,”
2020
Cited alongside, same era.
N. Vithayathil Varghese and Q. H. Mahmoud, “A survey of multi-task deep reinforcement learning,”
2020
Cited alongside, same era.
R. Julian, B. Swanson, G. S. Sukhatme, S. Levine, C. Finn, and K. Hausman, “Efficient adaptation for end-to-end vision-based robotic manipulation,” in
2020
Cited alongside, same era.
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet, “Learning latent plans from play,” in
2020
Cited alongside, same era.
2022
Later among the works it cites.
A. Kolesnikov, A. Susano Pinto, L. Beyer, X. Zhai, J. Harmsen, and N. Houlsby, “Uvim: A unified modeling approach for vision with learned guiding codes,”
2022
Later among the works it cites.
I. Uchendu, T. Xiao, Y. Lu, B. Zhu, M. Yan, J. Simon, M. Bennice, C. Fu, C. Ma, J. Jiao,
2022
Later among the works it cites.
2022
Later among the works it cites.
H. Hoshino, K. Ota, A. Kanezaki, and R. Yokota, “Opirl: Sample efficient off-policy inverse reinforcement learning via distribution matching,” in
2022
Later among the works it cites.
C. Xu and J. McAuley, “A survey on model compression for natural language processing,”
2022
Later among the works it cites.
S. Cohen, B. Amos, M. P. Deisenroth, M. Henaff, E. Vinitsky, and D. Yarats, “Imitation learning from pixel observations for continuous control,” 2022. [Online]. Available:
2022
Later among the works it cites.
A. Xie and C. Finn, “Lifelong robotic reinforcement learning by retaining experiences,” in
2022
Later among the works it cites.
2023
Closest in time.
M. Shridhar, L. Manuelli, and D. Fox, “Perceiver-actor: A multi-task transformer for robotic manipulation,” in
2023
Closest in time.
2023
Closest in time.
S. Haldar, V. Mathur, D. Yarats, and L. Pinto, “Watch and match: Supercharging imitation with regularized optimal transport,” in
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
P. Wu, A. Majumdar, K. Stone, Y. Lin, I. Mordatch, P. Abbeel, and A. Rajeswaran, “Masked trajectory models for prediction, representation, and control,” in
2023
Closest in time.
2023
Closest in time.
S. Auddy, J. Hollenstein, M. Saveriano, A. Rodríguez-Sánchez, and J. Piater, “Continual learning from demonstration of robotics skills,”
2023
Closest in time.