Fetching the paper…
Reading the bibliography…
Efficient machine learning (ML) has become increasingly important as models grow larger and data volumes expand.
Multitask learning
Rich Caruana · 1997
Earlier work this paper cites.
A survey on large-scale machine learning
Meng Wang, Weijie Fu, Xiangnan He, Shijie Hao, and Xindong Wu · 2008
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-Łojasiewicz condition
Hessam Karimi, Julie Nutini, and Mark Schmidt · 2017
Earlier work this paper cites.
An overview of multi-task learning in deep neural networks
Sebastian Ruder · 2017
Earlier work this paper cites.
Multi-task learning for document ranking and query suggestion
Wasi Uddin Ahmad, Kai-Wei Chang, and Hongning Wang · 2018
Earlier work this paper cites.
Multi-task learning as multi-objective optimization
Ozan Sener and Vladimir Koltun · 2018
Earlier work this paper cites.
Bag-of-words transfer: Non-contextual techniques for multi-task learning
Seth Ebner, Felicity Wang, and Benjamin Van Durme · 2019
Earlier work this paper cites.
End-to-end multi-task learning with attention
Shikun Liu, Edward Johns, and Andrew J. Davison · 2019
Earlier work this paper cites.
Which tasks should be learned together in multi-task learning?
Trevor Darrell Standley, Amir R Zamir, Dahun Chen, Leonidas J Guibas, Jitendra Malik, Silvio Savarese, and Yuke Zhang · 2020
Cited alongside, same era.
Gradient surgery for multi-task learning
Tianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine, Karol Hausman, and Chelsea Finn · 2020
Cited alongside, same era.
Vilt: Vision-and-language transformer without convolution or region supervision
Wonjae Kim, Bokyung Son, and Ildoo Kim · 2021
Cited alongside, same era.
Conflict-averse gradient descent for multi-task learning, 2021
Bo Liu, Xingchao Liu, Xiaojie Jin, Peter Stone, and Qiang Liu · 2021
Cited alongside, same era.
Neural ranking models for document retrieval
Mohamed Trabelsi, Zhiyu Chen, Brian D. Davison, and Jeff Heflin · 2021
Cited alongside, same era.
Multi-task pre-training for plug-and-play task-oriented dialogue system
Yixuan Su, Lei Shu, Elman Mansimov, Arshit Gupta, Deng Cai, Yi-An Lai, and Yi Zhang · 2022
Later among the works it cites.
Efficient Methods for Natural Language Processing: A Survey
Marcos Treviso, Ji-Ung Lee, Tianchu Ji, Betty van Aken, Qingqing Cao, Manuel R. Ciosici, Michael Hassid, Kenneth Heafield, Sara Hooker, Colin Raffel, Pedro H. Martins, André F. T. Martins, Jessica Zosa Forde, Peter Milder, Edwin Simpson, Noam Slonim, Jesse Dodge, Emma Strubell, Niranjan Balasubramanian, Leon Derczynski, Iryna Gurevych, and Roy Schwartz · 2023
Later among the works it cites.
A survey of multi-task learning in natural language processing: Regarding task relatedness and training methods
Zhihan Zhang, Wenhao Yu, Mengxia Yu, Zhichun Guo, and Meng Jiang · 2023
Later among the works it cites.
Embedding in recommender systems: A survey
Xiangyu Zhao, Maolin Wang, Xinjian Zhao, Jiansheng Li, Shucheng Zhou, Dawei Yin, Qing Li, Jiliang Tang, and Ruocheng Guo · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yan Zeng and Jian-Yun Nie · 2021
Cited alongside, same era.
A survey on multi-task learning
Yu Zhang and Qiang Yang · 2021
Cited alongside, same era.
The bearable lightness of big data: Towards massive public datasets in scientific machine learning
Wai Tong Chung, Ki Sung Jung, Jacqueline H. Chen, and Matthias Ihme · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Cited alongside, same era.
Hao Ban and Kaiyi Ji · 2024
Closest in time.
Llmeasyquant – an easy to use toolkit for llm quantization, 2024
Dong Liu and Kaiser Pister · 2024
Closest in time.
Graphsnapshot: Graph machine learning acceleration with fast storage and retrieval, 2024
Dong Liu, Roger Waleffe, Meng Jiang, and Shivaram Venkataraman · 2024
Closest in time.
Densemtl: Cross-task attention mechanism for dense multi-task learning, 2024
Ivan Lopes, Tuan-Hung Vu, and Raoul de Charette · 2024
Closest in time.