Fetching the paper…
Reading the bibliography…
In this paper, we present the practical problems and the lessons learned at short-video services from Kuaishou.
Adaptive Mixtures of Local Experts
Robert A. Jacobs, Michael I. Jordan, Steven J. Nowlan, and Geoffrey E. Hinton. 1991 · 1991
Earlier work this paper cites.
Multi-Task Learning for Stock Selection. In Advances in Neural Information Processing Systems (NeurIPS)
Joumana Ghosn and Yoshua Bengio. 1996 · 1996
Earlier work this paper cites.
Multitask Learning
Rich Caruana. 1997 · 1997
Earlier work this paper cites.
A Unified Architecture for Natural Language Processing: Deep Neural Networks with Multitask Learning. In International Conference on Machine Learning (ICML)
Ronan Collobert and Jason Weston. 2008 · 2008
Earlier work this paper cites.
The YouTube Video Recommendation System. In ACM Conference on Recommender Systems (RecSys)
James Davidson, Benjamin Liebald, Junning Liu, Palash Nandy, Taylor Van Vleet, Ullas Gargi, Sujoy Gupta, Yu He, Mike Lambert, Blake Livingston, and Dasarathi Sampath. 2010 · 2010
Earlier work this paper cites.
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. In International Conference on Machine Learning (ICML)
Sergey Ioffe and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Layer Normalization
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016 · 2016
Earlier work this paper cites.
Ask the GRU: Multi-task Learning for Deep Text Recommendations. In ACM Conference on Recommender Systems (RecSys)
Trapit Bansal, David Belanger, and Andrew McCallum. 2016 · 2016
Earlier work this paper cites.
Deep Neural Networks for YouTube Recommendations. In ACM Conference on Recommender Systems (RecSys)
Paul Covington, Jay Adams, and Emre Sargin. 2016 · 2016
Earlier work this paper cites.
Cross-Stitch Networks for Multi-Task Learning. In IEEE/CVF Computer Vision and Pattern Recognition Conference (CVPR)
Ishan Misra, Abhinav Shrivastava, Abhinav Gupta, and Martial Hebert. 2016 · 2016
Earlier work this paper cites.
HD-MTL: Hierarchical Deep Multi-Task Learning for Large-Scale Visual Recognition
Jianping Fan, Tianyi Zhao, Zhenzhong Kuang, Yu Zheng, Ji Zhang, Jun Yu, and Jinye Peng. 2017 · 2017
Earlier work this paper cites.
Traffic Sign Recognition via Multi-Modal Tree-Structure Embedded Multi-Task Learning
Xiao Lu, Yaonan Wang, Xuanyu Zhou, Zhenjun Zhang, and Zhigang Ling. 2017 · 2017
Earlier work this paper cites.
Searching for Activation Functions
Prajit Ramachandran, Barret Zoph, and Quoc V Le. 2017 · 2017
Earlier work this paper cites.
Sluice Networks: Learning What to Share Between Loosely Related Tasks
Sebastian Ruder, Joachim Bingel, Isabelle Augenstein, and Anders Søgaard. 2017 · 2017
Cited alongside, same era.
Multi-Task Learning Using Uncertainty to Weigh Losses for Scene Geometry and Semantics. In IEEE/CVF Computer Vision and Pattern Recognition Conference (CVPR)
Alex Kendall, Yarin Gal, and Roberto Cipolla. 2018 · 2018
Cited alongside, same era.
Deep Interest Network for Click-Through Rate Prediction. In ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD)
Guorui Zhou, Xiaoqiang Zhu, Chenru Song, Ying Fan, Han Zhu, Xiao Ma, Yanghui Yan, Junqi Jin, Han Li, and Kun Gai. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Multi-Task Learning of Hierarchical Vision-Language Representation. In IEEE/CVF Computer Vision and Pattern Recognition Conference (CVPR)
Deep Feedback Network for Recommendation. In International Joint Conference on Artificial Intelligence (IJCAI)
Ruobing Xie, Cheng Ling, Yalong Wang, Rui Wang, Feng Xia, and Leyu Lin. 2021 · 2021
Later among the works it cites.
Positive, Negative and Neutral: Modeling Implicit Feedback in Session-based News Recommendation. In International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR)
Shansan Gong and Kenny Q Zhu. 2022 · 2022
Later among the works it cites.
A Survey on Multi-Task Learning
Yu Zhang and Qiang Yang. 2022 · 2022
Later among the works it cites.
AdaTT: Adaptive Task-to-Task Fusion Network for Multitask Learning in Recommendations. In ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD)
Danwei Li, Zhengyu Zhang, Siyang Yuan, Mingze Gao, Weilin Zhang, Chaofei Yang, Xi Liu, and Jiyan Yang. 2023 · 2023
Later among the works it cites.
Deep Task-specific Bottom Representation Network for Multi-Task Recommendation. In ACM International Conference on Information and Knowledge Management (CIKM)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Duy-Kien Nguyen and Takayuki Okatani. 2019 · 2019
Cited alongside, same era.
A Hierarchical Multi-task Approach for Learning Embeddings from Semantic Tasks. In AAAI Conference on Artificial Intelligence (AAAI)
Victor Sanh, Thomas Wolf, and Sebastian Ruder. 2019 · 2019
Cited alongside, same era.
Recommending What Video to Watch Next: A Multitask Ranking System. In ACM Conference on Recommender Systems (RecSys)
Zhe Zhao, Lichan Hong, Li Wei, Jilin Chen, Aniruddh Nath, Shawn Andrews, Aditee Kumthekar, Maheswaran Sathiamoorthy, Xinyang Yi, and Ed Chi. 2019 · 2019
Cited alongside, same era.
Deep Interest Evolution Network for Click-Through Rate Prediction. In AAAI Conference on Artificial Intelligence (AAAI)
Guorui Zhou, Na Mou, Ying Fan, Qi Pi, Weijie Bian, Chang Zhou, Xiaoqiang Zhu, and Kun Gai. 2019 · 2019
Cited alongside, same era.
Search-based User Interest Modeling with Lifelong Sequential Behavior Data for Click-Through Rate Prediction. In ACM International Conference on Information and Knowledge Management (CIKM)
Qi Pi, Guorui Zhou, Yujing Zhang, Zhe Wang, Lejian Ren, Ying Fan, Xiaoqiang Zhu, and Kun Gai. 2020 · 2020
Cited alongside, same era.
Progressive Layered Extraction (PLE): A Novel Multi-Task Learning (MTL) Model for Personalized Recommendations. In ACM Conference on Recommender Systems (RecSys)
Hongyan Tang, Junning Liu, Ming Zhao, and Xudong Gong. 2020 · 2020
Cited alongside, same era.
MSSM: A Multiple-level Sparse Sharing Model for Efficient Multi-Task Learning. In International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR)
Ke Ding, Xin Dong, Yong He, Lei Cheng, Chilin Fu, Zhaoxin Huan, Hai Li, Tan Yan, Liang Zhang, Xiaolu Zhang, et al · 2021
Cited alongside, same era.
LoRA: Low-Rank Adaptation of Large Language Models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
Qi Liu, Zhilong Zhou, Gangwei Jiang, Tiezheng Ge, and Defu Lian. 2023 · 2023
Later among the works it cites.
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Damai Dai, Chengqi Deng, Chenggang Zhao, RX Xu, Huazuo Gao, Deli Chen, Jiashi Li, Wangding Zeng, Xingkai Yu, Y Wu, et al · 2024
Closest in time.
Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models
Yixin Liu, Kai Zhang, Yuan Li, Zhiling Yan, Chujie Gao, Ruoxi Chen, Zhengqing Yuan, Yue Huang, Hanchi Sun, Jianfeng Gao, et al · 2024
Closest in time.
STEM: Unleashing the Power of Embeddings for Multi-task Recommendation. In AAAI Conference on Artificial Intelligence (AAAI)
Liangcai Su, Junwei Pan, Ximei Wang, Xi Xiao, Shijie Quan, Xihua Chen, and Jie Jiang. 2024 · 2024
Closest in time.
Trinity: Syncretizing Multi-/Long-tail/Long-term Interests All in One
Jing Yan, Liu Jiang, Jianfei Cui, Zhichen Zhao, Xingyan Bin, Feng Zhang, and Zuotao Liu. 2024 · 2024
Closest in time.
Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations
Jiaqi Zhai, Lucy Liao, Xing Liu, Yueming Wang, Rui Li, Xuan Cao, Leon Gao, Zhaojie Gong, Fangda Gu, Michael He, et al · 2024
Closest in time.
Multi-Behavior Collaborative Filtering with Partial Order Graph Convolutional Networks
Yijie Zhang, Yuanchen Bei, Hao Chen, Qijie Shen, Zheng Yuan, Huan Gong, Senzhang Wang, Feiran Huang, and Xiao Huang. 2024a · 2024
Closest in time.
SpeechLM: Enhanced Speech Pre-Training With Unpaired Textual Data
Ziqiang Zhang, Sanyuan Chen, Long Zhou, Yu Wu, Shuo Ren, Shujie Liu, Zhuoyuan Yao, Xun Gong, Lirong Dai, Jinyu Li, and Furu Wei. 2024b · 2024
Closest in time.
Exploring Training on Heterogeneous Data with Mixture of Low-rank Adapters
Yuhang Zhou, Zihua Zhao, Haolin Li, Siyuan Du, Jiangchao Yao, Ya Zhang, and Yanfeng Wang. 2024 · 2024
Closest in time.