Fetching the paper…
Reading the bibliography…
Reinforcement Learning (RL)-based recommender systems (RSs) have garnered considerable attention due to their ability to learn optimal recommendation policies and maximize long-term user rewards.
Asynchronous methods for deep reinforcement learning. In International conference on machine learning . PMLR, 1928–1937
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
The cross-entropy method: a unified approach to combinatorial optimization, Monte-Carlo simulation, and machine learning . Vol. 133
Reuven Y Rubinstein and Dirk P Kroese. 2004 · 2004
Earlier work this paper cites.
Reinforcement Learning based Control of Traffic Lights in Non-stationary Environments: A Case Study in a Microscopic Simulator.. In EUMAS
Denise de Oliveira, Ana LC Bazzan, Bruno Castro da Silva, Eduardo W Basso, Luis Nunes, Rosaldo Rossetti, Eugénio de Oliveira, Roberto da Silva, and Luis Lamb. 2006 · 2006
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra. 2015 · 2015
Earlier work this paper cites.
Wide & deep learning for recommender systems. In Proceedings of the 1st workshop on deep learning for recommender systems . 7–10
Heng-Tze Cheng, Levent Koc, Jeremiah Harmsen, Tal Shaked, Tushar Chandra, Hrishi Aradhye, Glen Anderson, Greg Corrado, Wei Chai, Mustafa Ispir, et al · 2016
Earlier work this paper cites.
When recurrent neural networks meet the neighborhood for session-based recommendation. In Proceedings of the eleventh ACM conference on recommender systems . 306–310
Dietmar Jannach and Malte Ludewig. 2017 · 2017
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction. In International conference on machine learning . PMLR, 2778–2787
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell. 2017 · 2017
Earlier work this paper cites.
Deep reinforcement learning for list-wise recommendations
Xiangyu Zhao, Liang Zhang, Long Xia, Zhuoye Ding, Dawei Yin, and Jiliang Tang. 2017 · 2017
Earlier work this paper cites.
Learning a deep listwise context model for ranking refinement. In The 41st international ACM SIGIR conference on research & development in information retrieval . 135–144
Qingyao Ai, Keping Bi, Jiafeng Guo, and W Bruce Croft. 2018 · 2018
Earlier work this paper cites.
Addressing function approximation error in actor-critic methods. In International conference on machine learning . PMLR, 1587–1596
Scott Fujimoto, Herke Hoof, and David Meger. 2018 · 2018
Earlier work this paper cites.
Beyond greedy ranking: Slate optimization via list-CVAE
Ray Jiang, Sven Gowal, Timothy A Mann, and Danilo J Rezende. 2018 · 2018
Earlier work this paper cites.
Self-attentive sequential recommendation. In 2018 IEEE international conference on data mining (ICDM) . IEEE, 197–206
Wang-Cheng Kang and Julian McAuley. 2018 · 2018
Earlier work this paper cites.
David Rohde, Stephen Bonner, Travis Dunlop, Flavian Vasile, and Alexandros Karatzoglou. 2018 · 2018
Earlier work this paper cites.
Personalized top-n sequential recommendation via convolutional sequence embedding. In Proceedings of the eleventh ACM international conference on web search and data mining . 565–573
Jiaxi Tang and Ke Wang. 2018 · 2018
Earlier work this paper cites.
Supervised reinforcement learning with recurrent neural network for dynamic treatment recommendation. In Proceedings of the 24th ACM SIGKDD international conference on knowledge discovery & data mining . 2447–2456
Lu Wang, Wei Zhang, Xiaofeng He, and Hongyuan Zha. 2018 · 2018
Earlier work this paper cites.
Recsim: A configurable simulation platform for recommender systems
Eugene Ie, Chih-wei Hsu, Martin Mladenov, Vihan Jain, Sanmit Narvekar, Jing Wang, Rui Wu, and Craig Boutilier. 2019 · 2019
Earlier work this paper cites.
Continual lifelong learning with neural networks: A review
German I Parisi, Ronald Kemker, Jose L Part, Christopher Kanan, and Stefan Wermter. 2019 · 2019
Earlier work this paper cites.
Virtual-taobao: Virtualizing real-world online retail environment for reinforcement learning. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 4902–4909
Jing-Cheng Shi, Yang Yu, Qing Da, Shi-Yong Chen, and An-Xiang Zeng. 2019 · 2019
Earlier work this paper cites.
BERT4Rec: Sequential recommendation with bidirectional encoder representations from transformer. In Proceedings of the 28th ACM international conference on information and knowledge management . 1441–1450
Fei Sun, Jun Liu, Jian Wu, Changhua Pei, Xiao Lin, Wenwu Ou, and Peng Jiang. 2019 · 2019
Earlier work this paper cites.
" Deep reinforcement learning for search, recommendation, and online advertising: a survey" by Xiangyu Zhao, Long Xia, Jiliang Tang, and Dawei Yin with Martin Vesely as coordinator
Xiangyu Zhao, Long Xia, Jiliang Tang, and Dawei Yin. 2019a · 2019
Cited alongside, same era.
Toward simulating environments in reinforcement learning based recommendations
Xiangyu Zhao, Long Xia, Lixin Zou, Dawei Yin, and Jiliang Tang. 2019b · 2019
Cited alongside, same era.
Reinforcement learning to optimize long-term user engagement in recommender systems. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 2810–2818
Lixin Zou, Long Xia, Zhuoye Ding, Jiaxing Song, Weidong Liu, and Dawei Yin. 2019 · 2019
Cited alongside, same era.
Keeping dataset biases out of the simulation: A debiased simulator for reinforcement learning based recommender systems. In Proceedings of the 14th ACM Conference on Recommender Systems . 190–199
Jin Huang, Harrie Oosterhuis, Maarten De Rijke, and Herke Van Hoof. 2020 · 2020
Cited alongside, same era.
KuaiRand: An Unbiased Sequential Recommendation Dataset with Randomly Exposed Videos. In Proceedings of the 31st ACM International Conference on Information & Knowledge Management . 3953–3957
Chongming Gao, Shijun Li, Yuan Zhang, Jiawei Chen, Biao Li, Wenqiang Lei, Peng Jiang, and Xiangnan He. 2022 · 2022
Later among the works it cites.
Personalized recommendation system based on knowledge embedding and historical behavior
Bei Hui, Lizong Zhang, Xue Zhou, Xiao Wen, and Yuhui Nian. 2022 · 2022
Later among the works it cites.
An overview of and recommendations for more accessible digital mental health services
Emily G Lattie, Colleen Stiles-Shields, and Andrea K Graham. 2022 · 2022
Later among the works it cites.
MLP4Rec: A pure MLP architecture for sequential recommendations
Muyang Li, Xiangyu Zhao, Chuan Lyu, Minghao Zhao, Runze Wu, and Ruocheng Guo. 2022c · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Social attentive deep Q-networks for recommender systems
Yu Lei, Zhitao Wang, Wenjie Li, Hongbin Pei, and Quanyu Dai. 2020 · 2020
Cited alongside, same era.
Rl-cyclegan: Reinforcement learning aware simulation-to-real. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11157–11166
Kanishka Rao, Chris Harris, Alex Irpan, Sergey Levine, Julian Ibarz, and Mohi Khansari. 2020 · 2020
Cited alongside, same era.
Self-supervised reinforcement learning for recommender systems. In Proceedings of the 43rd International ACM SIGIR conference on research and development in Information Retrieval . 931–940
Xin Xin, Alexandros Karatzoglou, Ioannis Arapakis, and Joemon M Jose. 2020 · 2020
Cited alongside, same era.
Xiaocong Chen, Lina Yao, Julian McAuley, Guanglin Zhou, and Xianzhi Wang. 2021 · 2021
Cited alongside, same era.
A deep reinforcement learning based long-term recommender system
Liwei Huang, Mingsheng Fu, Fan Li, Hong Qu, Yangjun Liu, and Wenyu Chen. 2021 · 2021
Cited alongside, same era.
Advances in collaborative filtering
Yehuda Koren, Steffen Rendle, and Robert Bell. 2021 · 2021
Cited alongside, same era.
Variation control and evaluation for generative slate recommendations. In Proceedings of the Web Conference 2021 . 436–448
Shuchang Liu, Fei Sun, Yingqiang Ge, Changhua Pei, and Yongfeng Zhang. 2021 · 2021
Cited alongside, same era.
Crossing the reality gap: A survey on sim-to-real transferability of robot controllers in reinforcement learning
Erica Salvato, Gianfranco Fenu, Eric Medvet, and Felice Andrea Pellegrino. 2021 · 2021
Cited alongside, same era.
Yunqi Li, Hanxiong Chen, Shuyuan Xu, Yingqiang Ge, Juntao Tan, Shuchang Liu, and Yongfeng Zhang. 2022a · 2022
Later among the works it cites.
Reinforcement learning over sentiment-augmented knowledge graphs towards accurate and explainable recommendation. In Proceedings of the Fifteenth ACM International Conference on Web Search and Data Mining . 784–793
Sung-Jun Park, Dong-Kyu Chae, Hong-Kyun Bae, Sumin Park, and Sang-Wook Kim. 2022 · 2022
Later among the works it cites.
Supervised advantage actor-critic for recommender systems. In Proceedings of the Fifteenth ACM International Conference on Web Search and Data Mining . 1186–1196
Xin Xin, Alexandros Karatzoglou, Ioannis Arapakis, and Joemon M Jose. 2022 · 2022
Later among the works it cites.
PrefRec: Preference-based Recommender Systems for Reinforcing Long-term User Engagement
Wanqi Xue, Qingpeng Cai, Zhenghai Xue, Shuo Sun, Shuchang Liu, Dong Zheng, Peng Jiang, and Bo An. 2022a · 2022
Later among the works it cites.
ResAct: Reinforcing Long-term Engagement in Sequential Recommendation with Residual Actor
Wanqi Xue, Qingpeng Cai, Ruohan Zhan, Dong Zheng, Peng Jiang, and Bo An. 2022b · 2022
Later among the works it cites.
MAE4Rec: Storage-saving Transformer for Sequential Recommendations. In Proceedings of the 31st ACM International Conference on Information & Knowledge Management . 2681–2690
Kesen Zhao, Xiangyu Zhao, Zijian Zhang, and Muyang Li. 2022 · 2022
Later among the works it cites.
Reinforcing User Retention in a Billion Scale Short Video Recommender System. In Companion Proceedings of the ACM Web Conference 2023 . 421–426
Qingpeng Cai, Shuchang Liu, Xueliang Wang, Tianyou Zuo, Wentao Xie, Bin Yang, Dong Zheng, Peng Jiang, and Kun Gai. 2023a · 2023
Closest in time.
Two-Stage Constrained Actor-Critic for Short Video Recommendation. In Proceedings of the ACM Web Conference 2023 . 865–875
Qingpeng Cai, Zhenghai Xue, Chi Zhang, Wanqi Xue, Shuchang Liu, Ruohan Zhan, Xueliang Wang, Tianyou Zuo, Wentao Xie, Dong Zheng, et al · 2023
Closest in time.
STRec: Sparse Transformer for Sequential Recommendations. In Proceedings of the 17th ACM Conference on Recommender Systems . 101–111
Chengxi Li, Yejing Wang, Qidong Liu, Xiangyu Zhao, Wanyu Wang, Yiqi Wang, Lixin Zou, Wenqi Fan, and Qing Li. 2023 · 2023
Closest in time.
MMMLP: Multi-modal Multilayer Perceptron for Sequential Recommendations. In Proceedings of the ACM Web Conference 2023 . 1109–1117
Jiahao Liang, Xiangyu Zhao, Muyang Li, Zijian Zhang, Wanyu Wang, Haochen Liu, and Zitao Liu. 2023 · 2023
Closest in time.
Exploration and Regularization of the Latent Action Space in Recommendation
Shuchang Liu, Qingpeng Cai, Bowen Sun, Yuhao Wang, Ji Jiang, Dong Zheng, Kun Gai, Peng Jiang, Xiangyu Zhao, and Yongfeng Zhang. 2023a · 2023
Closest in time.
Multi-Task Recommendations with Reinforcement Learning. In Proceedings of the ACM Web Conference 2023 . 1273–1282
Ziru Liu, Jiejie Tian, Qingpeng Cai, Xiangyu Zhao, Jingtong Gao, Shuchang Liu, Dayou Chen, Tonghao He, Dong Zheng, Peng Jiang, et al · 2023
Closest in time.
A survey on the fairness of recommender systems
Yifan Wang, Weizhi Ma, Min Zhang, Yiqun Liu, and Shaoping Ma. 2023 · 2023
Closest in time.
User Retention-oriented Recommendation with Decision Transformer. In Proceedings of the ACM Web Conference 2023 . 1141–1149
Kesen Zhao, Lixin Zou, Xiangyu Zhao, Maolin Wang, and Dawei Yin. 2023 · 2023
Closest in time.