Fetching the paper…
Reading the bibliography…
In session-based or sequential recommendation, it is important to consider a number of factors like long-term user engagement, multiple types of user-item interactions such as clicks, purchases etc.
Dynamic programming
Richard Bellman. 1966 · 1966
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Actor-critic algorithms. In Advances in neural information processing systems . 1008–1014
Vijay R Konda and John N Tsitsiklis. 2000 · 2000
Earlier work this paper cites.
Improving recommendation diversity. In Proceedings of the Twelfth Irish Conference on Artificial Intelligence and Cognitive Science, Maynooth, Ireland . Citeseer, 85–94
Keith Bradley and Barry Smyth. 2001 · 2001
Earlier work this paper cites.
Cumulated gain-based evaluation of IR techniques
Kalervo Järvelin and Jaana Kekäläinen. 2002 · 2002
Earlier work this paper cites.
An MDP-based recommender system
Guy Shani, David Heckerman, and Ronen I Brafman. 2005 · 2005
Earlier work this paper cites.
Matrix factorization techniques for recommender systems
Yehuda Koren, Robert Bell, and Chris Volinsky. 2009 · 2009
Earlier work this paper cites.
BPR: Bayesian personalized ranking from implicit feedback. In Proceedings of the twenty-fifth conference on uncertainty in artificial intelligence . AUAI Press, 452–461
Steffen Rendle, Christoph Freudenthaler, Zeno Gantner, and Lars Schmidt-Thieme. 2009 · 2009
Earlier work this paper cites.
Double Q-learning. In Advances in Neural Information Processing Systems . 2613–2621
Hado V Hasselt. 2010 · 2010
Earlier work this paper cites.
A contextual-bandit approach to personalized news article recommendation. In Proceedings of the 19th international conference on World wide web . ACM, 661–670
Lihong Li, Wei Chu, John Langford, and Robert E Schapire. 2010 · 2010
Earlier work this paper cites.
Factorization machines. In 2010 IEEE International Conference on Data Mining . IEEE, 995–1000
Steffen Rendle. 2010 · 2010
Earlier work this paper cites.
Factorizing personalized markov chains for next-basket recommendation. In Proceedings of the 19th international conference on World wide web . ACM, 811–820
Steffen Rendle, Christoph Freudenthaler, and Lars Schmidt-Thieme. 2010 · 2010
Earlier work this paper cites.
Unbiased offline evaluation of contextual-bandit-based news article recommendation algorithms. In Proceedings of the fourth ACM international conference on Web search and data mining . ACM, 297–306
Lihong Li, Wei Chu, John Langford, and Xuanhui Wang. 2011 · 2011
Earlier work this paper cites.
Where you like to go next: Successive point-of-interest recommendation. In Twenty-Third international joint conference on Artificial Intelligence
Chen Cheng, Haiqin Yang, Michael R Lyu, and Irwin King. 2013 · 2013
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart Van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Generative adversarial nets. In Advances in neural information processing systems . 2672–2680
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Session-based recommendations with recurrent neural networks
Balázs Hidasi, Alexandros Karatzoglou, Linas Baltrunas, and Domonkos Tikk. 2015 · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Vista: A visually, socially, and temporally-aware model for artistic recommendation. In Proceedings of the 10th ACM Conference on Recommender Systems . ACM, 309–316
Reinforcement learning to rank in e-commerce search engine: Formalization, analysis, and application. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . ACM, 368–377
Yujing Hu, Qing Da, Anxiang Zeng, Yang Yu, and Yinghui Xu. 2018 · 2018
Later among the works it cites.
Self-attentive sequential recommendation. In 2018 IEEE International Conference on Data Mining (ICDM) . IEEE, 197–206
Wang-Cheng Kang and Julian McAuley. 2018 · 2018
Later among the works it cites.
David Rohde, Stephen Bonner, Travis Dunlop, Flavian Vasile, and Alexandros Karatzoglou. 2018 · 2018
Later among the works it cites.
Personalized top-n sequential recommendation via convolutional sequence embedding. In Proceedings of the Eleventh ACM International Conference on Web Search and Data Mining . ACM, 565–573
Jiaxi Tang and Ke Wang. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ruining He, Chen Fang, Zhaowen Wang, and Julian McAuley. 2016 · 2016
Cited alongside, same era.
Fusing similarity models with markov chains for sparse sequential recommendation. In 2016 IEEE 16th International Conference on Data Mining (ICDM) . IEEE, 191–200
Ruining He and Julian McAuley. 2016 · 2016
Cited alongside, same era.
General factorization framework for context-aware recommendations
Balázs Hidasi and Domonkos Tikk. 2016 · 2016
Cited alongside, same era.
Generative adversarial imitation learning. In Advances in neural information processing systems . 4565–4573
Jonathan Ho and Stefano Ermon. 2016 · 2016
Cited alongside, same era.
Model-free imitation learning with policy optimization. In International Conference on Machine Learning . 2760–2769
Jonathan Ho, Jayesh Gupta, and Stefano Ermon. 2016 · 2016
Cited alongside, same era.
Safe and efficient off-policy reinforcement learning. In Advances in Neural Information Processing Systems . 1054–1062
Rémi Munos, Tom Stepleton, Anna Harutyunyan, and Marc Bellemare. 2016 · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
Attention is all you need. In Advances in neural information processing systems . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Faraz Torabi, Garrett Warnell, and Peter Stone. 2018 · 2018
Later among the works it cites.
fBGD: Learning embeddings from positive unlabeled data with BGD
Fajie Yuan, Xin Xin, Xiangnan He, Guibing Guo, Weinan Zhang, Chua Tat-Seng, and Joemon M Jose. 2018 · 2018
Later among the works it cites.
Recommendations with negative feedback via pairwise deep reinforcement learning. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . ACM, 1040–1048
Xiangyu Zhao, Liang Zhang, Zhuoye Ding, Long Xia, Jiliang Tang, and Dawei Yin. 2018 · 2018
Later among the works it cites.
Exact-K Recommendation via Maximal Clique Optimization
Yu Gong, Yu Zhu, Lu Duan, Qingwen Liu, Ziyu Guan, Fei Sun, Wenwu Ou, and Kenny Q Zhu. 2019 · 2019
Later among the works it cites.
SlateQ: A tractable decomposition for reinforcement learning with recommendation sets
Eugene Ie, Vihan Jain, Jing Wang, Sanmit Narvekar, Ritesh Agarwal, Rui Wu, Heng-Tze Cheng, Tushar Chandra, and Craig Boutilier. 2019 · 2019
Later among the works it cites.
Stabilizing Transformers for Reinforcement Learning
Emilio Parisotto, H Francis Song, Jack W Rae, Razvan Pascanu, Caglar Gulcehre, Siddhant M Jayakumar, Max Jaderberg, Raphael Lopez Kaufman, Aidan Clark, Seb Noury, et al · 2019
Later among the works it cites.
Environment Reconstruction with Hidden Confounders for Reinforcement Learning based Recommendation. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . ACM, 566–576
Wenjie Shang, Yang Yu, Qingyang Li, Zhiwei Qin, Yiping Meng, and Jieping Ye. 2019 · 2019
Later among the works it cites.
Virtual-taobao: Virtualizing real-world online retail environment for reinforcement learning. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 4902–4909
Jing-Cheng Shi, Yang Yu, Qing Da, Shi-Yong Chen, and An-Xiang Zeng. 2019 · 2019
Later among the works it cites.
A Simple Convolutional Generative Network for Next Item Recommendation. In Proceedings of the Twelfth ACM International Conference on Web Search and Data Mining . ACM, 582–590
Fajie Yuan, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M Jose, and Xiangnan He. 2019 · 2019
Later among the works it cites.
Reinforcement Learning to Optimize Long-term User Engagement in Recommender Systems
Lixin Zou, Long Xia, Zhuoye Ding, Jiaxing Song, Weidong Liu, and Dawei Yin. 2019 · 2019
Later among the works it cites.