Fetching the paper…
Reading the bibliography…
E-commerce platforms usually display a mixed list of ads and organic items in feed.
Introduction to reinforcement learning . Vol. 135
Richard S Sutton, Andrew G Barto, et al · 1998
Earlier work this paper cites.
Constrained Markov decision processes . Vol. 7
Eitan Altman. 1999 · 1999
Earlier work this paper cites.
Dueling network architectures for deep reinforcement learning. In International conference on machine learning . PMLR, 1995–2003
Ziyu Wang, Tom Schaul, Matteo Hessel, Hado Hasselt, Marc Lanctot, and Nando Freitas. 2016 · 2003
Earlier work this paper cites.
Internet advertising and the generalized second-price auction: Selling billions of dollars worth of keywords
Benjamin Edelman, Michael Ostrovsky, and Michael Schwarz. 2007 · 2007
Earlier work this paper cites.
An Empirical Analysis of Search Engine Advertising: Sponsored Search in Electronic Markets
A. Ghose and Sha Yang. 2009 · 2009
Earlier work this paper cites.
Learning to Advertise: How Many Ads Are Enough?. In PAKDD
B. Wang, Zhaonan Li, Jie Tang, Kuo Zhang, Songcan Chen, and Liyun Ru. 2011 · 2011
Earlier work this paper cites.
Online Matching and Ad Allocation
Aranyak Mehta. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Optimal advertisement allocation in online social media feeds. In Proceedings of the 8th ACM International Workshop on Hot Topics in Planet-scale mObile computing and online Social neTworking . 43–48
Iordanis Koutsopoulos. 2016 · 2016
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Learning to Collaborate: Multi-Scenario Ranking via Multi-Agent Reinforcement Learning
Jun Feng, H. Li, Minlie Huang, Shichen Liu, Wenwu Ou, Zhirong Wang, and Xiaoyan Zhu. 2018 · 2018
Cited alongside, same era.
The whole-page optimization via dynamic ad allocation. In Companion Proceedings of the The Web Conference . 1407–1411
Weiru Zhang, Chao Wei, Xiaonan Meng, Yi Hu, and Hao Wang. 2018 · 2018
Cited alongside, same era.
Impression Allocation for Combating Fraud in E-commerce Via Deep Reinforcement Learning with Action Norm Penalty. In IJCAI
Generator and Critic: A Deep Reinforcement Learning Approach for Slate Re-ranking in E-commerce
Jianxiong Wei, Anxiang Zeng, Yueqiu Wu, Pengxin Guo, Q. Hua, and Qingpeng Cai. 2020 · 2020
Later among the works it cites.
Ads Allocation in Feed via Constrained Optimization. In Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 3386–3394
Jinyun Yan, Zhiyuan Xu, Birjodh Tiwana, and Shaunak Chatterjee. 2020 · 2020
Later among the works it cites.
Jointly learning to recommend and advertise. In Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 3319–3327
Xiangyu Zhao, Xudong Zheng, Xiwang Yang, Xiaobing Liu, and Jiliang Tang. 2020 · 2020
Later among the works it cites.
Blending Advertising with Organic Content in E-Commerce: A Virtual Bids Optimization Approach
Carlos Carrion, Zenan Wang, Harikesh Nair, Xianghong Luo, Yulin Lei, Xiliang Lin, Wenlong Chen, Qiyu Hu, Changping Peng, Yongjun Bao, and Weipeng P. Yan. 2021 · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mengchen Zhao, Z. Li, Bo An, Haifeng Lu, Yifan Yang, and Chen Chu. 2018 · 2018
Cited alongside, same era.
Deep interest network for click-through rate prediction. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 1059–1068
Guorui Zhou, Xiaoqiang Zhu, Chenru Song, Ying Fan, Han Zhu, Xiao Ma, Yanghui Yan, Junqi Jin, Han Li, and Kun Gai. 2018 · 2018
Cited alongside, same era.
Learning Adaptive Display Exposure for Real-Time Advertising. In Proceedings of the 28th ACM International Conference on Information and Knowledge Management . 2595–2603
Weixun Wang, Junqi Jin, Jianye Hao, Chunjie Chen, Chuan Yu, Weinan Zhang, Jun Wang, Xiaotian Hao, Yixi Wang, Han Li, et al · 2019
Cited alongside, same era.
Deep Time-Aware Item Evolution Network for Click-Through Rate Prediction. In Proceedings of the 29th ACM International Conference on Information & Knowledge Management . 785–794
Xiang Li, Chao Wang, Bin Tong, Jiwei Tan, Xiaoyi Zeng, and Tao Zhuang. 2020 · 2020
Cited alongside, same era.
MiNet: Mixed Interest Network for Cross-Domain Click-Through Rate Prediction. In Proceedings of the 29th ACM International Conference on Information & Knowledge Management . 2669–2676
Wentao Ouyang, Xiuwu Zhang, Lei Zhao, Jinmei Luo, Yu Zhang, Heng Zou, Zhaojie Liu, and Yanlong Du. 2020 · 2020
Cited alongside, same era.
Revisit Recommender System in the Permutation Prospective
Yufei Feng, Yu Gong, Fei Sun, Qingwen Liu, and Wenwu Ou. 2021a · 2021
Closest in time.
GRN: Generative Rerank Network for Context-wise Recommendation
Yufei Feng, Binbin Hu, Yu Gong, Fei Sun, Qingwen Liu, and Wenwu Ou. 2021b · 2021
Closest in time.
Hierarchical Reinforcement Learning for Integrated Recommendation. In Proceedings of AAAI
Ruobing Xie, Shaoliang Zhang, Rui Wang, Feng Xia, and Leyu Lin. 2021 · 2021
Closest in time.
DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender Systems. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 35. 750–758
Xiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang, Xiaobing Liu, Hui Liu, and Jiliang Tang. 2021 · 2021
Closest in time.