Fetching the paper…
Reading the bibliography…
The development of recommender systems that optimize multi-turn interaction with users, and model the interactions of different agents (e.g., users, content providers, vendors) in the recommender ecosystem have drawn increasing attention in recent years.
Latent Variable Session-Based Recommendation. (2019)
Stephen Bonner and David Rohde. 2019 · 1904
Earlier work this paper cites.
Challenges of Real-World Reinforcement Learning. (2019)
Gabriel Dulac-Arnold, Daniel Mankowitz, and Todd Hester. 2019 · 1904
Earlier work this paper cites.
Eugene Ie, Vihan Jain, Jing Wang, Sanmit Navrekar, Ritesh Agarwal, Rui Wu, Heng-Tze Cheng, Morgane Lustman, Vincent Gatto, Paul Covington, Jim McFadden, Tushar Chandra, and Craig Boutilier. 2019b · 1905
Earlier work this paper cites.
Toward Simulating Environments in Reinforcement Learning Based Recommendations. (2019)
Xiangyu Zhao, Long Xia, Zhuoye Ding, Dawei Yin, and Jiliang Tang. 2019 · 1906
Earlier work this paper cites.
RecSim: A Configurable Simulation Platform for Recommender Systems. (2019)
Eugene Ie, Chih wei Hsu, Martin Mladenov, Vihan Jain, Sanmit Narvekar, Jing Wang, Rui Wu, and Craig Boutilier. 2019c · 1909
Earlier work this paper cites.
A Model for Reasoning about Persistence and Causation
Thomas Dean and Keiji Kanazawa. 1989 · 1989
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning. In Machine Learning
Ronald J. Williams. 1992 · 1992
Earlier work this paper cites.
GroupLens: Applying Collaborative Filtering to Usenet News
Joseph A. Konstan, Bradley N. Miller, David Maltz, Jonathan L. Herlocker, Lee R. Gordon, and John Riedl. 1997 · 1997
Earlier work this paper cites.
Object-oriented Bayesian Networks. In Proceedings of the Thirteenth Conference on Uncertainty in Artificial Intelligence (UAI-97)
Avi Pfeffer and Daphne Koller. 1997 · 1997
Earlier work this paper cites.
Empirical Analysis of Predictive Algorithms for Collaborative Filtering. In Proceedings of the 14th Conference on Uncertainty in Artificial Intelligence
Jack S. Breese, David Heckerman, and Carl Kadie. 1998 · 1998
Earlier work this paper cites.
Monitoring a Complex Physical System using a Hybrid Dynamic Bayes Net. In Proceedings of the Eighteenth Conference on Uncertainty in Artificial Intelligence (UAI-02)
Uri Lerner, Brooks Moses, Maricia Scott, Sheila McIlraith, and Daphne Koller. 2002 · 2002
Earlier work this paper cites.
WES: Agent-based User Interaction Simulation on Real Infrastructure
John Ahlgren, Maria Eugenia Berezin, Kinga Bojarczuk, Elena Dulskyte, Inna Dvortsova, Johann George, Natalija Gucevska, Mark Harman, Ralf Lämmel, Erik Meijer, Silvia Sapora, and Justin Spahr-Summers. 2020 · 2004
Earlier work this paper cites.
The AI Economist: Improving Equality and Productivity with AI-Driven Tax Policies
Stephan Zheng, Alexander Trott, Sunil Srinivasa, Nikhil Naik, Melvin Gruesbeck, David C. Parkes, and Richard Socher. 2020 · 2004
Earlier work this paper cites.
An MDP-based Recommender System
Guy Shani, David Heckerman, and Ronen I. Brafman. 2005 · 2005
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Christopher M Bishop. 2006 · 2006
Earlier work this paper cites.
Probabilistic Matrix Factorization. In Advances in Neural Information Processing Systems 20 (NIPS-07)
Ruslan Salakhutdinov and Andriy Mnih. 2007 · 2007
Earlier work this paper cites.
Usage-based Web Recommendations: A Reinforcement Learning Approach. In Proceedings of the First ACM Conference on Recommender Systems (RecSys07)
Nima Taghipour, Ahmad Kardan, and Saeed Shiry Ghidary. 2007 · 2007
Earlier work this paper cites.
MCMC using Hamiltonian dynamics
Radford Neal. 2011 · 2011
Earlier work this paper cites.
A Hidden Markov Model for Collaborative Filtering
Nachiketa Sahoo, Param Vir Singh, and Tridas Mukhopadhyay. 2012 · 2012
Earlier work this paper cites.
The Arcade Learning Environment: An Evaluation Platform for General Agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling. 2013 · 2013
Earlier work this paper cites.
Deep Content-based Music Recommendation. In Advances in Neural Information Processing Systems 26 (NIPS-13)
Aaron van den Oord, Sander Dieleman, and Benjamin Schrauwen. 2013 · 2013
Earlier work this paper cites.
Online Clustering of Bandits. In Proceedings of the Thirty-first International Conference on Machine Learning (ICML-14)
Claudio Gentile, Shuai Li, and Giovanni Zappella. 2014 · 2014
Cited alongside, same era.
Latent Bandits. In Proceedings of the Thirty-first International Conference on Machine Learning (ICML-14)
Odalric-Ambrym Maillard and Shie Mannor. 2014 · 2014
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Cited alongside, same era.
Towards Conversational Recommender Systems. In Proceedings of the 22Nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining
Konstantina Christakopoulou, Filip Radlinski, and Katja Hofmann. 2016 · 2016
Cited alongside, same era.
Deep Neural Networks for YouTube Recommendations. In Proceedings of the 10th ACM Conference on Recommender Systems
Horizon: Facebook’s Open Source Applied Reinforcement Learning Platform. (2018)
Jason Gauci, Edoardo Conti, Yitao Liang, Kittipat Virochsiri, Yuchen He, Zachary Kaden, Vivek Narayanan, and Xiaohui Ye. 2018 · 2018
Later among the works it cites.
Deep Dyna-Q: Integrating Planning for Task-Completion Dialogue Policy Learning. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Baolin Peng, Xiujun Li, Jianfeng Gao, Jingjing Liu, and Kam-Fai Wong. 2018 · 2018
Later among the works it cites.
David Rohde, Stephen Bonner, Travis Dunlop, Flavian Vasile, and Alexandros Karatzoglou. 2018 · 2018
Later among the works it cites.
Fairness of Exposure in Rankings. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining
Ashudeep Singh and Thorsten Joachims. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Paul Covington, Jay Adams, and Emre Sargin. 2016 · 2016
Cited alongside, same era.
The MovieLens Datasets: History and Context
F. Maxwell Harper and Joseph A. Konstan. 2016 · 2016
Cited alongside, same era.
Collaborative Filtering Bandits. In Proceedings of the 39th International ACM SIGIR Conference on Research and Development in Information Retrieval
Shuai Li, Alexandros Karatzoglou, and Claudio Gentile. 2016 · 2016
Cited alongside, same era.
Improved Recurrent Neural Networks for Session-based Recommendations. In Proceedings of the 1st Workshop on Deep Learning for Recommender Systems
Yong Kiam Tan, Xinxing Xu, and Yong Liu. 2016 · 2016
Cited alongside, same era.
Neural Collaborative Filtering. In Proceedings of the 26th International Conference on World Wide Web (WWW-17)
Xiangnan He, Lizi Liao, Hanwang Zhang, Liqiang Nie, Xia Hu, and Tat-Seng Chua. 2017 · 2017
Cited alongside, same era.
ELF: An Extensive, Lightweight and Flexible Research Platform for Real-time Strategy Games. In Advances in Neural Information Processing Systems 30 (NIPS-17)
Yuandong Tian, Qucheng Gong, Wenling Shang, Yuxin Wu, and C. Lawrence Zitnick. 2017 · 2017
Cited alongside, same era.
Deep probabilistic programming
Dustin Tran, Matthew D Hoffman, Rif A Saurous, Eugene Brevdo, Kevin Murphy, and David M Blei. 2017 · 2017
Cited alongside, same era.
Recurrent Recommender Networks. In Proceedings of the Tenth ACM International Conference on Web Search and Data Mining (WSDM-17)
Chao-Yuan Wu, Amr Ahmed, Alex Beutel, Alexander J. Smola, and How Jing. 2017 · 2017
Cited alongside, same era.
Conversational Recommender System. (2018)
Yueming Sun and Yi Zhang. 2018 · 2018
Later among the works it cites.
Simple, Distributed, and Accelerated Probabilistic Programming. In Advances in Neural Information Processing Systems 31 (NeurIPS-18)
Dustin Tran, Matthew D. Hoffman, Dave Moore, Christopher Suter, Srinivas Vasudevan, Alexey Radul, Matthew Johnson, and Rif A. Saurous. 2018 · 2018
Later among the works it cites.
AirDialogue: An Environment for Goal-oriented Dialogue Research. In Proceedings of 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP-18)
Wei Wei, Quoc Le, Andrew Dai, and Jia Li. 2018 · 2018
Later among the works it cites.
Deep Reinforcement Learning for Page-wise Recommendations. In Proceedings of the 12th ACM Conference on Recommender Systems (RecSys-18)
Xiangyu Zhao, Long Xia, Liang Zhang, Zhuoye Ding, Dawei Yin, and Jiliang Tang. 2018 · 2018
Later among the works it cites.
From Recommendation Systems to Facility Location Games. In Proceedings of the Thirty-third AAAI Conference on Artificial Intelligence (AAAI-19)
Omer Ben-Porat, Gregory Goren, Itay Rosenberg, and Moshe Tennenholtz. 2019 · 2019
Later among the works it cites.
Pyro: Deep universal probabilistic programming
Eli Bingham, Jonathan P Chen, Martin Jankowiak, Fritz Obermeyer, Neeraj Pradhan, Theofanis Karaletsos, Rohit Singh, Paul Szerlip, Paul Horsfall, and Noah D Goodman. 2019 · 2019
Later among the works it cites.
Building Personalized Simulator for Interactive Search. In Proceedings of the Twenty-eighth International Joint Conference on Artificial Intelligence (IJCAI-19)
Qianlong Liu, Baoliang Cui, Zhongyu Wei, Baolin Peng, Haikuan Huang, Hongbo Deng, Jianye Hao, Xuanjing Huang, and Kam-Fai Wong. 2019 · 2019
Later among the works it cites.
Advantage Amplification in Slowly Evolving Latent-State Environments. In International Joint Conference on Artifical Intelligence (IJCAI)
Martin Mladenov, Ofer Meshi, Jayden Ooi, Dale Schuurmans, and Craig Boutilier. 2019 · 2019
Later among the works it cites.
Virtual-Taobao: Virtualizing Real-world Online Retail Environment for Reinforcement Learning. In Proceedings of the Thirty-third AAAI Conference on Artificial Intelligence (AAAI-19)
Jing-Cheng Shi, Yang Yu, Qing Da, Shi-Yong Chen, and An-Xiang Zeng. 2019 · 2019
Later among the works it cites.
Bayesian layers: A module for neural network uncertainty. In Advances in Neural Information Processing Systems
Dustin Tran, Mike Dusenberry, Mark van der Wilk, and Danijar Hafner. 2019 · 2019
Later among the works it cites.
Differentiable Meta-Learning of Bandit Policies. In Advances in Neural Information Processing Systems 33 (NeurIPS-20)
Craig Boutilier, Chih-Wei Hsu, Branislav Kveton, Martin Mladenov, Csaba Szepesvári, and Manzil Zaheer. 2020 · 2020
Later among the works it cites.
Modeling Agents with Probabilistic Programs
Owain Evans, Andreas Stuhlmüller, John Salvatier, and Daniel Filan. 2017 · 2020
Later among the works it cites.
The Design and Implementation of Probabilistic Programming Languages
Noah D Goodman and Andreas Stuhlmüller. 2014 · 2020
Later among the works it cites.
Optimizing Long-term Social Welfare in Recommender Systems: A Constrained Matching Approach. In Proceedings of the Thirty-seventh International Conference on Machine Learning (ICML-20)
Martin Mladenov, Elliot Creager, Kevin Swerksy, Omer Ben-Porat, Richard S. Zemel, and Craig Boutilier. 2020 · 2020
Later among the works it cites.
Optimization of Large-Scale Agent-Based Simulations Through Automated Abstraction and Simplification. In 21st International Workshop on Multi-Agent-Based Simulation (MABS-20)
Alexey Tregubov and Jim Blythe. 2020 · 2020
Later among the works it cites.
Mixed Negative Sampling for Learning Two-tower Neural Networks in Recommendations. In Proceedings of the Web Conference (WWW-20)
Ji Yang, Xinyang Yi, Derek Zhiyuan Cheng, Lichan Hong, Yang Li, Simon Xiaoming Wang, Taibai Xu, and Ed H. Chi. 2020 · 2020
Later among the works it cites.
Measuring Recommender System Effects with Simulated Users. In 2nd Workshop on Fairness, Accountability, Transparency, Ethics and Society on the Web (FATES-20)
Siriu Yao, Yoni Halpern, Nithum Thain, Xuezhi Wang, Kang Lee, Flavien Prost, Ed H. Chi, Jilin Chen, and Alex Beutel. 2020 · 2020
Later among the works it cites.