Fetching the paper…
Reading the bibliography…
A personalized conversational sales agent could have much commercial potential.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 1998 · 1998
Earlier work this paper cites.
An MDP-based recommender system
Guy Shani, David Heckerman, and Ronen I Brafman. 2005 · 2005
Earlier work this paper cites.
An experimental comparison of click position-bias models. In Proceedings of the 2008 international conference on web search and data mining
Nick Craswell, Onno Zoeter, Michael Taylor, and Bill Ramsey. 2008 · 2008
Earlier work this paper cites.
Personalized interactive faceted search. In Proceedings of the 17th international conference on World Wide Web
Jonathan Koren, Yi Zhang, and Xue Liu. 2008 · 2008
Earlier work this paper cites.
Introduction to Information Retrieval
Christopher D. Manning, Prabhakar Raghavan, and Hinrich Schütze. 2008 · 2008
Earlier work this paper cites.
Probabilistic matrix factorization. In Advances in neural information processing systems
Andriy Mnih and Ruslan R Salakhutdinov. 2008 · 2008
Earlier work this paper cites.
Matrix factorization techniques for recommender systems
Yehuda Koren, Robert Bell, and Chris Volinsky. 2009 · 2009
Earlier work this paper cites.
Factorization machines. In Data Mining (ICDM), 2010 IEEE 10th International Conference on
Steffen Rendle. 2010 · 2010
Earlier work this paper cites.
Interactive retrieval based on faceted feedback. In Proceedings of the 33rd international ACM SIGIR conference on Research and development in information retrieval
Lanbo Zhang and Yi Zhang. 2010 · 2010
Earlier work this paper cites.
Content-based recommender systems: State of the art and trends
Pasquale Lops, Marco De Gemmis, and Giovanni Semeraro. 2011 · 2011
Earlier work this paper cites.
N-best error simulation for training spoken dialogue systems. In Spoken Language Technology Workshop (SLT), 2012 IEEE
Blaise Thomson, Milica Gasic, Matthew Henderson, Pirros Tsiakoulis, and Steve Young. 2012 · 2012
Cited alongside, same era.
Dialog State Tracking Challenge 2 & 3
Blaise Thomson Matthew Henderson and Jason Williams. 2013 · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller. 2013 · 2013
Cited alongside, same era.
The dialog state tracking challenge. In Proceedings of the SIGDIAL 2013 Conference
Jason Williams, Antoine Raux, Deepak Ramachandran, and Alan Black. 2013 · 2013
Cited alongside, same era.
Pomdp-based statistical spoken dialog systems: A review
Steve Young, Milica Gašić, Blaise Thomson, and Jason D Williams. 2013 · 2013
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
A network-based end-to-end trainable task-oriented dialogue system
Tsung-Hsien Wen, David Vandyke, Nikola Mrksic, Milica Gasic, Lina M Rojas-Barahona, Pei-Hao Su, Stefan Ultes, and Steve Young. 2016 · 2016
Later among the works it cites.
Tiancheng Zhao and Maxine Eskenazi. 2016 · 2016
Later among the works it cites.
Real-Time Bidding by Reinforcement Learning in Display Advertising. In Proceedings of the Tenth ACM International Conference on Web Search and Data Mining
Han Cai, Kan Ren, Weinan Zhang, Kleanthis Malialis, Jun Wang, Yong Yu, and Defeng Guo. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan. 2015 · 2015
Cited alongside, same era.
Using recurrent neural networks for slot filling in spoken language understanding
Grégoire Mesnil, Yann Dauphin, Kaisheng Yao, Yoshua Bengio, Li Deng, Dilek Hakkani-Tur, Xiaodong He, Larry Heck, Gokhan Tur, Dong Yu, et al · 2015
Cited alongside, same era.
Oriol Vinyals and Quoc Le. 2015 · 2015
Cited alongside, same era.
Learning end-to-end goal-oriented dialog
Antoine Bordes and Jason Weston. 2016 · 2016
Cited alongside, same era.
Towards Conversational Recommender Systems.. In KDD
Konstantina Christakopoulou, Filip Radlinski, and Katja Hofmann. 2016 · 2016
Cited alongside, same era.
Improving information extraction by acquiring external evidence with reinforcement learning
Karthik Narasimhan, Adam Yala, and Regina Barzilay. 2016 · 2016
Cited alongside, same era.
Towards End-to-End Reinforcement Learning of Dialogue Agents for Information Access. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Bhuwan Dhingra, Lihong Li, Xiujun Li, Jianfeng Gao, Yun-Nung Chen, Faisal Ahmed, and Li Deng. 2017 · 2017
Later among the works it cites.
Investigation of Language Understanding Impact for Reinforcement Learning Based Dialogue Systems
Xiujun Li, Yun-Nung Chen, Lihong Li, Jianfeng Gao, and Asli Celikyilmaz. 2017 · 2017
Later among the works it cites.
Task-oriented query reformulation with reinforcement learning
Rodrigo Nogueira and Kyunghyun Cho. 2017 · 2017
Later among the works it cites.
Mastering the game of Go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Later among the works it cites.
Dynamic Facet Ordering for Faceted Product Search Engines
Damir Vandic, Steven Aanen, Flavius Frasincar, and Uzay Kaymak. 2017 · 2017
Later among the works it cites.
A probabilistic framework for representing dialog systems and entropy-based dialog management through dynamic stochastic state evolution
Ji Wu, Miao Li, and Chin-Hui Lee. 2015 · 2035
Closest in time.