Fetching the paper…
Reading the bibliography…
Dialogue policy training for composite tasks, such as restaurant reservation in multiple places, is a practically important and challenging problem.
Intra-option learning about temporally abstract actions
Richard S Sutton, Doina Precup, and Satinder P Singh. 1998 · 1998
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini. 2009 · 2009
Earlier work this paper cites.
Bayesian update of dialogue state: A pomdp framework for spoken dialogue systems
Blaise Thomson and Steve Young. 2010 · 2010
Earlier work this paper cites.
Gaussian processes for pomdp-based dialogue manager optimization
Milica Gašić and Steve Young. 2013 · 2013
Earlier work this paper cites.
Pomdp-based statistical spoken dialog systems: A review
Steve Young, Milica Gašić, Blaise Thomson, and Jason D Williams. 2013 · 2013
Earlier work this paper cites.
Policy committee for adaptation in multi-domain spoken dialogue systems
M Gašić, N Mrkšić, Pei-hao Su, David Vandyke, Tsung-Hsien Wen, and Steve Young. 2015 · 2015
Earlier work this paper cites.
End-to-end lstm-based dialog control optimized with supervised and reinforcement learning
Jason D Williams and Geoffrey Zweig. 2016 · 2016
Earlier work this paper cites.
Towards end-to-end learning for dialog state tracking and management using deep reinforcement learning
Tiancheng Zhao and Maxine Eskenazi. 2016 · 2016
Cited alongside, same era.
Sub-domain modelling for dialogue management with hierarchical reinforcement learning
Paweł Budzianowski, Stefan Ultes, Pei-Hao Su, Nikola Mrkšić, Tsung-Hsien Wen, Inigo Casanueva, Lina Rojas-Barahona, and Milica Gašić. 2017 · 2017
Cited alongside, same era.
Affordable on-line dialogue policy learning
Cheng Chang, Runzhe Yang, Lu Chen, Xiang Zhou, and Kai Yu. 2017 · 2017
Cited alongside, same era.
Agent-aware dropout dqn for safe and efficient on-line dialogue policy learning
Lu Chen, Xiang Zhou, Cheng Chang, Runzhe Yang, and Kai Yu. 2017 · 2017
Cited alongside, same era.
Dialogue manager domain adaptation using gaussian process reinforcement learning
Milica Gašić, Nikola Mrkšić, Lina M Rojas-Barahona, Pei-Hao Su, Stefan Ultes, David Vandyke, Tsung-Hsien Wen, and Steve Young. 2017 · 2017
Cited alongside, same era.
Composite task-completion dialogue policy learning via hierarchical deep reinforcement learning
Baolin Peng, Xiujun Li, Lihong Li, Jianfeng Gao, Asli Celikyilmaz, Sungjin Lee, and Kam-Fai Wong. 2017 · 2017
Later among the works it cites.
Pydial: A multi-domain statistical dialogue system toolkit
Stefan Ultes, Lina M Rojas Barahona, Pei-Hao Su, David Vandyke, Dongho Kim, Inigo Casanueva, Paweł Budzianowski, Nikola Mrkšić, Tsung-Hsien Wen, Milica Gasic, et al. 2017 · 2017
Later among the works it cites.
Hybrid code networks: practical and efficient end-to-end dialog control with supervised and reinforcement learning
Jason D Williams, Kavosh Asadi, and Geoffrey Zweig. 2017 · 2017
Later among the works it cites.
Structured dialogue policy with graph neural networks
Lu Chen, Bowen Tan, Sishan Long, and Kai Yu. 2018 · 2018
Later among the works it cites.
Adversarial advantage actor-critic model for task-completion dialogue policy learning
Baolin Peng, Xiujun Li, Jianfeng Gao, Jingjing Liu, Yun-Nung Chen, and Kam-Fai Wong. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
End-to-end task-completion neural dialogue systems
Xiujun Li, Yun-Nung Chen, Lihong Li, Jianfeng Gao, and Asli Celikyilmaz. 2017 · 2017
Cited alongside, same era.
Iterative policy learning in end-to-end trainable task-oriented neural dialog models
Bing Liu and Ian Lane. 2017 · 2017
Cited alongside, same era.
Sample-efficient actor-critic reinforcement learning with supervised data for dialogue management
Pei-Hao Su, Paweł Budzianowski, Stefan Ultes, Milica Gašic, and Steve Young. ????
Cited in the paper.
Later among the works it cites.
Subgoal discovery for hierarchical dialogue policy learning
Da Tang, Xiujun Li, Jianfeng Gao, Chong Wang, Lihong Li, and Tony Jebara. 2018 · 2018
Later among the works it cites.
Nervenet: Learning structured policy with graph neural networks
Tingwu Wang, Renjie Liao, Jimmy Ba, and Sanja Fidler. 2018 · 2018
Later among the works it cites.