Fetching the paper…
Reading the bibliography…
Building a dialogue agent to fulfill complex tasks, such as travel planning, is challenging because the agent has to learn to collectively complete multiple subtasks.
Reinforcement learning with a hierarchy of abstract models
Satinder P. Singh. 1992 · 1992
Earlier work this paper cites.
Reinforcement learning with hierarchies of machines
Ronald Parr and Stuart J. Russell. 1997 · 1997
Earlier work this paper cites.
Intra-option learning about temporally abstract actions
Richard S. Sutton, Doina Precup, and Satinder P. Singh. 1998 · 1998
Earlier work this paper cites.
Between MDPs and Semi-MDPs: A framework for temporal abstraction in reinforcement learning
Richard S. Sutton, Doina Precup, and Satinder P. Singh. 1999 · 1999
Earlier work this paper cites.
Hierarchical reinforcement learning with the MAXQ value function decomposition
Thomas G. Dietterich. 2000 · 2000
Earlier work this paper cites.
A stochastic model of human-machine interaction for learning dialog strategies
Esther Levin, Roberto Pieraccini, and Wieland Eckert. 2000 · 2000
Earlier work this paper cites.
Probabilistic simulation of human-machine dialogues
Konrad Scheffler and Steve J. Young. 2000 · 2000
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
Andrew G. Barto and Sridhar Mahadevan. 2003 · 2003
Earlier work this paper cites.
Agenda-based user simulation for bootstrapping a POMDP dialogue system
Jost Schatzmann, Blaise Thomson, Karl Weilhammer, Hui Ye, and Steve J. Young. 2007 · 2007
Earlier work this paper cites.
Hierarchical reinforcement learning for spoken dialogue systems
Heriberto Cuayáhuitl. 2009 · 2009
Earlier work this paper cites.
The hidden agenda user simulation model
Jost Schatzmann and Steve Young. 2009 · 2009
Cited alongside, same era.
Multi-policy dialogue management
Pierre Lison. 2011 · 2011
Cited alongside, same era.
POMDP-based statistical spoken dialog systems: A review
Steve J. Young, Milica Gasic, Blaise Thomson, and Jason D. Williams. 2013 · 2013
Cited alongside, same era.
Spoken language understanding using long short-term memory neural networks
Kaisheng Yao, Baolin Peng, Yu Zhang, Dong Yu, Geoffrey Zweig, and Yangyang Shi. 2014 · 2014
Cited alongside, same era.
Distributed dialogue policies for multi-domain statistical dialogue management
Milica Gasic, Dongho Kim, Pirros Tsiakoulis, and Steve J. Young. 2015a · 2015
Cited alongside, same era.
Policy committee for adaptation in multi-domain spoken dialogue systems
Deep reinforcement learning for multi-domain dialogue systems
Heriberto Cuayáhuitl, Seunghak Yu, Ashley Williamson, and Jacob Carse. 2016 · 2016
Later among the works it cites.
Multi-domain joint semantic frame parsing using bi-directional RNN-LSTM
Dilek Hakkani-Tür, Gokhan Tur, Asli Celikyilmaz, Yun-Nung Chen, Jianfeng Gao, Li Deng, and Ye-Yi Wang. 2016 · 2016
Later among the works it cites.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
Tejas D. Kulkarni, Karthik Narasimhan, Ardavan Saeedi, and Josh Tenenbaum. 2016 · 2016
Later among the works it cites.
A user simulator for task-completion dialogues
Xiujun Li, Zachary C Lipton, Bhuwan Dhingra, Lihong Li, Jianfeng Gao, and Yun-Nung Chen. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Milica Gasic, Nikola Mrksic, Pei-hao Su, David Vandyke, Tsung-Hsien Wen, and Steve J. Young. 2015b · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin A. Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis. 2015 · 2015
Cited alongside, same era.
Semantically conditioned LSTM-based natural language generation for spoken dialogue systems
Tsung-Hsien Wen, Milica Gasic, Nikola Mrksic, Pei-hao Su, David Vandyke, and Steve J. Young. 2015 · 2015
Cited alongside, same era.
A sequence-to-sequence model for user simulation in spoken dialogue systems
Layla El Asri, Jing He, and Kaheer Suleman. 2016 · 2016
Cited alongside, same era.
End-to-end task-completion neural dialogue systems
Xiujun Li, Yun-Nung Chen, Lihong Li, and Jianfeng Gao. 2017a
Cited in the paper.
Investigation of language understanding impact for reinforcement learning based dialogue systems
Xiujun Li, Yun-Nung Chen, Lihong Li, Jianfeng Gao, and Asli Celikyilmaz. 2017b
Cited in the paper.
Pei-Hao Su, Milica Gasic, Nikola Mrksic, Lina Rojas-Barahona, Stefan Ultes, David Vandyke, Tsung-Hsien Wen, and Steve Young. 2016 · 2016
Later among the works it cites.
SimpleDS: A simple deep reinforcement learning dialogue system
Heriberto Cuayáhuitl. 2017 · 2017
Closest in time.
End-to-end reinforcement learning of dialogue agents for information access
Bhuwan Dhingra, Lihong Li, Xiujun Li, Jianfeng Gao, Yun-Nung Chen, Faisal Ahmed, and Li Deng. 2017 · 2017
Closest in time.
Frames: A corpus for adding memory to goal-oriented dialogue systems
Layla El Asri, Hannes Schulz, Shikhar Sharma, Jeremie Zumer, Justin Harris, Emery Fine, Rahul Mehrotra, and Kaheer Suleman. 2017 · 2017
Closest in time.
Hybrid code networks: Practical and efficient end-to-end dialog control with supervised and reinforcement learning
Jason D Williams, Kavosh Asadi, and Geoffrey Zweig. 2017 · 2017
Closest in time.