Fetching the paper…
Reading the bibliography…
Despite widespread interests in reinforcement-learning for task-oriented dialogue systems, several obstacles can frustrate research and development progress.
User modeling for spoken dialogue system evaluation
Wieland Eckert, Esther Levin, and Roberto Pieraccini · 1997
Earlier work this paper cites.
A stochastic model of human-machine interaction for learning dialog strategies
Esther Levin, Roberto Pieraccini, and Wieland Eckert · 2000
Earlier work this paper cites.
Automatic learning of dialogue strategy using dialogue simulation and reinforcement learning
Konrad Scheffler and Steve Young · 2002
Earlier work this paper cites.
Automatic design of spoken dialogue systems
Konrad Haarhoff Scheffler · 2003
Earlier work this paper cites.
Human-computer dialogue simulation using hidden markov models
Heriberto Cuayáhuitl, Steve Renals, Oliver Lemon, and Hiroshi Shimodaira · 2005
Earlier work this paper cites.
Learning user simulations for information state update dialogue systems
Kallirroi Georgila, James Henderson, and Oliver Lemon · 2005
Earlier work this paper cites.
Learning more effective dialogue strategies using limited dialogue move features
Matthew Frampton and Oliver Lemon · 2006
Earlier work this paper cites.
Consistent goal-directed user model for realisitc man-machine task-oriented spoken dialogue simulation
Olivier Pietquin · 2006
Earlier work this paper cites.
A probabilistic framework for dialog simulation and optimal strategy learning
Olivier Pietquin and Thierry Dutoit · 2006
Earlier work this paper cites.
A survey of statistical user simulation techniques for reinforcement-learning of dialogue management strategies
Jost Schatzmann, Karl Weilhammer, Matt Stuttle, and Steve Young · 2006
Cited alongside, same era.
Error simulation for training statistical dialogue systems
Jost Schatzmann, Blaise Thomson, and Steve Young · 2007
Cited alongside, same era.
Data-driven user simulation for automated evaluation of spoken dialog systems
Sangkeun Jung, Cheongjae Lee, Kyungduk Kim, Minwoo Jeong, and Gary Geunbae Lee · 2009
Cited alongside, same era.
The hidden agenda user simulation model
Jost Schatzmann and Steve Young · 2009
Cited alongside, same era.
A survey on metrics for the evaluation of user simulations
Olivier Pietquin and Helen Hastie · 2013
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
A sequence-to-sequence model for user simulation in spoken dialogue systems
Layla El Asri, Jing He, and Kaheer Suleman · 2016
Closest in time.
End-to-end reinforcement learning of dialogue agents for information access
Bhuwan Dhingra, Lihong Li, Xiujun Li, Jianfeng Gao, Yun-Nung Chen, Faisal Ahmed, and Li Deng · 2016
Closest in time.
Multi-domain joint semantic frame parsing using bi-directional rnn-lstm
Dilek Hakkani-Tür, Gokhan Tur, Asli Celikyilmaz, Yun-Nung Chen, Jianfeng Gao, Li Deng, and Ye-Yi Wang · 2016
Closest in time.
Efficient exploration for dialogue policy learning with BBQ networks & replay buffer spiking
Zachary C Lipton, Jianfeng Gao, Lihong Li, Xiujun Li, Faisal Ahmed, and Li Deng · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Cited alongside, same era.
Semantically conditioned lstm-based natural language generation for spoken dialogue systems
Tsung-Hsien Wen, Milica Gašić, Nikola Mrkšić, Pei-Hao Su, David Vandyke, and Steve Young · 2015
Cited alongside, same era.
Pei-Hao Su, Milica Gasic, Nikola Mrksic, Lina Rojas-Barahona, Stefan Ultes, David Vandyke, Tsung-Hsien Wen, and Steve Young · 2016
Closest in time.
Conditional generation and snapshot learning in neural dialogue systems
Tsung-Hsien Wen, Milica Gašić, Nikola Mrkšić, Lina M. Rojas-Barahona, Pei-Hao Su, Stefan Ultes, David Vandyke, and Steve Young · 2016
Closest in time.
End-to-end lstm-based dialog control optimized with supervised and reinforcement learning
Jason D Williams and Geoffrey Zweig · 2016
Closest in time.
Tiancheng Zhao and Maxine Eskenazi · 2016
Closest in time.