Fetching the paper…
Reading the bibliography…
User Simulators are one of the major tools that enable offline training of task-oriented dialogue systems.
User modeling for spoken dialogue system evaluation
Wieland Eckert, Esther Levin, and Roberto Pieraccini. 1997 · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Spoken dialogue management using probabilistic reasoning
Nicholas Roy, Joelle Pineau, and Sebastian Thrun. 2000 · 2000
Earlier work this paper cites.
Probabilistic simulation of human-machine dialogues
Konrad Scheffler and Steve Young. 2000 · 2000
Earlier work this paper cites.
Corpus-based dialogue simulation for automatic strategy learning and evaluation
Konrad Scheffler and Steve Young. 2001 · 2001
Earlier work this paper cites.
Human-computer dialogue simulation using hidden markov models
Heriberto Cuayáhuitl, Steve Renals, Oliver Lemon, and Hiroshi Shimodaira. 2005 · 2005
Earlier work this paper cites.
Learning user simulations for information state update dialogue systems
Kallirroi Georgila, James Henderson, and Oliver Lemon. 2005 · 2005
Earlier work this paper cites.
Effects of the user model on simulation-based learning of dialogue strategies
Jost Schatztmann, Matthew N Stuttle, Karl Weilhammer, and Steve Young. 2005 · 2005
Earlier work this paper cites.
A probabilistic framework for dialog simulation and optimal strategy learning
Olivier Pietquin and Thierry Dutoit. 2006 · 2006
Earlier work this paper cites.
Agenda-based user simulation for bootstrapping a pomdp dialogue system
Jost Schatzmann, Blaise Thomson, Karl Weilhammer, Hui Ye, and Steve Young. 2007 · 2007
Earlier work this paper cites.
Partially observable markov decision processes for spoken dialog systems
Jason D Williams and Steve Young. 2007 · 2007
Earlier work this paper cites.
Evaluating user simulations with the cramér–von mises divergence
Jason D Williams. 2008 · 2008
Earlier work this paper cites.
Which words are hard to recognize? prosodic, lexical, and disfluency factors that increase speech recognition error rates
Sharon Goldwater, Dan Jurafsky, and Christopher D Manning. 2010 · 2010
Earlier work this paper cites.
User simulation in dialogue systems using inverse reinforcement learning
Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefevre, and Olivier Pietquin. 2011 · 2011
Earlier work this paper cites.
On-line policy optimisation of spoken dialogue systems via live interaction with human subjects
M. Gašić, F. Jurčíček, B. Thomson, K. Yu, and S. Young. 2011 · 2011
Cited alongside, same era.
Real user evaluation of spoken dialogue systems using amazon mechanical turk
Filip Jurčíček, Simon Keizer, Milica Gašić, Francois Mairesse, Blaise Thomson, Kai Yu, and Steve Young. 2011 · 2011
Cited alongside, same era.
Pomdp-based statistical spoken dialog systems: A review
Steve Young, Milica Gašić, Blaise Thomson, and Jason D Williams. 2013 · 2013
Cited alongside, same era.
Gaussian processes for pomdp-based dialogue manager optimization
Milica Gašić and Steve Young. 2014 · 2014
Cited alongside, same era.
The second dialog state tracking challenge
Matthew Henderson, Blaise Thomson, and Jason D Williams. 2014 · 2014
Cited alongside, same era.
The sjtu system for dialog state tracking challenge 2
A benchmarking environment for reinforcement learning based task oriented dialogue management
Iñigo Casanueva, Paweł Budzianowski, Pei-Hao Su, Nikola Mrkšić, Tsung-Hsien Wen, Stefan Ultes, Lina Rojas-Barahona, Steve Young, and Milica Gašić. 2017 · 2017
Later among the works it cites.
Affordable on-line dialogue policy learning
Cheng Chang, Runzhe Yang, Lu Chen, Xiang Zhou, and Kai Yu. 2017 · 2017
Later among the works it cites.
Agent-aware dropout dqn for safe and efficient on-line dialogue policy learning
Lu Chen, Xiang Zhou, Cheng Chang, Runzhe Yang, and Kai Yu. 2017 · 2017
Later among the works it cites.
Sequence to sequence modeling for user simulation in dialog systems
Paul Crook and Alex Marin. 2017 · 2017
Later among the works it cites.
End-to-end task-completion neural dialogue systems
Xiujun Li, Yun-Nung Chen, Lihong Li, Jianfeng Gao, and Asli Celikyilmaz. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kai Sun, Lu Chen, Su Zhu, and Kai Yu. 2014 · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Cited alongside, same era.
TensorFlow: Large-scale machine learning on heterogeneous systems
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dandelion Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2015 · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
Implementation of generic positive-negative tracker in extensible dialog system
Sangjun Koo, Seonghan Ryu, and Gary Geunbae Lee. 2015 · 2015
Cited alongside, same era.
Incremental lstm-based dialog state tracker
Lukas Zilka and Filip Jurcicek. 2015 · 2015
Cited alongside, same era.
A sequence-to-sequence model for user simulation in spoken dialogue systems
Layla El Asri, Jing He, and Kaheer Suleman. 2016 · 2016
Cited alongside, same era.
Bing Liu and Ian Lane. 2017 · 2017
Later among the works it cites.
Neural belief tracker: Data-driven dialogue state tracking
Nikola Mrkšić, Diarmuid Ó Séaghdha, Tsung-Hsien Wen, Blaise Thomson, and Steve Young. 2017 · 2017
Later among the works it cites.
Regularized neural user model for goal oriented spoken dialogue systems
Manex Serras, María Inés Torres Torres, and Arantza del Pozo. 2017 · 2017
Later among the works it cites.
PyDial: A Multi-domain Statistical Dialogue System Toolkit
Stefan Ultes, Lina M. Rojas Barahona, Pei-Hao Su, David Vandyke, Dongho Kim, Iñigo Casanueva, Paweł Budzianowski, Nikola Mrkšić, Tsung-Hsien Wen, Milica Gašić, and Steve Young. 2017 · 2017
Later among the works it cites.
Feudal reinforcement learning for dialogue management in large domains
Iñigo Casanueva, Paweł Budzianowski, Pei-Hao Su, Stefan Ultes, Lina Rojas-Barahona, Bo-Hsiang Tseng, and Milica Gašić. 2018 · 2018
Closest in time.
Large-scale multi-domain belief tracking with knowledge sharing
Osman Ramadan, Paweł Budzianowski, and Milica Gašić. 2018 · 2018
Closest in time.
Building a conversational agent overnight with dialogue self-play
Pararth Shah, Dilek Hakkani-Tür, Gokhan Tür, Abhinav Rastogi, Ankur Bapna, Neha Nayak, and Larry Heck. 2018 · 2018
Closest in time.
Sample efficient deep reinforcement learning for dialogue systems with large action spaces
Gellért Weisz, Paweł Budzianowski, Pei-Hao Su, and Milica Gašić. 2018 · 2018
Closest in time.