Fetching the paper…
Reading the bibliography…
Dialogue assistants are rapidly becoming an indispensable daily aid.
Using markov decision process for learning dialogue strategies
Esther Levin, Roberto Pieraccini, and Wieland Eckert · 1998
Earlier work this paper cites.
Asr system modeling for automatic evaluation and optimization of dialogue systems
Olivier Pietquin and Steve Renals · 2002
Earlier work this paper cites.
Hybrid reinforcement/supervised learning for dialogue policies from communicator data
James Henderson, Oliver Lemon, and Kallirroi Georgila · 2005
Earlier work this paper cites.
A probabilistic framework for dialog simulation and optimal strategy learning
Olivier Pietquin and Thierry Dutoit · 2006
Earlier work this paper cites.
Agenda-based user simulation for bootstrapping a pomdp dialogue system
Jost Schatzmann, Blaise Thomson, Karl Weilhammer, Hui Ye, and Steve Young · 2007
Earlier work this paper cites.
Partially observable Markov decision processes for spoken dialog systems
Jason D. Williams and Steve Young · 2007
Earlier work this paper cites.
The best of both worlds: unifying conventional dialog systems and pomdps
Jason D Williams · 2008
Earlier work this paper cites.
Back-off action selection in summary space-based pomdp dialogue systems
M Gašić, Fabrice Lefevre, F Jurčíček, Simon Keizer, Francois Mairesse, Blaise Thomson, Kai Yu, and Steve Young · 2009
Earlier work this paper cites.
The hidden agenda user simulation model
J. Schatzmann and S. Young · 2009
Earlier work this paper cites.
Parameter estimation for agenda-based user simulation
Simon Keizer, Milica Gašić, Filip Jurčíček, François Mairesse, Blaise Thomson, Kai Yu, and Steve Young · 2010
Earlier work this paper cites.
Natural belief-critic: a reinforcement algorithm for parameter estimation in statistical spoken dialogue systems
Blaise Thomson, Simon Keizer, François Mairesse, Kai Yu, and Steve J Young · 2010
Earlier work this paper cites.
Demonstration of at&t “let’s go”: A production-grade statistical spoken dialog system
Jason D Williams, Iker Arizmendi, and Alistair Conkie · 2010
Earlier work this paper cites.
The hidden information state model: A practical framework for pomdp-based spoken dialogue management
Steve Young, Milica Gašić, Simon Keizer, François Mairesse, Jost Schatzmann, Blaise Thomson, and Kai Yu · 2010
Earlier work this paper cites.
On-line policy optimisation of spoken dialogue systems via live interaction with human subjects
Milica Gašić, Filip Jurcicek, Blaise. Thomson, Kai Yu, and Steve Young · 2011
Earlier work this paper cites.
Sample efficient on-line learning of optimal dialogue policies with kalman temporal differences
Olivier Pietquin, Matthieu Geist, Senthilkumar Chandramohan, et al · 2011
Earlier work this paper cites.
N-best error simulation for training spoken dialogue systems
Blaise Thomson, Milica Gasic, Matthew Henderson, Pirros Tsiakoulis, and Steve Young · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Cited alongside, same era.
Statistical methods for spoken dialogue management
Blaise Thomson · 2013
Cited alongside, same era.
The dialog state tracking challenge
Jason Williams, Antoine Raux, Deepak Ramachandran, and Alan Black · 2013
Cited alongside, same era.
Pomdp-based statistical spoken dialog systems: A review
Steve Young, Milica Gašić, Blaise Thomson, and Jason D Williams · 2013
Cited alongside, same era.
Adaptive speech recognition and dialogue management for users with speech disorders
Benchmarking deep reinforcement learning for continuous control
Yan Duan, Xi Chen, Rein Houthooft, John Schulman, and Pieter Abbeel · 2016
Later among the works it cites.
Policy networks with two-stage training for dialogue systems
Mehdi Fatemi, Layla El Asri, Hannes Schulz, Jing He, and Kaheer Suleman · 2016
Later among the works it cites.
Opendial: A toolkit for developing spoken dialogue systems with probabilistic rules
P. Lison and C. Kennington · 2016
Later among the works it cites.
Building end-to-end dialogue systems using generative hierarchical neural network models
Iulian Vlad Serban, Alessandro Sordoni, Yoshua Bengio, Aaron C Courville, and Joelle Pineau · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, et al · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Inigo Casanueva, Heidi Christensen, Thomas Hain, and Phil D Green · 2014
Cited alongside, same era.
Gaussian processes for pomdp-based dialogue manager optimization
Milica Gasic and Steve Young · 2014
Cited alongside, same era.
Word-based Dialog State Tracking with Recurrent Neural Networks
M. Henderson, B. Thomson, and S. J. Young · 2014
Cited alongside, same era.
Knowledge transfer between speakers for personalised dialogue management
Inigo Casanueva, Thomas Hain, Heidi Christensen, Ricard Marxer, and Phil Green · 2015
Cited alongside, same era.
Hyper-parameter optimisation of gaussian process reinforcement learning for statistical dialogue management
Lu Chen, Pei-Hao Su, and Milica Gašić · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Ziyu Wang, Victor Bapst, Nicolas Heess, Volodymyr Mnih, Remi Munos, Koray Kavukcuoglu, and Nando de Freitas · 2016
Later among the works it cites.
Tiancheng Zhao and Maxine Eskenazi · 2016
Later among the works it cites.
Learning end-to-end goal-oriented dialog
Antoine Bordes, Y-Lan Boureau, and Jason Weston · 2017
Closest in time.
Sub-domain modelling for dialogue management with hierarchical reinforcement learning
Paweł Budzianowski, Stefan Ultes, Pei-Hao Su, Nikola Mrkšić, Tsung-Hsien Wen, Inigo Casanueva, Lina M. Rojas Barahona, and Milica Gašić · 2017
Closest in time.
Neural belief tracker: Data-driven dialogue state tracking
Nikola Mrkšić, Diarmuid O Séaghdha, Tsung-Hsien Wen, Blaise Thomson, and Steve Young · 2017
Closest in time.
Single-model multi-domain dialogue management with deep learning
Alexandros Papangelis and Yannis Stylianou · 2017
Closest in time.
Sample-efficient actor-critic reinforcement learning with supervised data for dialogue management
Pei-Hao Su, Pawel Budzianowski, Stefan Ultes, Milica Gasic, and Steve Young · 2017
Closest in time.
Pydial: A multi-domain statistical dialogue system toolkit
Stefan Ultes, Lina M. Rojas-Barahona, Pei-Hao Su, David Vandyke, Dongho Kim, Iñigo Casanueva, Paweł Budzianowski, Nikola Mrkšić, Tsung-Hsien Wen, Milica Gašić, and Steve J. Young · 2017
Closest in time.
Starcraft ii: A new challenge for reinforcement learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, et al · 2017
Closest in time.
Hybrid code networks: practical and efficient end-to-end dialog control with supervised and reinforcement learning
Jason D Williams, Kavosh Asadi, and Geoffrey Zweig · 2017
Closest in time.
End-to-end joint learning of natural language understanding and dialogue manager
Xuesong Yang, Yun-Nung Chen, Dilek Hakkani-Tür, Paul Crook, Xiujun Li, Jianfeng Gao, and Li Deng · 2017
Closest in time.