Fetching the paper…
Reading the bibliography…
End-to-end learning of recurrent neural networks (RNNs) is an attractive solution for dialog systems; however, current techniques are data-intensive and require thousands of dialogs to learn simple behaviors.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jurgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
A stochastic model of human-machine interaction for learning dialogue strategies
Esther Levin, Roberto Pieraccini, and Wieland Eckert. 2000 · 2000
Earlier work this paper cites.
Optimizing dialogue management with reinforcement leaning: experiments with the NJFun system
Satinder Singh, Diane J Litman, Michael Kearns, and Marilyn A Walker. 2002 · 2002
Earlier work this paper cites.
Policy gradient reinforcement learning for fast quadrupedal locomotion
Nate Kohl and Peter Stone. 2004 · 2004
Earlier work this paper cites.
Partially observable Markov decision processes for spoken dialog systems
Jason D. Williams and Steve Young. 2007 · 2007
Earlier work this paper cites.
A statistical approach to spoken dialog systems design and evaluation
David Griol, Lluís F. Hurtado, Encarna Segarra, and Emilio Sanchis. 2008 · 2008
Earlier work this paper cites.
The best of both worlds: Unifying conventional dialog systems and POMDPs
Jason D. Williams. 2008 · 2008
Earlier work this paper cites.
Statistical dialog management applied to WFST-based dialog systems
Chiori Hori, Kiyonori Ohtake, Teruhisa Misu, Hideki Kashioka, and Satoshi Nakamura. 2009 · 2009
Earlier work this paper cites.
Example-based dialog modeling for practical multi-domain dialog system
Cheongjae Lee, Sangkeun Jung, Seokhwan Kim, and Gary Geunbae Lee. 2009 · 2009
Earlier work this paper cites.
Natural actor and belief critic: Reinforcement algorithm for learning parameters of dialogue systems modelled as pomdps
Filip Jurčíček, Blaise Thomson, and Steve Young. 2011 · 2011
Earlier work this paper cites.
ADADELTA: an adaptive learning rate method
Matthew D. Zeiler. 2012 · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Chris Manning, Andrew Ng, and Chris Potts. 2013 · 2013
Cited alongside, same era.
POMDP-based Statistical Spoken Dialogue Systems: a Review
Steve Young, Milica Gasic, Blaise Thomson, and Jason D. Williams. 2013 · 2013
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Temporal supervised learning for inferring a dialog policy from example conversations
Lihong Li, He He, and Jason D. Williams. 2014 · 2014
Cited alongside, same era.
Neural responding machine for short-text conversation
Lifeng Shang, Zhengdong Lu, , and Hang Li. 2015 · 2015
Cited alongside, same era.
A neural network approach to context-sensitive generation of conversational responses
Coherent dialogue with attention-based language models
Hongyuan Mei, Mohit Bansal, and Matthew R. Walter. 2016 · 2016
Later among the works it cites.
Query-regression networks for machine comprehension
Min Joon Seo, Hannaneh Hajishirzi, and Ali Farhadi. 2016 · 2016
Later among the works it cites.
Building end-to-end dialogue systems using generative hierarchical neural network models
Iulian V. Serban, Alessandro Sordoni, Yoshua Bengio, Aaron Courville, and Joelle Pineau. 2016 · 2016
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al. 2016 · 2016
Later among the works it cites.
Continuously learning neural dialogue management
Pei-Hao Su, Milica Gašić, Nikola Mrkšić, Lina Rojas-Barahona, Stefan Ultes, David Vandyke, Tsung-Hsien Wen, and Steve Young. 2016 · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alessandro Sordoni, Michel Galley, Michael Auli, Chris Brockett, Yangfeng Ji, Meg Mitchell, Jian-Yun Nie, Jianfeng Gao, and Bill Dolan. 2015 · 2015
Cited alongside, same era.
End-to-end memory networks
Sainbayar Sukhbaatar, Arthur Szlam, Jason Weston, and Rob Fergus. 2015 · 2015
Cited alongside, same era.
A neural conversational model
Oriol Vinyals and Quoc Le. 2015 · 2015
Cited alongside, same era.
Attention with intention for a neural network conversation model
Kaisheng Yao, Geoffrey Zweig, and Baolin Peng. 2015 · 2015
Cited alongside, same era.
Learning end-to-end goal-oriented dialog
Antoine Bordes and Jason Weston. 2016 · 2016
Cited alongside, same era.
Gated end-to-end memory networks
Fei Liu and Julien Perez. 2016 · 2016
Cited alongside, same era.
LSTM based conversation models
Yi Luan, Yangfeng Ji, and Mari Ostendorf. 2016 · 2016
Cited alongside, same era.
Later among the works it cites.
Theano: A Python framework for fast computation of mathematical expressions
Theano Development Team. 2016 · 2016
Later among the works it cites.
A network-based end-to-end trainable task-oriented dialogue system
Tsung-Hsien Wen, Milica Gasic, Nikola Mrksic, Lina Maria Rojas-Barahona, Pei-Hao Su, Stefan Ultes, David Vandyke, and Steve J. Young. 2016 · 2016
Later among the works it cites.
Incorporating loose-structured knowledge into LSTM with recall gate for conversation modeling
Zhen Xu, Bingquan Liu, Baoxun Wang, Chengjie Sun, and Xiaolong Wang. 2016 · 2016
Later among the works it cites.
Towards end-to-end reinforcement learning of dialogue agents for information access
Bhuwan Dhingra, Lihong Li, Xiujun Li, Jianfeng Gao, Yun-Nung Chen, Faisal Ahmed, and Li Deng. 2017 · 2017
Closest in time.
A copy-augmented sequence-to-sequence architecture gives good performance on task-oriented dialogue
Mihail Eric and Christopher D Manning. 2017 · 2017
Closest in time.
Training end-to-end dialogue systems with the ubuntu dialogue corpus
Ryan Thomas Lowe, Nissan Pow, Iulian Vlad Serban, Laurent Charlin, Chia-Wei Liu, and Joelle Pineau. 2017 · 2017
Closest in time.
A hierarchical latent variable encoder-decoder model for generating dialogues
Iulian Vlad Serban, Alessandro Sordoni, Ryan Lowe, Laurent Charlin, Joelle Pineau, Aaron Courville, and Yoshua Bengio. 2017 · 2017
Closest in time.