Fetching the paper…
Reading the bibliography…
User simulators are essential for training reinforcement learning (RL) based dialog models.
Tiancheng Zhao, Kaige Xie, and Maxine Eskenazi. 2019 · 1902
Earlier work this paper cites.
Unifying human and statistical evaluation for natural language generation
Tatsunori B Hashimoto, Hugh Zhang, and Percy Liang. 2019 · 1904
Earlier work this paper cites.
Convlab: Multi-domain end-to-end dialog system platform
Sungjin Lee, Qi Zhu, Ryuichi Takanobu, Xiang Li, Yaoqin Zhang, Zheng Zhang, Jinchao Li, Baolin Peng, Xiujun Li, Minlie Huang, et al. 2019 · 1904
Earlier work this paper cites.
Unsupervised dialog structure learning
Weiyan Shi, Tiancheng Zhao, and Zhou Yu. 2019 · 1904
Earlier work this paper cites.
Domain adaptive dialog generation via meta learning
Kun Qian and Zhou Yu. 2019 · 1906
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
User modeling and user-adapted interaction
Alfred Kobsa. 1994 · 1994
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Let’s go public! taking a spoken dialog system to the real world
Antoine Raux, Brian Langner, Dan Bohus, Alan W Black, and Maxine Eskenazi. 2005 · 2005
Earlier work this paper cites.
Effects of the user model on simulation-based learning of dialogue strategies
Jost Schatztnann, Matthew N Stuttle, Karl Weilhammer, and Steve Young. 2005 · 2005
Earlier work this paper cites.
A survey of statistical user simulation techniques for reinforcement-learning of dialogue management strategies
Jost Schatzmann, Karl Weilhammer, Matt Stuttle, and Steve Young. 2006 · 2006
Earlier work this paper cites.
Agenda-based user simulation for bootstrapping a pomdp dialogue system
Jost Schatzmann, Blaise Thomson, Karl Weilhammer, Hui Ye, and Steve Young. 2007 · 2007
Earlier work this paper cites.
Analysis of a new simulation approach to dialog system evaluation
Klaus-Peter Engelbrecht, Michael Quade, and Sebastian Möller. 2009 · 2009
Earlier work this paper cites.
Data-driven user simulation for automated evaluation of spoken dialog systems
Sangkeun Jung, Cheongjae Lee, Kyungduk Kim, Minwoo Jeong, and Gary Geunbae Lee. 2009 · 2009
Earlier work this paper cites.
The hidden agenda user simulation model
Jost Schatzmann and Steve Young. 2009 · 2009
Cited alongside, same era.
Adaptive ε \varepsilon -greedy exploration in reinforcement learning based on value differences
Michel Tokic. 2010 · 2010
Cited alongside, same era.
A survey on metrics for the evaluation of user simulations
Olivier Pietquin and Helen Hastie. 2013 · 2013
Cited alongside, same era.
Pomdp-based statistical spoken dialog systems: A review
Steve Young, Milica Gašić, Blaise Thomson, and Jason D Williams. 2013 · 2013
Cited alongside, same era.
The third dialog state tracking challenge
Matthew Henderson, Blaise Thomson, and Jason D Williams. 2014 · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Cited alongside, same era.
Iterative policy learning in end-to-end trainable task-oriented neural dialog models
Bing Liu and Ian Lane. 2017 · 2017
Later among the works it cites.
Parlai: A dialog research software platform
Alexander H Miller, Will Feng, Adam Fisch, Jiasen Lu, Dhruv Batra, Antoine Bordes, Devi Parikh, and Jason Weston. 2017 · 2017
Later among the works it cites.
Sample-efficient actor-critic reinforcement learning with supervised data for dialogue management
Pei-Hao Su, Pawel Budzianowski, Stefan Ultes, Milica Gasic, and Steve Young. 2017 · 2017
Later among the works it cites.
Semantic refinement gru-based neural language generation for spoken dialogue systems
Van-Khanh Tran and Le-Minh Nguyen. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Strategic dialogue management via deep reinforcement learning
Heriberto Cuayáhuitl, Simon Keizer, and Oliver Lemon. 2015 · 2015
Cited alongside, same era.
A sequence-to-sequence model for user simulation in spoken dialogue systems
Layla El Asri, Jing He, and Kaheer Suleman. 2016 · 2016
Cited alongside, same era.
Sequence-to-sequence generation for spoken dialogue via deep syntax trees and strings
Ondřej Dušek and Filip Jurčíček. 2016 · 2016
Cited alongside, same era.
A user simulator for task-completion dialogues
Xiujun Li, Zachary C Lipton, Bhuwan Dhingra, Lihong Li, Jianfeng Gao, and Yun-Nung Chen. 2016 · 2016
Cited alongside, same era.
Chia-Wei Liu, Ryan Lowe, Iulian V Serban, Michael Noseworthy, Laurent Charlin, and Joelle Pineau. 2016 · 2016
Cited alongside, same era.
Yu Wu, Wei Wu, Chen Xing, Ming Zhou, and Zhoujun Li. 2016 · 2016
Cited alongside, same era.
Jason D Williams, Kavosh Asadi, and Geoffrey Zweig. 2017 · 2017
Later among the works it cites.
Multiwoz-a large-scale multi-domain wizard-of-oz dataset for task-oriented dialogue modelling
Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng, Iñigo Casanueva, Stefan Ultes, Osman Ramadan, and Milica Gašić. 2018 · 2018
Later among the works it cites.
Decoupling strategy and generation in negotiation dialogues
He He, Derek Chen, Anusha Balakrishnan, and Percy Liang. 2018 · 2018
Later among the works it cites.
Neural user simulation for corpus-based policy optimisation for spoken dialogue systems
Florian Kreyssig, Inigo Casanueva, Pawel Budzianowski, and Milica Gasic. 2018 · 2018
Later among the works it cites.
Sequicity: Simplifying task-oriented dialogue systems with single sequence-to-sequence architectures
Wenqiang Lei, Xisen Jin, Min-Yen Kan, Zhaochun Ren, Xiangnan He, and Dawei Yin. 2018 · 2018
Later among the works it cites.
Building a conversational agent overnight with dialogue self-play
Pararth Shah, Dilek Hakkani-Tür, Gokhan Tür, Abhinav Rastogi, Ankur Bapna, Neha Nayak, and Larry Heck. 2018 · 2018
Later among the works it cites.
Sentiment adaptive end-to-end dialog systems
Weiyan Shi and Zhou Yu. 2018 · 2018
Later among the works it cites.
Adversarial domain adaptation for variational neural language generation in dialogue systems
Van-Khanh Tran and Le-Minh Nguyen. 2018 · 2018
Later among the works it cites.
Convolutional neural network architectures for matching natural language sentences
Baotian Hu, Zhengdong Lu, Hang Li, and Qingcai Chen. 2014 · 2050
Closest in time.