Fetching the paper…
Reading the bibliography…
The present paper surveys neural approaches to conversational AI that have been developed in the last few years.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
Thompson, W. R. (1933) · 1933
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Mnih, V., Adrià, Badia, P., Mirza, M., Graves, A., Lillicrap, T. P., Harley, T., Silver, D., and Kavukcuoglu, K. (2016) · 1937
Earlier work this paper cites.
The perceptron: A perceiving and recognizing automaton
Rosenblatt, F. (1957) · 1957
Earlier work this paper cites.
Principles of Neurodynamics
Rosenblatt, F. (1962) · 1962
Earlier work this paper cites.
ELIZA: a computer program for the study of natural language communication between man and machine
Weizenbaum, J. (1966) · 1966
Earlier work this paper cites.
Perceptrons: An Introduction to Computational Geometry
Minsky, M. and Papert, S. (1969) · 1969
Earlier work this paper cites.
Artificial Paranoia: A Computer Simulation of Paranoid Processes
Colby, K. M. (1975) · 1975
Earlier work this paper cites.
A plan recognition model for subdialogues in conversations
Litman, D. J. and Allen, J. F. (1987) · 1987
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
Sutton, R. S. (1988) · 1988
Earlier work this paper cites.
Learning from Delayed Rewards
Watkins, C. J. (1989) · 1989
Earlier work this paper cites.
Integrated architectures for learning, planning, and reacting based on approximating dynamic programming
Sutton, R. S. (1990) · 1990
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen Netzen. Diploma thesis, Institut für Informatik, Lehrstuhl Prof. Brauer, Technische Universität München
Hochreiter, S. (1991) · 1991
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Lin, L.-J. (1992) · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J. (1992) · 1992
Earlier work this paper cites.
Feudal reinforcement learning
Dayan, P. and Hinton, G. E. (1993) · 1993
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Hinton, G. E. and Van Camp, D. (1993) · 1993
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Puterman, M. L. (1994) · 1994
Earlier work this paper cites.
Text chunking using transformation based learning
Ramshaw, L. A. and Marcus, M. (1995) · 1995
Earlier work this paper cites.
Temporal difference learning and TD-Gammon
Tesauro, G. (1995) · 1995
Earlier work this paper cites.
Neuro-Dynamic Programming
Bertsekas, D. P. and Tsitsiklis, J. N. (1996) · 1996
Earlier work this paper cites.
Reinforcement learning: A survey
Kaelbling, L. P., Littman, M. L., and Moore, A. P. (1996) · 1996
Earlier work this paper cites.
Coding dialogs with the DAMSL annotation scheme
Core, M. G. and Allen, J. F. (1997) · 1997
Earlier work this paper cites.
User modeling for spoken dialogue system evaluation
Eckert, W., Levin, E., and Pieraccini, R. (1997) · 1997
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Machine Learning
Mitchell, T. (1997) · 1997
Earlier work this paper cites.
PARADISE: A framework for evaluating spoken dialogue agents
Walker, M. A., Litman, D. J., Kamm, C. A., and Abella, A. (1997) · 1997
Earlier work this paper cites.
Multitask learning
Caruana, R. (1998) · 1998
Earlier work this paper cites.
Generation that exploits corpus-based statistical knowledge
Langkilde, I. and Knight, K. (1998) · 1998
Earlier work this paper cites.
Reinforcement learning with hierarchies of machines
Parr, R. and Russell, S. J. (1998) · 1998
Earlier work this paper cites.
MindNet: acquiring and structuring semantic information from text
Richardson, S. D., Dolan, W. B., and Vanderwende, L. (1998) · 1998
Earlier work this paper cites.
Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning
Sutton, R. S., Precup, D., and Singh, S. P. (1999b) · 1998
Earlier work this paper cites.
Evaluating spoken dialogue agents with PARADISE: Two case studies
Walker, M. A., Litman, D. J., Kamm, C. A., and Abella, A. (1998) · 1998
Earlier work this paper cites.
Pragmatic reasoning: Inferring contexts
Bell, J. (1999) · 1999
Earlier work this paper cites.
Tools for research and education in speech science
Cole, R. A. (1999) · 1999
Earlier work this paper cites.
Principles of mixed-initiative user interfaces
Horvitz, E. (1999) · 1999
Earlier work this paper cites.
Actor-critic algorithms
Konda, V. R. and Tsitsiklis, J. N. (1999) · 1999
Earlier work this paper cites.
Speech acts for dialogue agents
Traum, D. R. (1999) · 1999
Earlier work this paper cites.
Hierarchical reinforcement learning with the MAXQ value function decomposition
Dietterich, T. G. (2000) · 2000
Earlier work this paper cites.
Information state and dialogue management in the TRINDI dialogue move engine toolkit
Larsson, S. and Traum, D. R. (2000) · 2000
Earlier work this paper cites.
A stochastic model of human-machine interaction for learning dialog strategies
Levin, E., Pieraccini, R., and Eckert, W. (2000) · 2000
Earlier work this paper cites.
Eligibility traces for off-policy policy evaluation
Precup, D., Sutton, R. S., and Singh, S. P. (2000) · 2000
Earlier work this paper cites.
Spoken dialogue management using probabilistic reasoning
Roy, N., Pineau, J., and Thrun, S. (2000) · 2000
Earlier work this paper cites.
BoosTexter: A boosting-based system for text categorization
Schapire, R. E. and Singer, Y. (2000) · 2000
Earlier work this paper cites.
Towards developing general models of usability with PARADISE
Walker, M., Kamm, C., and Litman, D. (2000) · 2000
Earlier work this paper cites.
An application of reinforcement learning to dialogue strategy selection in a spoken dialogue system for email
Walker, M. A. (2000) · 2000
Earlier work this paper cites.
Toward conversational human-computer interaction
Allen, J. F., Byron, D. K., Dzikovska, M. O., Ferguson, G., Galescu, L., and Stent, A. (2001) · 2001
Earlier work this paper cites.
Infinite-horizon policy-gradient estimation
Baxter, J. and Bartlett, P. (2001) · 2001
Earlier work this paper cites.
Experiments with infinite-horizon, policy-gradient estimation
Baxter, J., Bartlett, P., and Weaver, L. (2001) · 2001
Earlier work this paper cites.
A machine learning approach to the automatic evaluation of machine translation
Corston-Oliver, S., Gamon, M., and Brockett, C. (2001) · 2001
Earlier work this paper cites.
Spoken language processing: A guide to theory, algorithm, and system development
Huang, X., Acero, A., and Hon, H.-W. (2001) · 2001
Earlier work this paper cites.
A natural policy gradient
Kakade, S. (2001) · 2001
Earlier work this paper cites.
Empirical methods for evaluating dialog systems
Paek, T. (2001) · 2001
Earlier work this paper cites.
COLLAGEN: Applying collaborative discourse theory to human-computer interaction
Rich, C., Sidner, C. L., and Lesh, N. (2001) · 2001
Earlier work this paper cites.
Spoken dialogue technology: Enabling the conversational user interface
McTear, M. F. (2002) · 2002
Earlier work this paper cites.
Stochastic natural language generation for spoken dialog systems
Oh, A. H. and Rudnicky, A. I. (2002) · 2002
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J. (2002) · 2002
Earlier work this paper cites.
Optimizing dialogue management with reinforcement learning: Experiments with the NJFun system
Singh, S. P., Litman, D., Kearns, M. J., and Walker, M. (2002) · 2002
Earlier work this paper cites.
DIPPER: Description and formalisation of an information-state update dialogue system architecture
Bos, J., Klein, E., Lemon, O., and Oka, T. (2003) · 2003
Earlier work this paper cites.
Statistical phrase-based translation
Koehn, P., Och, F. J., and Marcu, D. (2003) · 2003
Earlier work this paper cites.
Minimum error rate training in statistical machine translation
Och, F. J. (2003) · 2003
Earlier work this paper cites.
A systematic comparison of various statistical alignment models
Och, F. J. and Ney, H. (2003) · 2003
Earlier work this paper cites.
Microplanning with communicative intentions: The SPUD system
Stone, M., Doran, C., Webber, B. L., Bleam, T., and Palmer, M. (2003) · 2003
Earlier work this paper cites.
Dueling network architectures for deep reinforcement learning
Wang, Z., Schaul, T., Hessel, M., van Hasselt, H., Lanctot, M., and de Freitas, N. (2016) · 2003
Earlier work this paper cites.
Subjective evaluation of spoken dialogue systems using SERVQUAL method
Hartikainen, M., Salonen, E.-P., and Turunen, M. (2004) · 2004
Earlier work this paper cites.
Statistical significance tests for machine translation evaluation
Koehn, P. (2004) · 2004
Earlier work this paper cites.
A learning approach to improving sentence-level MT evaluation
Kulesza, A. and Shieber, S. M. (2004) · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Lin, C.-Y. (2004) · 2004
Earlier work this paper cites.
The alignment template approach to statistical machine translation
Och, F. J. and Ney, H. (2004) · 2004
Earlier work this paper cites.
Trainable sentence planning for complex information presentations in spoken dialog systems
Stent, A., Prasad, R., and Walker, M. A. (2004) · 2004
Earlier work this paper cites.
METEOR: An automatic metric for mt evaluation with improved correlation with human judgments
Banerjee, S. and Lavie, A. (2005) · 2005
Earlier work this paper cites.
Reinforcement learning with Gaussian processes
Engel, Y., Mannor, S., and Meir, R. (2005) · 2005
Earlier work this paper cites.
Tree-based batch mode reinforcement learning
Ernst, D., Geurts, P., and Wehenkel, L. (2005) · 2005
Earlier work this paper cites.
Framewise phoneme classification with bidirectional LSTM and other neural network architectures
Graves, A. and Schmidhuber, J. (2005) · 2005
Earlier work this paper cites.
BLANC: Learning evaluation metrics for MT
Lita, L. V., Rogati, M., and Lavie, A. (2005) · 2005
Earlier work this paper cites.
Natural actor-critic
Peters, J., Vijayakumar, S., and Schaal, S. (2005) · 2005
Earlier work this paper cites.
Spoken language understanding: An introduction to the statistical framework
Wang, Y.-Y., Deng, L., and Acero, A. (2005) · 2005
Earlier work this paper cites.
User simulation for spoken dialogue systems: Learning and evaluation
Georgila, K., Henderson, J., and Lemon, O. (2006) · 2006
Earlier work this paper cites.
Multi-domain spoken dialogue system with extensibility and robustness against speech recognition errors
Komatani, K., Kanda, N., Nakano, M., Nakadai, K., Tsujino, H., Ogata, T., and Okuno, H. G. (2006) · 2006
Earlier work this paper cites.
A probabilistic framework for dialog simulation and optimal strategy learning
Pietquin, O. and Dutoit, T. (2006) · 2006
Earlier work this paper cites.
A survey of statistical user simulation techniques for reinforcement-learning of dialogue management strategies
Schatzmann, J., Weilhammer, K., Stuttle, M., and Young, S. (2006) · 2006
Earlier work this paper cites.
Combinatorics of Go
Tromp, J. and Farnebäck, G. (2006) · 2006
Earlier work this paper cites.
Partially Observable Markov Decision Processes for Spoken Dialogue Management
Williams, J. D. (2006) · 2006
Earlier work this paper cites.
Comparing spoken dialog corpora collected with recruited subjects versus real users
Ai, H., Raux, A., Bohus, D., Eskenazi, M., and Litman, D. (2007) · 2007
Earlier work this paper cites.
A re-examination of machine learning approaches for sentence-level mt evaluation
Albrecht, J. and Hwa, R. (2007) · 2007
Earlier work this paper cites.
DBpedia: A nucleus for a web of open data
Auer, S., Bizer, C., Kobilarov, G., Lehmann, J., Cyganiak, R., and Ives, Z. (2007) · 2007
Earlier work this paper cites.
Overview of the TREC 2007 question answering track
Dang, H. T., Kelly, D., and Lin, J. J. (2007) · 2007
Earlier work this paper cites.
Sentence generation as a planning problem
Koller, A. and Stone, M. (2007) · 2007
Earlier work this paper cites.
YAGO: a core of semantic knowledge
Suchanek, F. M., Kasneci, G., and Weikum, G. (2007) · 2007
Earlier work this paper cites.
Individual and domain adaptation in sentence planning for dialogue
Walker, M. A., Stent, A., Mairesse, F., and Prasad, R. (2007) · 2007
Earlier work this paper cites.
Partially observable Markov decision processes for spoken dialog systems
Williams, J. D. and Young, S. J. (2007) · 2007
Earlier work this paper cites.
Assessing dialog system user simulation evaluation measures using human judges
Ai, H. and Litman, D. J. (2008) · 2008
Earlier work this paper cites.
Freebase: a collaboratively created graph database for structuring human knowledge
Bollacker, K., Evans, C., Paritosh, P., Sturge, T., and Taylor, J. (2008) · 2008
Earlier work this paper cites.
A smorgasbord of features for automatic MT evaluation
Giménez, J. and Màrquez, L. (2008) · 2008
Earlier work this paper cites.
Hybrid reinforcement/supervised learning of dialogue policies from fixed data sets
Henderson, J., Lemon, O., and Georgila, K. (2008) · 2008
Earlier work this paper cites.
Finite-time bounds for sampling-based fitted value iteration
Munos, R. and Szepesvári, C. (2008) · 2008
Earlier work this paper cites.
Automating spoken dialogue management design using machine learning: An industry perspective
Paek, T. and Pieraccini, R. (2008) · 2008
Earlier work this paper cites.
The NIST 2008 metrics for machine translation challenge—overview, methodology, metrics, and results
Przybocki, M., Peterson, K., Bronsart, S., and Sanders, G. (2009) · 2008
Earlier work this paper cites.
Learning effective multimodal dialogue strategies from wizard-of-oz data: Bootstrapping and evaluation
Rieser, V. and Lemon, O. (2008) · 2008
Earlier work this paper cites.
Evaluating user simulations with the Cramér-von Mises divergence
Williams, J. D. (2008) · 2008
Earlier work this paper cites.
An integrative and discriminative technique for spoken utterance classification
Yaman, S., Deng, L., Yu, D., Wang, Y.-Y., and Acero, A. (2008) · 2008
Earlier work this paper cites.
Natural Language Processing with Python
Bird, S., Klein, E., and Loper, E. (2009) · 2009
Earlier work this paper cites.
Models for multiparty engagement in open-world dialog
Bohus, D. and Horvitz, E. (2009) · 2009
Earlier work this paper cites.
The RavenClaw dialog management framework: Architecture and systems
Bohus, D. and Rudnicky, A. I. (2009) · 2009
Earlier work this paper cites.
Findings of the 2009 Workshop on Statistical Machine Translation
Callison-Burch, C., Koehn, P., Monz, C., and Schroeder, J. (2009) · 2009
Earlier work this paper cites.
Data-driven user simulation for automated evaluation of spoken dialog systems
Jung, S., Lee, C., Kim, K., Jeong, M., and Lee, G. G. (2009) · 2009
Earlier work this paper cites.
Speech & language processing
Jurafsky, D. and Martin, J. H. (2009) · 2009
Earlier work this paper cites.
Reinforcement learning for spoken dialog management using least-squares policy iteration and fast feature selection
Li, L., Williams, J. D., and Balakrishnan, S. (2009) · 2009
Earlier work this paper cites.
Measuring machine translation quality as semantic equivalence: A metric based on entailment features
Pado, S., Cer, D., Galley, M., Jurafsky, D., and Manning, C. D. (2009) · 2009
Earlier work this paper cites.
The hidden agenda user simulation model
Schatzmann, J. and Young, S. (2009) · 2009
Earlier work this paper cites.
Reinforcement learning in finite MDPs: PAC analysis
Strehl, A. L., Li, L., and Littman, M. L. (2009) · 2009
Earlier work this paper cites.
A simple domain-independent probabilistic approach to generation
Angeli, G., Liang, P., and Klein, D. (2010) · 2010
Earlier work this paper cites.
Spoken dialog challenge 2010: Comparison of live and control test results
Black, A. W., Burger, S., Conkie, A., Hastie, H. W., Keizer, S., Lemon, O., Merigaud, N., Parent, G., Schubiner, G., Thomson, B., Williams, J. D., Yu, K., Young, S. J., and Eskénazi, M. (2011) · 2010
Earlier work this paper cites.
The best lexical metric for phrase-based statistical MT system optimization
Cer, D., Manning, C. D., and Jurafsky, D. (2010) · 2010
Earlier work this paper cites.
Evaluation of a hierarchical reinforcement learning spoken dialogue system
Cuayáhuitl, H., Renals, S., Lemon, O., and Shimodaira, H. (2010) · 2010
Earlier work this paper cites.
Filter, rank, and transfer the knowledge: Learning to chat
Jafarpour, S., Burges, C. J., and Ritter, A. (2010) · 2010
Earlier work this paper cites.
Near-optimal regret bounds for reinforcement learning
Jaksch, T., Ortner, R., and Auer, P. (2010) · 2010
Earlier work this paper cites.
Relational retrieval using a combination of path-constrained random walks
Lao, N. and Cohen, W. W. (2010) · 2010
Earlier work this paper cites.
Natural language generation as planning under uncertainty for spoken dialogue systems
Rieser, V. and Lemon, O. (2010) · 2010
Earlier work this paper cites.
Optimising information presentation for spoken dialogue systems
Rieser, V., Lemon, O., and Liu, X. (2010) · 2010
Earlier work this paper cites.
Algorithms for Reinforcement Learning
Szepesvári, C. (2010) · 2010
Earlier work this paper cites.
Adapting boosting for information retrieval measures
Wu, Q., Burges, C. J., Svore, K. M., and Gao, J. (2010) · 2010
Earlier work this paper cites.
The Hidden Information State model: A practical framework for POMDP-based spoken dialogue management
Young, S. J., Gašić, M., Keizer, S., Mairesse, F., Schatzmann, J., Thomson, B., and Yu, K. (2010) · 2010
Earlier work this paper cites.
Multiparty turn taking in situated dialog: Study, lessons, and directions
Bohus, D. and Horvitz, E. (2011) · 2011
Cited alongside, same era.
User simulation in dialogue systems using inverse reinforcement learning
Chandramohan, S., Geist, M., Lefèvre, F., and Pietquin, O. (2011) · 2011
Cited alongside, same era.
Uncertainty management for on-line optimisation of a POMDP-based large-scale spoken dialogue system
Daubigney, L., Gašić, M., Chandramohan, S., Geist, M., Pietquin, O., and Young, S. J. (2011) · 2011
Cited alongside, same era.
Random walk inference and learning in a large scale knowledge base
Lao, N., Mitchell, T., and Cohen, W. W. (2011) · 2011
Cited alongside, same era.
Learning what to say and how to say it: Joint optimisation of spoken dialogue management and natural language generation
Lemon, O. (2011) · 2011
Cited alongside, same era.
Improved techniques for training GANs
Salimans, T., Goodfellow, I. J., Zaremba, W., Cheung, V., Radford, A., and Chen, X. (2016) · 2016
Later among the works it cites.
Bidirectional attention flow for machine comprehension
Seo, M., Kembhavi, A., Farhadi, A., and Hajishirzi, H. (2016) · 2016
Later among the works it cites.
Building end-to-end dialogue systems using generative hierarchical neural network models
Serban, I. V., Sordoni, A., Bengio, Y., Courville, A. C., and Pineau, J. (2016) · 2016
Later among the works it cites.
Implicit ReasoNet: Modeling large-scale structured relationships with shared memory
Shen, Y., Huang, P., Chang, M., and Gao, J. (2016) · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pietquin, O., Geist, M., Chandramohan, S., and Frezza-Buet, H. (2011) · 2011
Cited alongside, same era.
Learning and evaluation of dialogue strategies for new applications: Empirical methods for optimization from small data sets
Rieser, V. and Lemon, O. (2011) · 2011
Cited alongside, same era.
Data-driven response generation in social media
Ritter, A., Cherry, C., and Dolan, W. (2011) · 2011
Cited alongside, same era.
Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Roemmele, M., Bejan, C., and S. Gordon, A. (2011) · 2011
Cited alongside, same era.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G. J., and Bagnell, D. (2011) · 2011
Cited alongside, same era.
Learning from real users: Rating dialogue success with neural networks for reinforcement learning in spoken dialogue systems
Su, P.-H., Vandyke, D., Gašić, M., Kim, D., Mrkšić, N., Wen, T.-H., and Young, S. (2015) · 2011
Cited alongside, same era.
Spoken language understanding: Systems for extracting semantic information from speech
Tur, G. and De Mori, R. (2011) · 2011
Cited alongside, same era.
Sordoni, A., Bachman, P., Trischler, A., and Bengio, Y. (2016) · 2016
Later among the works it cites.
Data-efficient off-policy policy evaluation for reinforcement learning
Thomas, P. S. and Brunskill, E. (2016) · 2016
Later among the works it cites.
Compositional learning of embeddings for relation paths in knowledge base and text
Toutanova, K., Lin, V., Yih, W.-t., Poon, H., and Quirk, C. (2016) · 2016
Later among the works it cites.
NewsQA: A machine comprehension dataset
Trischler, A., Wang, T., Yuan, X., Harris, J., Sordoni, A., Bachman, P., and Suleman, K. (2016) · 2016
Later among the works it cites.
WaveNet: A generative model for raw audio
van den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A., and Kavukcuoglu, K. (2016) · 2016
Later among the works it cites.
Multi-domain neural network language generation for spoken dialogue systems
Wen, T.-H., Gašić, M., Mrkšić, N., Rojas-Barahona, L. M., Su, P.-H., Vandyke, D., and Young, S. J. (2016) · 2016
Later among the works it cites.
End-to-end LSTM-based dialog control optimized with supervised and reinforcement learning
Williams, J. D. and Zweig, G. (2016) · 2016
Later among the works it cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., Klingner, J., Shah, A., Johnson, M., Liu, X., Kaiser, L., Gouws, S., Kato, Y., Kudo, T., Kazawa, H., Stevens, K., Kurian, G., Patil, N., Wang, W., Young, C., Smith, J., Riesa, J., Rudnick, A., Vinyals, O., Corrado, G., Hughes, M., and Dean, J. (2016) · 2016
Later among the works it cites.
Dynamic coattention networks for question answering
Xiong, C., Zhong, V., and Socher, R. (2016) · 2016
Later among the works it cites.
Learning to respond with deep neural networks for retrieval-based human-computer conversation system
Yan, R., Song, Y., and Wu, H. (2016) · 2016
Later among the works it cites.
Deep learning and continuous representations for natural language processing
Yih, W.-t., He, X., and Gao, J. (2016) · 2016
Later among the works it cites.
Evaluation of statistical POMDP-based dialogue systems in noisy environments
Young, S., Breslin, C., Gašić, M., Henderson, M., Kim, D., Szummer, M., Thomson, B., Tsiakoulis, P., and Hancock, E. T. (2016) · 2016
Later among the works it cites.
Towards end-to-end learning for dialog state tracking and management using deep reinforcement learning
Zhao, T. and Eskénazi, M. (2016) · 2016
Later among the works it cites.
Towards zero-shot frame semantic parsing for domain scaling
Bapna, A., Tür, G., Hakkani-Tür, D., and Heck, L. P. (2017) · 2017
Later among the works it cites.
soc2seq: Social embedding meets conversation model
Bhatia, P., Gavaldà, M., and Einolghozati, A. (2017) · 2017
Later among the works it cites.
Learning end-to-end goal-oriented dialog
Bordes, A., Boureau, Y.-L., and Weston, J. (2017) · 2017
Later among the works it cites.
Sub-domain modelling for dialogue management with hierarchical reinforcement learning
Budzianowski, P., Ultes, S., Su, P.-H., Mrkšić, N., Wen, T.-H., nigo Casanueva, I., Rojas-Barahona, L. M., and Gašić, M. (2017) · 2017
Later among the works it cites.
Agent-aware dropout DQN for safe and efficient on-line dialogue policy learning
Chen, L., Zhou, X., Chang, C., Yang, R., and Yu, K. (2017d) · 2017
Later among the works it cites.
Open-domain neural dialogue systems
Chen, Y.-N. and Gao, J. (2017) · 2017
Later among the works it cites.
Unifying PAC and regret: Uniform PAC bounds for episodic reinforcement learning
Dann, C., Lattimore, T., and Brunskill, E. (2017) · 2017
Later among the works it cites.
GuessWhat?! visual object discovery through multi-modal dialogue
de Vries, H., Strub, F., Chandar, S., Pietquin, O., Larochelle, H., and Courville, A. C. (2017) · 2017
Later among the works it cites.
Towards end-to-end reinforcement learning of dialogue agents for information access
Dhingra, B., Li, L., Li, X., Gao, J., Chen, Y.-N., Ahmed, F., and Deng, L. (2017) · 2017
Later among the works it cites.
Visualizing and understanding neural machine translation
Ding, Y., Liu, Y., Luan, H., and Sun, M. (2017) · 2017
Later among the works it cites.
SearchQA: A new Q&A dataset augmented with context from a search engine
Dunn, M., Sagun, L., Higgins, M., Guney, U., Cirik, V., and Cho, K. (2017) · 2017
Later among the works it cites.
Key-value retrieval networks for task-oriented dialogue
Eric, M., Krishnan, L., Charette, F., and Manning, C. D. (2017) · 2017
Later among the works it cites.
Avoiding echo-responses in a retrieval-based conversation system
Fedorenko, D. G., Smetanin, N., and Rodichev, A. (2017) · 2017
Later among the works it cites.
An introduction to deep learning for natural language processing
Gao, J. (2017) · 2017
Later among the works it cites.
Q-Prop: Sample-efficient policy gradient with an off-policy critic
Gu, S., Lillicrap, T., Ghahramani, Z., Turner, R. E., and Levine, S. (2017) · 2017
Later among the works it cites.
End-to-end conversation modeling track in DSTC6
Hori, C. and Hori, T. (2017) · 2017
Later among the works it cites.
The sixth dialog state tracking challenge
Hori, C., Perez, J., Yoshino, K., and Kim, S. (2017) · 2017
Later among the works it cites.
Mnemonic reader for machine comprehension
Hu, M., Peng, Y., and Qiu, X. (2017) · 2017
Later among the works it cites.
FusionNet: Fusing via fully-aware attention with application to machine comprehension
Huang, H.-Y., Zhu, C., Shen, Y., and Chen, W. (2017) · 2017
Later among the works it cites.
Search-based neural structured learning for sequential question answering
Iyyer, M., Yih, W.-t., and Chang, M.-W. (2017) · 2017
Later among the works it cites.
Adversarial examples for evaluating reading comprehension systems
Jia, R. and Liang, P. (2017) · 2017
Later among the works it cites.
Contextual decision processes with low Bellman rank are PAC-learnable
Jiang, N., Krishnamurthy, A., Agarwal, A., Langford, J., and Schapire, R. E. (2017) · 2017
Later among the works it cites.
TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension
Joshi, M., Choi, E., Weld, D. S., and Zettlemoyer, L. (2017) · 2017
Later among the works it cites.
The NarrativeQA reading comprehension challenge
Kočisky, T., Schwarz, J., Blunsom, P., Dyer, C., Hermann, K. M., Melis, G., and Grefenstette, E. (2017) · 2017
Later among the works it cites.
RACE: Large-scale reading comprehension dataset from examinations
Lai, G., Xie, Q., Liu, H., Yang, Y., and Hovy, E. (2017) · 2017
Later among the works it cites.
Deal or no deal? End-to-end learning of negotiation dialogues
Lewis, M., Yarats, D., Dauphin, Y., Parikh, D., and Batra, D. (2017) · 2017
Later among the works it cites.
Adversarial learning for neural dialogue generation
Li, J., Monroe, W., Shi, T., Jean, S., Ritter, A., and Jurafsky, D. (2017c) · 2017
Later among the works it cites.
Iterative policy learning in end-to-end trainable task-oriented neural dialog models
Liu, B. and Lane, I. (2017) · 2017
Later among the works it cites.
Phase conductor on multi-layered attentions for machine comprehension
Liu, R., Wei, W., Mao, W., and Chikina, M. (2017) · 2017
Later among the works it cites.
Towards an automatic turing test: Learning to evaluate dialogue responses
Lowe, R., Noseworthy, M., Serban, I. V., Angelard-Gontier, N., Bengio, Y., and Pineau, J. (2017) · 2017
Later among the works it cites.
Multi-task learning for speaker-role adaptation in neural conversation models
Luan, Y., Brockett, C., Dolan, B., Gao, J., and Galley, M. (2017) · 2017
Later among the works it cites.
Learned in translation: Contextualized word vectors
McCann, B., Bradbury, J., Xiong, C., and Socher, R. (2017) · 2017
Later among the works it cites.
Coherent dialogue with attention-based language models
Mei, H., Bansal, M., and Walter, M. R. (2017) · 2017
Later among the works it cites.
Image-grounded conversations: Multimodal context for natural question and response generation
Mostafazadeh, N., Brockett, C., Dolan, B., Galley, M., Gao, J., Spithourakis, G., and Vanderwende, L. (2017) · 2017
Later among the works it cites.
Neural belief tracker: Data-driven dialogue state tracking
Mrkšić, N., Séaghdha, D. O., Wen, T.-H., Thomson, B., and Young, S. J. (2017) · 2017
Later among the works it cites.
An overview of embedding models of entities and relationships for knowledge base completion
Nguyen, D. Q. (2017) · 2017
Later among the works it cites.
Why is posterior sampling better than optimism for reinforcement learning?
Osband, I. and Roy, B. V. (2017) · 2017
Later among the works it cites.
Composite task-completion dialogue policy learning via hierarchical deep reinforcement learning
Peng, B., Li, X., Li, L., Gao, J., Celikyilmaz, A., Lee, S., and Wong, K.-F. (2017) · 2017
Later among the works it cites.
Get to the point: Summarization with pointer-generator networks
See, A., Liu, P. J., and Manning, C. D. (2017) · 2017
Later among the works it cites.
A hierarchical latent variable encoder-decoder model for generating dialogues
Serban, I. V., Sordoni, A., Lowe, R., Charlin, L., Pineau, J., Courville, A., and Bengio, Y. (2017) · 2017
Later among the works it cites.
Generating high-quality and informative conversation responses with sequence-to-sequence models
Shao, Y., Gouws, S., Britz, D., Goldie, A., Strope, B., and Kurzweil, R. (2017) · 2017
Later among the works it cites.
End-to-end optimization of goal-driven and visually grounded dialogue systems
Strub, F., de Vries, H., Mary, J., Piot, B., Courville, A. C., and Pietquin, O. (2017) · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I. (2017) · 2017
Later among the works it cites.
FastQA: A simple and efficient neural architecture for question answering
Weissenborn, D., Wiese, G., and Seiffe, L. (2017) · 2017
Later among the works it cites.
Constructing datasets for multi-hop reading comprehension across documents
Welbl, J., Stenetorp, P., and Riedel, S. (2017) · 2017
Later among the works it cites.
A network-based end-to-end trainable task-oriented dialogue system
Wen, T.-H., Vandyke, D., Mrkšić, N., Gašić, M., Rojas-Barahona, L. M., Su, P.-H., Ultes, S., and Young, S. J. (2017) · 2017
Later among the works it cites.
Hybrid code networks: Practical and efficient end-to-end dialog control with supervised and reinforcement learning
Williams, J. D., Asadi, K., and Zweig, G. (2017) · 2017
Later among the works it cites.
Nora the empathetic psychologist
Winata, G. I., Kampman, O., Yang, Y., Dey, A., and Fung, P. (2017) · 2017
Later among the works it cites.
DeepPath: A reinforcement learning method for knowledge graph reasoning
Xiong, W., Hoang, T., and Wang, W. Y. (2017) · 2017
Later among the works it cites.
Neural response generation via GAN with an approximate embedding layer
Xu, Z., Liu, B., Wang, B., Sun, C., Wang, X., Wang, Z., and Qi, C. (2017) · 2017
Later among the works it cites.
End-to-end joint learning of natural language understanding and dialogue manager
Yang, X., Chen, Y.-N., Hakkani-Tür, D. Z., Crook, P., Li, X., Gao, J., and Deng, L. (2017b) · 2017
Later among the works it cites.
Generative encoder-decoder models for task-oriented spoken dialog systems with chatting capability
Zhao, T., Lu, A., Lee, K., and Eskenazi, M. (2017) · 2017
Later among the works it cites.
Comparison of an end-to-end trainable dialogue system with a modular statistical dialogue system
Braunschweiler, N. and Papangelis, A. (2018) · 2018
Closest in time.
Feudal reinforcement learning for dialogue management in large domains
Casanueva, I., Budzianowski, P., Su, P.-H., Ultes, S., Rojas-Barahona, L. M., Tseng, B.-H., and Gašić, M. (2018) · 2018
Closest in time.
Structured dialogue policy with graph neural networks
Chen, L., Tan, B., Long, S., and Yu, K. (2018) · 2018
Closest in time.
QuAC: Question answering in context
Choi, E., He, H., Iyyer, M., Yatskar, M., Yih, W.-t., Choi, Y., Liang, P., and Zettlemoyer, L. (2018) · 2018
Closest in time.
Think you have solved question answering? try ARC, the AI2 reasoning challenge
Clark, P., Cowhey, I., Etzioni, O., Khot, T., Sabharwal, A., Schoenick, C., and Tafjord, O. (2018) · 2018
Closest in time.
TextWorld: A learning environment for text-based games
Côté, M.-A., Ákos Kádár, Yuan, X., Kybartas, B., Barnes, T., Fine, E., Moore, J., Hausknecht, M., El Asri, L., Adada, M., Tay, W., and Trischler, A. (2018) · 2018
Closest in time.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2018) · 2018
Closest in time.
Neural approaches to conversational AI
Gao, J., Galley, M., and Li, L. (2018b) · 2018
Closest in time.
A knowledge-grounded neural conversation model
Ghazvininejad, M., Brockett, C., Chang, M.-W., Dolan, B., Gao, J., Yih, W.-t., and Galley, M. (2018) · 2018
Closest in time.
Learning to write with cooperative discriminators
Holtzman, A., Buys, J., Forbes, M., Bosselut, A., Golub, D., and Choi, Y. (2018) · 2018
Closest in time.
Hierarchically structured reinforcement learning for topically coherent visual story generation
Huang, Q., Gan, Z., Celikyilmaz, A., Wu, D. O., Wang, J., and He, X. (2018) · 2018
Closest in time.
Emotional dialogue generation using image-grounded language models
Huber, B., McDuff, D., Brockett, C., Galley, M., and Dolan, B. (2018) · 2018
Closest in time.
Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition
Jurafsky, D. and Martin, J. H. (2018) · 2018
Closest in time.
Sharp nearby, fuzzy far away: How neural language models use context
Khandelwal, U., He, H., Qi, P., and Jurafsky, D. (2018) · 2018
Closest in time.
A case study on the importance of belief state representation for dialogue policy management
Kotti, M., Diakoloukas, V., Papangelis, A., Lagoudakis, M., and Stylianou, Y. (2018) · 2018
Closest in time.
Sequicity: Simplifying task-oriented dialogue systems with single sequence-to-sequence architectures
Lei, W., Jin, X., Ren, Z., He, X., Kan, M.-Y., and Yin, D. (2018) · 2018
Closest in time.
Microsoft dialogue challenge: Building end-to-end task-completion dialogue systems
Li, X., Panda, S., Liu, J., and Gao, J. (2018) · 2018
Closest in time.
Deep reinforcement learning
Li, Y. (2018) · 2018
Closest in time.
BBQ-networks: Efficient exploration in deep reinforcement learning for task-oriented dialogue systems
Lipton, Z. C., Gao, J., Li, L., Li, X., Ahmed, F., and Deng, L. (2018) · 2018
Closest in time.
Adversarial learning of task-oriented neural dialog models
Liu, B. and Lane, I. (2018) · 2018
Closest in time.
Mem2Seq: Effectively incorporating knowledge bases into end-to-end task-oriented dialog systems
Madotto, A., Wu, C.-S., and Fung, P. (2018) · 2018
Closest in time.
The natural language decathlon: Multitask learning as question answering
McCann, B., Keskar, N. S., Xiong, C., and Socher, R. (2018) · 2018
Closest in time.
Towards scalable information-seeking multi-domain dialogue
Papangelis, A., Kotti, M., and Stylianou, Y. (2018a) · 2018
Closest in time.
Integrating planning for task-completion dialogue policy learning
Peng, B., Li, X., Gao, J., Liu, J., and Wong, K.-F. (2018) · 2018
Closest in time.
Deep contextualized word representations
Peters, M. E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., and Zettlemoyer, L. (2018) · 2018
Closest in time.
Know what you don’t know: Unanswerable questions for SQuAD
Rajpurkar, P., Jia, R., and Liang, P. (2018) · 2018
Closest in time.
Conversational AI: the science behind the alexa prize
Ram, A., Prasad, R., Khatri, C., Venkatesh, A., Gabriel, R., Liu, Q., Nunn, J., Hedayatnia, B., Cheng, M., Nagar, A., King, E., Bland, K., Wartick, A., Pan, Y., Song, H., Jayadevan, S., Hwang, G., and Pettigrue, A. (2018) · 2018
Closest in time.
CoQA: A conversational question answering challenge
Reddy, S., Chen, D., and Manning, C. D. (2018) · 2018
Closest in time.
Conversational query understanding using sequence to sequence modeling
Ren, G., Ni, X., Malik, M., and Ke, Q. (2018a) · 2018
Closest in time.
Towards universal dialogue state tracking
Ren, L., Xie, K., Chen, L., and Yu, K. (2018b) · 2018
Closest in time.
A tutorial on Thompson sampling
Russo, D. J., Van Roy, B., Kazerouni, A., Osband, I., and Wen, Z. (2018) · 2018
Closest in time.
Saha, A., Pahuja, V., Khapra, M. M., Sankaranarayanan, K., and Chandar, S. (2018) · 2018
Closest in time.
A survey of available corpora for building data-driven dialogue systems: The journal version
Serban, I. V., Lowe, R., Henderson, P., Charlin, L., and Pineau, J. (2018) · 2018
Closest in time.
Bootstrapping a neural conversational agent with dialogue self-play, crowdsourcing and on-line reinforcement learning
Shah, P., Hakkani-Tür, D. Z., Liu, B., and Tür, G. (2018) · 2018
Closest in time.
DRCD: a chinese machine reading comprehension dataset
Shao, C. C., Liu, T., Lai, Y., Tseng, Y., and Tsai, S. (2018) · 2018
Closest in time.
M-walk: Learning to walk in graph with monte carlo tree search
Shen, Y., Chen, J., Huang, P., Guo, Y., and Gao, J. (2018) · 2018
Closest in time.
From Eliza to XiaoIce: Challenges and opportunities with social chatbots
Shum, H., He, X., and Li, D. (2018) · 2018
Closest in time.
Discriminative deep Dyna-Q: Robust planning for dialogue policy learning
Su, S.-Y., Li, X., Gao, J., Liu, J., and Chen, Y.-N. (2018b) · 2018
Closest in time.
Natural language generation by hierarchical decoding with linguistic patterns
Su, S.-Y., Lo, K.-L., Yeh, Y. T., and Chen, Y.-N. (2018c) · 2018
Closest in time.
Learning to map context-dependent sentences to executable formal queries
Suhr, A., Iyer, S., and Artzi, Y. (2018) · 2018
Closest in time.
Reinforcement Learning: An Introduction
Sutton, R. S. and Barto, A. G. (2018) · 2018
Closest in time.
The web as a knowledge-base for answering complex questions
Talmor, A. and Berant, J. (2018) · 2018
Closest in time.
Subgoal discovery for hierarchical dialogue policy learning
Tang, D., Li, X., Gao, J., Wang, C., Li, L., and Jebara, T. (2018) · 2018
Closest in time.
AirDialogue: An environment for goal-oriented dialogue research
Wei, W., Le, Q. V., Dai, A. M., and Li, L.-J. (2018) · 2018
Closest in time.
End-to-end dynamic query memory network for entity-value independent task-oriented dialog
Wu, C.-S., Madotto, A., Winata, G. I., and Fung, P. (2018) · 2018
Closest in time.
Hierarchical recurrent attention network for response generation
Xing, C., Wu, W., Wu, Y., Zhou, M., Huang, Y., and Ma, W. (2018) · 2018
Closest in time.
Emo2vec: Learning generalized emotion representation by multi-task training
Xu, P., Madotto, A., Wu, C.-S., Park, J. H., and Fung, P. (2018) · 2018
Closest in time.
QANet: Combining local convolution with global self-attention for reading comprehension
Yu, A. W., Dohan, D., Luong, M.-T., Zhao, R., Chen, K., Norouzi, M., and Le, Q. V. (2018) · 2018
Closest in time.
The design and implementation of XiaoIce, an empathetic social chatbot
Zhou, L., Gao, J., Li, D., and Shum, H.-Y. (2018) · 2018
Closest in time.
Zero-shot adaptive transfer for conversational language understanding
Lee, S. and Jha, R. (2019) · 2019
Closest in time.
A probabilistic framework for representing dialog systems and entropy-based dialog management through dynamic stochastic state evolution
Wu, J., Li, M., and Lee, C.-H. (2015) · 2035
Closest in time.
Dialogue learning with human teaching and feedback in end-to-end trainable task-oriented dialogue systems
Liu, B., Tür, G., Hakkani-Tür, D., Shah, P., and Heck, L. P. (2018a) · 2069
Closest in time.
Deep reinforcement learning with double Q-learning
van Hasselt, H., Guez, A., and Silver, D. (2016) · 2094
Closest in time.