Fetching the paper…
Reading the bibliography…
We present MILABOT: a deep reinforcement learning chatbot developed by the Montreal Institute for Learning Algorithms (MILA) for the Amazon Alexa Prize competition.
Weizenbaum, J. (1966), ‘Eliza—a computer program for the study of natural language communication between man and machine’, ACM
1966
Earlier work this paper cites.
Colby, K. M. (1981), ‘Modeling a paranoid mind’, Behavioral and Brain Sciences
1981
Earlier work this paper cites.
McGlashan, S., Fraser, N., Gilbert, N., Bilange, E., Heisterkamp, P. & Youd, N. (1992), Dialogue management for telephone information systems, in
1992
Earlier work this paper cites.
Williams, R. J. (1992), ‘Simple statistical gradient-following algorithms for connectionist reinforcement learning’, Machine learning
1992
Earlier work this paper cites.
Lin, L.-J. (1993), Reinforcement learning for robots using neural networks, Technical report, Carnegie-Mellon Univ Pittsburgh PA School of Computer Science
1993
Earlier work this paper cites.
Simpson, A. & Eraser, N. M. (1993), Black box and glass box evaluation of the sundial system, in
1993
Earlier work this paper cites.
Aust, H., Oerder, M., Seide, F. & Steinbiss, V. (1995), ‘The Philips automatic train timetable information system’, Speech Communication
1995
Earlier work this paper cites.
Breiman, L. (1996), ‘Bagging predictors’, Machine learning
1996
Earlier work this paper cites.
Eckert, W., Levin, E. & Pieraccini, R. (1997), User modeling for spoken dialogue system evaluation, in
1997
Earlier work this paper cites.
Sutton, R. S. & Barto, A. G. (1998), Reinforcement learning: An introduction
1998
Earlier work this paper cites.
Ng, A. Y., Harada, D. & Russell, S. (1999), Policy invariance under reward transformations: Theory and application to reward shaping, in
1999
Earlier work this paper cites.
Singh, S. P., Kearns, M. J., Litman, D. J. & Walker, M. A. (1999), Reinforcement learning for spoken dialogue systems., in
1999
Earlier work this paper cites.
Levin, E., Pieraccini, R. & Eckert, W. (2000), ‘A stochastic model of human-machine interaction for learning dialog strategies’, IEEE Transactions on speech and audio processing
2000
Earlier work this paper cites.
Precup, D. (2000), ‘Eligibility traces for off-policy policy evaluation’, Computer Science Department Faculty Publication Series
2000
Earlier work this paper cites.
Stolcke, A., Ries, K., Coccaro, N., Shriberg, E., Bates, R., Jurafsky, D., Taylor, P., Martin, R., Van Ess-Dykema, C. & Meteer, M. (2000), ‘Dialogue act modeling for automatic tagging and recognition of conversational speech’, Computational linguistics
2000
Earlier work this paper cites.
Precup, D., Sutton, R. S. & Dasgupta, S. (2001), Off-policy temporal-difference learning with function approximation, in
2001
Earlier work this paper cites.
Singh, S., Litman, D., Kearns, M. & Walker, M. (2002), ‘Optimizing dialogue management with reinforcement learning: Experiments with the njfun system’, Journal of Artificial Intelligence Research
2002
Earlier work this paper cites.
Chung, G. (2004), Developing a flexible spoken dialog system using simulation, in
2004
Earlier work this paper cites.
Cuayáhuitl, H., Renals, S., Lemon, O. & Shimodaira, H. (2005), Human-computer dialogue simulation using hidden markov models, in
2005
Earlier work this paper cites.
Ward, N. G., Rivera, A. G., Ward, K. & Novick, D. G. (2005), ‘Root causes of lost time and user stress in a simple dialog system’
2005
Earlier work this paper cites.
Georgila, K., Henderson, J. & Lemon, O. (2006), User simulation for spoken dialogue systems: Learning and evaluation, in
2006
Earlier work this paper cites.
Paek, T. (2006), Reinforcement learning for spoken dialogue systems: Comparing strengths and weaknesses for practical deployment, in
2006
Earlier work this paper cites.
Raux, A., Bohus, D., Langner, B., Black, A. W. & Eskenazi, M. (2006), Doing research on a deployed spoken dialogue system: one year of let’s go! experience., in
2006
Earlier work this paper cites.
Bohus, D., Raux, A., Harris, T. K., Eskenazi, M. & Rudnicky, A. I. (2007), Olympus: an open-source framework for conversational spoken language interface research, in
2007
Earlier work this paper cites.
Schatzmann, J., Thomson, B., Weilhammer, K., Ye, H. & Young, S. (2007), Agenda-based user simulation for bootstrapping a pomdp dialogue system, in
2007
Earlier work this paper cites.
Shawar, B. A. & Atwell, E. (2007), Chatbots: are they really useful?, in
2007
Earlier work this paper cites.
Williams, J. D. & Young, S. (2007), ‘Partially observable markov decision processes for spoken dialog systems’, Computer Speech & Language
2007
Earlier work this paper cites.
Henderson, J., Lemon, O. & Georgila, K. (2008), ‘Hybrid reinforcement/supervised learning of dialogue policies from fixed data sets’, Computational Linguistics
2008
Earlier work this paper cites.
Traum, D., Marsella, S. C., Gratch, J., Lee, J. & Hartholt, A. (2008), Multi-party, multi-issue, multi-strategy negotiation for multi-modal virtual agents, in
2008
Earlier work this paper cites.
Bird, S., Klein, E. & Loper, E. (2009), Natural Language Processing with Python
2009
Earlier work this paper cites.
Heeman, P. A. (2009), Representing the reinforcement learning state in a negotiation dialogue, in
2009
Earlier work this paper cites.
Koren, Y., Bell, R. & Volinsky, C. (2009), ‘Matrix factorization techniques for recommender systems’, Computer
2009
Cited alongside, same era.
Pieraccini, R., Suendermann, D., Dayanidhi, K. & Liscombe, J. (2009), Are we there yet? research in commercial spoken dialog systems, in
2009
Cited alongside, same era.
Wallace, R. S. (2009), ‘The anatomy of alice’, Parsing the Turing Test
2009
Cited alongside, same era.
Ferrucci, D., Brown, E., Chu-Carroll, J., Fan, J., Gondek, D., Kalyanpur, A. A., Lally, A., Murdock, J. W., Nyberg, E., Prager, J. et al. (2010), ‘Building Watson: An overview of the DeepQA project’, AI magazine
2010
Cited alongside, same era.
Nair, V. & Hinton, G. E. (2010), Rectified linear units improve restricted boltzmann machines, in
2010
Cited alongside, same era.
Asri, L. E., He, J. & Suleman, K. (2016), A sequence-to-sequence model for user simulation in spoken dialogue systems, in
2016
Later among the works it cites.
Charras, F., Duplessis, G. D., Letard, V., Ligozat, A.-L. & Rosset, S. (2016), Comparing system-response retrieval models for open-domain and casual conversational agent, in
2016
Later among the works it cites.
Fatemi, M., Asri, L. E., Schulz, H., He, J. & Suleman, K. (2016), Policy networks with two-stage training for dialogue systems, in
2016
Later among the works it cites.
Foerster, J., Assael, Y. M., de Freitas, N. & Whiteson, S. (2016), Learning to communicate with deep multi-agent reinforcement learning, in
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Georgila, K. & Traum, D. (2011), Reinforcement learning of argumentation dialogue policies in negotiation, in
2011
Cited alongside, same era.
Glorot, X., Bordes, A. & Bengio, Y. (2011), Deep sparse rectifier neural networks, in
2011
Cited alongside, same era.
Williams, J. D. (2011), An empirical evaluation of a statistical dialog system in public use, in
2011
Cited alongside, same era.
Lee, S. & Eskenazi, M. (2012), Pomdp-based let’s go system for spoken dialog challenge, in
2012
Cited alongside, same era.
Mikolov, T., Sutskever, I., Chen, K., Corrado, G. S. & Dean, J. (2013), Distributed representations of words and phrases and their compositionality, in
2013
Cited alongside, same era.
2013
Cited alongside, same era.
Socher, R., Perelygin, A., Wu, J. Y., Chuang, J., Manning, C. D., Ng, A. Y., Potts, C. et al. (2013), Recursive deep models for semantic compositionality over a sentiment treebank, in
2013
Cited alongside, same era.
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
Liu, C.-W., Lowe, R., Serban, I. V., Noseworthy, M., Charlin, L. & Pineau, J. (2016), How NOT to evaluate your dialogue system: An empirical study of unsupervised evaluation metrics for dialogue response generation, in
2016
Later among the works it cites.
López-Cózar, R. (2016), ‘Automatic creation of scenarios for evaluating spoken dialogue systems via user-simulation’, Knowledge-Based Systems
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
Serban, I. V., Lowe, R., Charlin, L. & Pineau, J. (2016), Generative deep neural networks for dialogue: A short review, in
2016
Later among the works it cites.
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M. et al. (2016), ‘Mastering the game of go with deep neural networks and tree search’, Nature
2016
Later among the works it cites.
2016
Later among the works it cites.
Sukhbaatar, S., Fergus, R. et al. (2016), Learning multiagent communication with backpropagation, in
2016
Later among the works it cites.
Williams, J. D., Raux, A. & Henderson, M. (2016), ‘Introduction to the special issue on dialogue state tracking’, Dialogue & Discourse
2016
Later among the works it cites.
2016
Later among the works it cites.
Yu, Z., Xu, Z., Black, A. W. & Rudnicky, A. I. (2016), Strategy and policy learning for non-task-oriented conversational systems., in
2016
Later among the works it cites.
Zhao, T. & Eskenazi, M. (2016), Towards end-to-end learning for dialog state tracking and management using deep reinforcement learning, in
2016
Later among the works it cites.
Zhao, T., Lee, K. & Eskenazi, M. (2016), Dialport: Connecting the spoken dialog research community to real user data, in
2016
Later among the works it cites.
Das, A., Kottur, S., Moura, J. M., Lee, S. & Batra, D. (2017), Learning cooperative visual dialog agents with deep reinforcement learning, in
2017
Closest in time.
Im, J. (2017). http://search.aifounded.com/
2017
Closest in time.
Khouzaimi, H., Laroche, R. & Lefevre, F. (2017), Incremental human-machine dialogue simulation, in
2017
Closest in time.
Lewis, M., Yarats, D., Dauphin, Y. N., Parikh, D. & Batra, D. (2017), Deal or No Deal? End-to-End Learning for Negotiation Dialogues, in
2017
Closest in time.
Liu, B. & Lane, I. (2017), Iterative policy learning in end-to-end trainable task-oriented neural dialog models, in
2017
Closest in time.
Lowe, R., Noseworthy, M., Serban, I. V., Angelard-Gontier, N., Bengio, Y. & Pineau, J. (2017), Towards an automatic Turing test: Learning to evaluate dialogue responses, in
2017
Closest in time.
Lowe, R. T., Pow, N., Serban, I. V., Charlin, L., Liu, C.-W. & Pineau, J. (2017), ‘Training end-to-end dialogue systems with the ubuntu dialogue corpus’, Dialogue & Discourse
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
Serban, I. V., Sordoni, A., Lowe, R., Charlin, L., Pineau, J., Courville, A. & Bengio, Y. (2017), A Hierarchical Latent Variable Encoder-Decoder Model for Generating Dialogues, in
2017
Closest in time.