Fetching the paper…
Reading the bibliography…
We propose an approach towards natural language generation using a bidirectional encoder-decoder which incorporates external rewards through reinforcement learning (RL).
1907
Earlier work this paper cites.
Hochreiter, S. and Schmidhuber, J., 1997. Long short-term memory. Neural computation, 9(8), pp.1735-1780
1997
Earlier work this paper cites.
Papineni, K., Roukos, S., Ward, T. and Zhu, W.J., 2002, July. BLEU: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting on association for computational linguistics (pp. 311-318). Association for Computational Linguistics
2002
Earlier work this paper cites.
Mintz, M., Bills, S., Snow, R. and Jurafsky, D., 2009, August. Distant supervision for relation extraction without labeled data. In Proceedings of the Joint Conference of the 47th Annual Meeting of the ACL and the 4th International Joint Conference on Natural Language Processing of the AFNLP: Volume 2-Volume 2 (pp. 1003-1011). Association for Computational Linguistics
2009
Earlier work this paper cites.
Danescu-Niculescu-Mizil, C. and Lee, L., 2011, June. Chameleons in imagined conversations: A new approach to understanding coordination of linguistic style in dialogs. In Proceedings of the 2nd workshop on cognitive modeling and computational linguistics (pp. 76-87). Association for Computational Linguistics
2011
Earlier work this paper cites.
Warriner, A.B., Kuperman, V. and Brysbaert, M., 2013. Norms of valence, arousal, and dominance for 13,915 English lemmas. Behavior research methods, 45(4), pp.1191-1207
2013
Earlier work this paper cites.
Mikolov, T., Sutskever, I., Chen, K., Corrado, G.S. and Dean, J., 2013. Distributed representations of words and phrases and their compositionality. In Advances in neural information processing systems (pp. 3111-3119)
2013
Earlier work this paper cites.
Sequeira, P., Melo, F.S. and Paiva, A., 2014. Learning by appraising: an emotion-based approach to intrinsic reward design. Adaptive Behavior, 22(5), pp.330-349
2014
Earlier work this paper cites.
2014
Cited alongside, same era.
Lowe, R., Pow, N., Serban, I. and Pineau, J., 2015, September. The Ubuntu Dialogue Corpus: A Large Dataset for Research in Unstructured Multi-Turn Dialogue Systems. In Proceedings of the 16th Annual Meeting of the Special Interest Group on Discourse and Dialogue (pp. 285-294)
2015
Cited alongside, same era.
Vinyals, O. and Le, Q., 2015. A neural conversational model. arXiv preprint arXiv:1506.05869
2015
Cited alongside, same era.
Li, J., Monroe, W., Ritter, A., Jurafsky, D., Galley, M. and Gao, J., 2016, November. Deep Reinforcement Learning for Dialogue Generation. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing (pp. 1192-1202)
2016
Cited alongside, same era.
Li, J., Monroe, W., Shi, T., Jean, S., Ritter, A. and Jurafsky, D., 2017, September. Adversarial Learning for Neural Dialogue Generation. In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing (pp. 2157-2169)
2017
Later among the works it cites.
Chen, H., Liu, X., Yin, D. and Tang, J., 2017. A survey on dialogue systems: Recent advances and new frontiers. Acm Sigkdd Explorations Newsletter, 19(2), pp.25-35
2017
Later among the works it cites.
Niu, T. and Bansal, M., 2018. Polite dialogue generation without parallel data. Transactions of the Association for Computational Linguistics, 6, pp.373-389
2018
Later among the works it cites.
Guo, J., Lu, S., Cai, H., Zhang, W., Yu, Y. and Wang, J., 2018, April. Long text generation via adversarial training with leaked information. In Thirty-Second AAAI Conference on Artificial Intelligence
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Badoy Jr, W. and Teknomo, K., 2016. Q-learning with basic emotions. arXiv preprint arXiv:1609.01468
2016
Cited alongside, same era.
Li, J., Galley, M., Brockett, C., Gao, J. and Dolan, B., 2016. A Diversity-Promoting Objective Function for Neural Conversation Models. In Proceedings of NAACL-HLT (pp. 110-119)
2016
Cited alongside, same era.
Christiano, P.F., Leike, J., Brown, T., Martic, M., Legg, S. and Amodei, D., 2017. Deep reinforcement learning from human preferences. In Advances in Neural Information Processing Systems (pp. 4299-4307)
2017
Cited alongside, same era.
Jaques, N., Gu, S., Turner, R.E. and Eck, D., 2017. Tuning recurrent neural networks with reinforcement learning
2017
Cited alongside, same era.
Asghar, N., Poupart, P., Hoey, J., Jiang, X. and Mou, L., 2018, March. Affective neural response generation. In European Conference on Information Retrieval (pp. 154-166). Springer, Cham
2018
Later among the works it cites.
Zhou, H., Huang, M., Zhang, T., Zhu, X. and Liu, B., 2018, April. Emotional chatting machine: Emotional conversation generation with internal and external memory. In Thirty-Second AAAI Conference on Artificial Intelligence
2018
Later among the works it cites.
Moerland, T.M., Broekens, J. and Jonker, C.M., 2018. Emotion in reinforcement learning agents and robots: a survey. Machine Learning, 107(2), pp.443-480
2018
Later among the works it cites.