Fetching the paper…
Reading the bibliography…
Recently, two approaches, fine-tuning large pre-trained language models and variational training, have attracted significant interests, separately, for semi-supervised end-to-end task-oriented dialog (TOD) systems.
L. Gillick and S. J. Cox, “Some statistical issues in the comparison of speech recognition algorithms,” in International Conference on Acoustics, Speech, and Signal Processing, . IEEE, 1989, pp. 532–535
1989
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3-4, pp. 229–256, 1992
1992
Earlier work this paper cites.
X. Zhu, “Semi-supervised learning literature survey,” Technical report, University of Wisconsin-Madison , 2006
2006
Earlier work this paper cites.
2013
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in Proc. of Advances in neural information processing systems , 2014
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in 2nd International Conference on Learning Representations (ICLR) , Y. Bengio and Y. LeCun, Eds., 2014
2014
Earlier work this paper cites.
D. J. Rezende, S. Mohamed, and D. Wierstra, “Stochastic backpropagation and approximate inference in deep generative models,” in ICML , 2014
2014
Earlier work this paper cites.
N. Mrkšić, D. Ó. Séaghdha, T.-H. Wen, B. Thomson, and S. Young, “Neural belief tracker: Data-driven dialogue state tracking,” in Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (ACL) , 2017
2017
Earlier work this paper cites.
T. Wen, Y. Miao, P. Blunsom, and S. J. Young, “Latent intention dialogue models,” in Proceedings of the 34th International Conference on Machine Learning (ICML) , D. Precup and Y. W. Teh, Eds., 2017
2017
Earlier work this paper cites.
T.-H. Wen, D. Vandyke, N. Mrkšić, M. Gasic, L. M. R. Barahona, P.-H. Su, S. Ultes, and S. Young, “A network-based end-to-end trainable task-oriented dialogue system,” in Proceedings of the 15th Conference of the European Chapter of the Association for Computational Linguistics , 2017
2017
Earlier work this paper cites.
B. Liu and I. Lane, “An end-to-end trainable neural network model with belief tracking for task-oriented dialog,” Proc. Interspeech 2017 , pp. 2506–2510, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in Proc. of Advances in neural information processing systems , 2017
2017
Earlier work this paper cites.
E. Jang, S. Gu, and B. Poole, “Categorical reparameterization with gumbel-softmax,” in International Conference on Learning Representations (ICLR) , 2017
2017
Earlier work this paper cites.
W. Lei, X. Jin, M.-Y. Kan, Z. Ren, X. He, and D. Yin, “Sequicity: Simplifying task-oriented dialogue systems with single sequence-to-sequence architectures,” in 56th Annual Meeting of the Association for Computational Linguistics (ACL) , 2018
2018
Cited alongside, same era.
X. Jin, W. Lei, Z. Ren, H. Chen, S. Liang, Y. Zhao, and D. Yin, “Explicit state tracking with semi-supervision for neural dialogue generation,” in Proceedings of the 27th ACM International Conference on Information and Knowledge Management (CIKM) , 2018
2018
Cited alongside, same era.
A. Radford, K. Narasimhan, T. Salimans, and I. Sutskever, “Improving language understanding by generative pre-training,” 2018, http://openai-assets.s3.amazonaws.com/research-covers/language-unsupervised/language_understanding_paper.pdf
2018
Cited alongside, same era.
Y. Zhang, Z. Ou, M. Hu, and J. Feng, “A probabilistic end-to-end task-oriented dialog model with latent belief states towards semi-supervised learning,” in Proc. of the Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2020
2020
Later among the works it cites.
M. Heck, C. van Niekerk, N. Lubis, C. Geishauser, H.-C. Lin, M. Moresi, and M. Gasic, “Trippy: A triple copy strategy for value independent neural dialog state tracking,” in Proceedings of the 21th Annual Meeting of the Special Interest Group on Discourse and Dialogue , 2020
2020
Later among the works it cites.
D. Ham, J.-G. Lee, Y. Jang, and K.-E. Kim, “End-to-end neural pipeline for goal-oriented dialogue systems using GPT-2,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (ACL) , 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
P. Budzianowski, T.-H. Wen, B.-H. Tseng, I. Casanueva, U. Stefan, R. Osman, and M. Gašić, “MultiWOZ - a large-scale multi-domain wizard-of-oz dataset for task-oriented dialogue modelling,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2018
2018
Cited alongside, same era.
L. Shu, P. Molino, M. Namazifar, H. Xu, B. Liu, H. Zheng, and G. Tür, “Flexibly-structured model for task-oriented dialogues,” in Proceedings of the 20th Annual SIGdial Meeting on Discourse and Dialogue , 2019
2019
Cited alongside, same era.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Proc. of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , 2019
2019
Cited alongside, same era.
P. Budzianowski and I. Vulić, “Hello, it’s GPT-2 - how can I help you? towards the use of pretrained language models for task-oriented dialogue systems,” in Proceedings of the 3rd Workshop on Neural Generation and Translation . Association for Computational Linguistics, 2019
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” OpenAI Blog , vol. 1, no. 8, p. 9, 2019
2019
Cited alongside, same era.
T. Zhao, K. Xie, and M. Eskenazi, “Rethinking action spaces for reinforcement learning in end-to-end dialog agents with latent variable models,” in Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL-HLT) , 2019
2019
Cited alongside, same era.
S. Mehri, T. Srinivasan, and M. Eskenazi, “Structured fusion networks for dialog,” in Proceedings of the 20th Annual SIGdial Meeting on Discourse and Dialogue , 2019
2019
Cited alongside, same era.
Y. Zhang, Z. Ou, and Z. Yu, “Task-oriented dialog systems that consider multiple appropriate responses under the same context,” in The Thirty-Fourth AAAI Conference on Artificial Intelligence (AAAI) , 2020
2020
Cited alongside, same era.
2020
Later among the works it cites.
B. P. C. Li, J. Li, S. Shayandeh, L. Liden, and J. Gao, “SOLOIST: Building task bots at scale with transfer learning and machine teaching,” Transactions of the Association for Computational Linguistics (TACL), 2021 , 2020
2020
Later among the works it cites.
M. Eric, R. Goel, S. Paul, A. Sethi, S. Agarwal, S. Gao, A. Kumar, A. K. Goyal, P. Ku, and D. Hakkani-Tür, “MultiWOZ 2.1: A consolidated multi-domain dialogue dataset with state corrections and state tracking baselines,” in LREC , 2020
2020
Later among the works it cites.
Q. Zhu, K. Huang, Z. Zhang, X. Zhu, and M. Huang, “CrossWOZ: A large-scale chinese cross-domain task-oriented dialogue dataset,” Transactions of the Association for Computational Linguistics , vol. 8, pp. 281–295, 2020
2020
Later among the works it cites.
B. Kim, J. Ahn, and G. Kim, “Sequential latent knowledge selection for knowledge-grounded dialogue,” in International Conference on Learning Representations (ICLR) , 2020
2020
Later among the works it cites.
N. Lubis, C. Geishauser, M. Heck, H.-C. Lin, M. Moresi, C. van Niekerk, and M. Gasic, “LAVA: Latent action spaces via variational auto-encoding for dialogue policy optimization,” in Proceedings of the 28th International Conference on Computational Linguistics , 2020, pp. 465–479
2020
Later among the works it cites.
S. Bao, H. He, F. Wang, H. Wu, and H. Wang, “PLATO: Pre-trained dialogue generation model with discrete latent variable,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (ACL) , 2020
2020
Later among the works it cites.
2021
Closest in time.
Y. Yang, Y. Li, and X. Quan, “UBAR: Towards fully end-to-end task-oriented dialog system with gpt-2,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , 2021
2021
Closest in time.