Deep reinforcement learning for dialogue generation
Original
Li, J · 2016
Later among the works it cites.
How not to evaluate your dialogue system: An empirical study of unsupervised evaluation metrics for dialogue response generation
Original
Liu, C.-W · 2016
Later among the works it cites.
Strategy and policy learning for non-task-oriented conversational systems
Yu, Z · 2016
Later among the works it cites.
Boosting the actor with dual critic
Original
Dai, B · 2017
Later among the works it cites.
Towards an automatic turing test: Learning to evaluate dialogue responses
Lowe, R · 2017
Later among the works it cites.
Parlai: A dialog research software platform
Original
Miller, A. H · 2017
Later among the works it cites.
Ruber: An unsupervised method for automatic evaluation of open-domain dialog systems
Original
Tao, C · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A · 2017
Later among the works it cites.
On landscape of lagrangian functions and stochastic search for constrained nonconvex optimization
Original
Chen, Z · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Original
Devlin, J · 2018
Later among the works it cites.
Breaking the curse of horizon: Infinite-horizon off-policy estimation
Liu, Q · 2018
Later among the works it cites.
Bootstrapping a neural conversational agent with dialogue self-play, crowdsourcing and on-line reinforcement learning
Shah, P · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Sutton, R. S · 2018
Later among the works it cites.
Airdialogue: An environment for goal-oriented dialogue research
Wei, W · 2018
Later among the works it cites.
Personalizing dialogue agents: I have a dog, do you have pets too?
Original
Zhang, S · 2018
Later among the works it cites.
Approximating interactive human evaluation with self-play for open-domain dialog systems
Ghandeharioun, A · 2019
Later among the works it cites.
Huggingface’s transformers: State-of-the-art natural language processing
Wolf, T · 2019
Later among the works it cites.
Towards optimal off-policy evaluation for reinforcement learning with marginalized importance sampling
Xie, T · 2019
Later among the works it cites.
Survey on evaluation methods for dialogue systems
Deriu, J · 2020
Later among the works it cites.
The second conversational intelligence challenge (convai2)
Dinan, E · 2020
Later among the works it cites.
Statistically efficient off-policy policy gradients
Kallus, N · 2020
Later among the works it cites.
Towards holistic and automatic evaluation of open-domain dialogue generation
Pang, B · 2020
Later among the works it cites.
Improving dialog evaluation with a multi-reference adversarial dataset and large scale pretraining
Sai, A. B · 2020
Later among the works it cites.
Asymptotically efficient off-policy evaluation for tabular reinforcement learning
Yin, M · 2020
Later among the works it cites.
uBLEU: Uncertainty-aware automatic evaluation method for open-domain dialogue systems
Yuma, T · 2020
Later among the works it cites.