Understand
Automatically evaluating the quality of dialogue responses for unstructured domains is a challenging problem.
- ADEM(Lowe et al.
- 2017) formulated the automatic evaluation of dialogue systems as a learning problem and showed that such a model was able to predict responses which correlate significantly with human judgements, both at utterance and system level.
- Their system was shown to have beaten word-overlap metrics such as BLEU with large margins.
Built on
Nothing clear enough to list yet.
Similar
Nothing clear enough to list yet.
Then
Nothing clear enough to list yet.
Beyond the bibliography
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…