Fetching the paper…

Beyond User Self-Reported Likert Scale Ratings: A Comparison Model for Automatic Dialog Evaluation · Around