TY - RPRT TI - Beyond User Self-Reported Likert Scale Ratings: A Comparison Model for Automatic Dialog Evaluation AU - Weixin Liang AU - James Zou AU - Zhou Yu PY - 2020 UR - https://arxiv.org/abs/2005.10716 ID - 2005.10716 ER -