← Search

Iulian V. Serban

1 accepted papers

2017

Towards an automatic Turing test: Learning to evaluate dialogue responses

ICLR 2017workshop

Automatically evaluating the quality of dialogue responses for unstructured domains is a challenging problem. Unfortunately, existing automatic evaluation metrics are biased and correlate very poorly with human judgements of response quality (Liu et al., 2016). Yet having an accurate automatic evalu…

Cited by 453SourcecodeScholar