Skip to content

How NOT To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation.

Chia-Wei Liu, Ryan Lowe, Iulian Serban, Michael Noseworthy, Laurent Charlin, Joelle Pineau

VenueA*EMNLP
Year2016
ProceedingsEMNLP

Browse the full EMNLP paper archive.