Skip to content

What happens if you treat ordinal ratings as interval data? Human evaluations in NLP are even more under-powered than you think.

David M. Howcroft, Verena Rieser

VenueA*EMNLP
Year2021
ProceedingsEMNLP (1)

Browse the full EMNLP paper archive.