What happens if you treat ordinal ratings as interval data? Human evaluations in NLP are even more under-powered than you think.
David M. Howcroft, Verena Rieser
Browse the full EMNLP paper archive.
David M. Howcroft, Verena Rieser
Browse the full EMNLP paper archive.