Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation.
Noy Sternlicht, Ariel Gera, Roy Bar-Haim, Tom Hope, Noam Slonim
Browse the full EMNLP paper archive.
Noy Sternlicht, Ariel Gera, Roy Bar-Haim, Tom Hope, Noam Slonim
Browse the full EMNLP paper archive.