Skip to content

Exploring the Reliability of Large Language Models as Customized Evaluators for Diverse NLP Tasks.

Qintong Li, Leyang Cui, Lingpeng Kong, Wei Bi

VenueBCOLING
Year2025
ProceedingsCOLING

Browse the full COLING paper archive.