Exploring the Reliability of Large Language Models as Customized Evaluators for Diverse NLP Tasks.
Qintong Li, Leyang Cui, Lingpeng Kong, Wei Bi
Browse the full COLING paper archive.
Qintong Li, Leyang Cui, Lingpeng Kong, Wei Bi
Browse the full COLING paper archive.