Skip to content

OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models.

Hainiu Xu, Runcong Zhao, Lixing Zhu, Jinhua Du, Yulan He

VenueA*ACL
Year2024
ProceedingsACL (1)

Browse the full ACL paper archive.