TY - RPRT TI - OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models AU - Hainiu Xu AU - Runcong Zhao AU - Lixing Zhu AU - Jinhua Du AU - Yulan He PY - 2024 UR - https://arxiv.org/abs/2402.06044 ID - 2402.06044 ER -