Please avoid testing OpenToM questions in OpenAI playground or places where the data might be used for LLM training.
OpenToM is a new benchmark for assessing LLMs' Neural Theory-of-Mind (N-ToM) with the following key features:
(1) longer and clearer narrative stories
(2) characters with explicit personality traits
(3) actions that are triggered by character intentions
(4) questions designed to challenge LLMs' capabilities of modeling characters' mental states of both the physical and… See the full description on the dataset page:
https://huggingface.co/datasets/SeacowX/OpenToM.