The Children’s Book Test (CBT) is designed to measure directly how well language models can exploit wider linguistic context. The CBT is built from books that are freely available.
This dataset contains four different configurations:
V: where the answers to the questions are verbs.
P: where the answers to the questions are pronouns.
NE: where the answers to the questions are named entities.
CN: where the answers to the questions are… See the full description on the dataset page: https://huggingface.co/datasets/cam-cst/cbt.