This dataset contains 48 parallel multimodal samples (paired TXT↔PDF) derived from PISA studies up to 2012, plus 47 non-parallel samples (TXT-only or PDF-only). Each sample may include multiple questions. Content is available in German and English.
Source & usage: Materials are published by the OECD and are provided here for non-commercial use only. Please verify that your usage complies with OECD terms.… See the full description on the dataset page: https://huggingface.co/datasets/barthfab/PISA_tests.