This contains binary_correctness.json, for recreating the results of the paper titled "The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors".
This file includes outputs from GPT-5-mini labeling whether student is correct/incorrect on binary error & correctness questions, from DrawEduMath.
Please consult the datacard for DrawEduMath for detailed information about data source.
Quick links:… See the full description on the dataset page:
https://huggingface.co/datasets/lucy3/aftermath_binary_correctness.