math-ai-bench-sources-final is a per-source split of haowu89/math-ai-bench-sources-latest.
Each row keeps the original benchmark fields and adds:
context_metadata: token counts for generated_solutions, computed with tiktoken o200k_base
original_solution: copied over where an aligned subset existed in zechen-nlp/math_ai_bench_cot
aime25
aime26
apex_2025
arxivmath
cmimc_2025
gpqa_diamond
hmmt_feb_2026
hmmt_nov_2025
imobench… See the full description on the dataset page:
https://huggingface.co/datasets/haowu89/math-ai-bench-sources-final.