A supervised fine-tuning (SFT) dataset of math problems with full chain-of-thought solutions,
formatted for the Olmo 3 "Thinking" models.
2,813,055 examples · ~37.9 B tokens.
Three task families: proofs, numeric-answer problems, and tool-augmented (Python) problems.
Every assistant turn carries an explicit
… reasoning trace before the answer.
Olmo 3 native chat + function-calling format; every example fits within a 64k-token context.… See the full description on the dataset page:
https://huggingface.co/datasets/chankhavu/smolmo-sft-v2-seqlen64k.