Math-verified-correct reasoning trajectories from the expandA pool. HET = heterogeneous 3x32B roster
(Qwen3-32B + DeepSeek-R1-Distill-Qwen-32B + OpenReasoning-Nemotron-32B, true token-level continuation);
HOM = homogeneous single Qwen3-4B. Verified with the OMEGA math verifier (verify.py) against
data/sft_main_prompts.jsonl ground truth; only accepted=true candidates are included here.
Solve rates: HET ~42-45%… See the full description on the dataset page:
https://huggingface.co/datasets/shizhuo2/omega-het-expandA-verified.