Matched supervised-fine-tuning data for the OMEGA diversity experiment: for each math prompt,
reasoning trajectories are sampled two ways and only prompts solved (math-verified correct) in both
conditions are kept (matched HOM∩HET = 3,219 prompts), so HET and HOM are directly comparable.
HET (heterogeneous): true token-level continuation across a 3×32B roster
(Qwen3-32B + DeepSeek-R1-Distill-Qwen-32B +… See the full description on the dataset page:
https://huggingface.co/datasets/shizhuo2/omega-het-expandA-sft.