This dataset is a 29,963 sample subset of marin-community/open-thoughts-4-code-qwen3-32b-annotated, originally derived from mlfoundations-dev/hero_run_4_code curated by the OpenThoughts4 team.
We provide the responses from Qwen/Qwen3-32B in the generated_text column. These samples were generated using temperature = 0.8 and max output tokens = 7500.Note that many of the… See the full description on the dataset page:
https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-code-qwen3-32b-annotated.