A production-style synthetic dataset for reasoning-focused LLM/SLM training.
This dataset provides 1,000 instruction-response samples with explicit thinking_cot traces for multi-domain supervision:
Basic arithmetic reasoning (200)
Advanced math reasoning (200)
General chat + explanation behavior (400)
Coding and snippet generation tasks (200)
Deterministic numeric… See the full description on the dataset page:
https://huggingface.co/datasets/Sharjeelbaig/thinking-cot-1k.