This model is a fine-tuned version of
secmlr/DS-Clean_QWQ-Clean_Qwen2.5-7B-Instruct_full_sft_1e-5 on the ruizhe_simplier_reasoning_ds_clean_32k, the ruizhe_simplier_reasoning_ds_noisy_32k, the ruizhe_simplier_reasoning_qwq_clean_32k and the ruizhe_simplier_reasoning_qwq_noisy_small_32k datasets.