This model is a fine-tuned version of
Qwen/Qwen3-8B on the /e/scratch/jureap59/raoof1/sft_data/hf_hub/datasets--DCAgent--g1_min_episodes_e1_gpt_long_d1_original_40k_glm47_traces_3160/snapshots/8b28e56fb925489a4a5a61f5dd2ce2689e5d81b3_thinking_preprocessed dataset.