This model is a fine-tuned version of
Qwen/Qwen3-32B on the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--laion--nemotron-terminal-scientific_computing-10pct/snapshots/4a53d51d88fca48b36813cea0ce5585396edc6c2_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--exp_tas_repetition_penalty_1.05_traces-10pct/snapshots/e805e930a7dba8cfd0e93d640beb3176190b5ebb_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--a1_multifile_composition-10pct/snapshots/20cc36403c186eb54888cc4227e265ab81302888_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--exp-gfi-staqc-embedding-mean-filtered-10K_glm_4.7_traces_jupiter-10pct/snapshots/ad07ef81b0a5b4a2a05d99752c8fa080a41f735b_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--exp_tas_max_episodes_512_traces-10pct/snapshots/b1dff6ac3a4dc55be7dda8668407b84c78b6d13e_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--dev_set_part1_10k_glm_4.7_traces_jupiter-10pct/snapshots/d74f778df2c249d3e9b55a6f00c308b170f6b1b2_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--a1_repo_scaffold-10pct/snapshots/8636834124732cab4f3b9d2dabdce0fb05dcf01e_thinking_preprocessed, the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--DCAgent--swesmith-sandboxes-with_tests-gpt-5-mini-passed_glm_4.7_traces-10pct/snapshots/f65f170739f6abf9ec30ee35fc391ca1c34d5c81_thinking_preprocessed and the /e/data1/datasets/playground/ot-baf/hf_hub/datasets--penfever--Kimi-2.5-r2egym_sandboxes-maxeps-32k-10pct/snapshots/3a1b1dcf779a7a4dd15329c6a3124a97a289e11d_thinking_preprocessed datasets.