Views
No views yet
GLM-4-32B-0414 (instruct) and GLM-Z1-32B-0414 (reasoning) at t=0.3 (~70% instruct / 30% reasoning). No Korean fine-tuning — this is an experimental base artifact, shared transparently.| Model | GSM8K (5-shot) | n | stderr |
|---|---|---|---|
| This merge (SLERP t=0.3) | 0.940 | 300 | +/-0.014 |
| GLM-4-32B-0414 (instruct parent) | 0.883 | 300 | +/-0.019 |
| GLM-Z1-32B-0414 (reasoning parent) | 0.43* | 100 | +/-0.05 |
1merge_method: slerp
2base_model: zai-org/GLM-4-32B-0414
3models:
4 - model: zai-org/GLM-Z1-32B-0414
5parameters:
6 t: 0.3
7dtype: bfloat16dare_ties / ties with a base) collapsed this instruct+reasoning pair into incoherent multilingual/code output; only SLERP (direct interpolation, no base subtraction) yielded a coherent model.