Views
No views yet
downcast_to_mxfp from triton-kernels.| Model | GSM8K (strict-match) | GSM8K (flexible-extract) |
|---|---|---|
| Qwen3-Coder-30B-A3B-Instruct (BF16) | 90.67% ± 0.80% | 89.92% ± 0.83% |
| Qwen3-Coder-30B-A3B-Instruct_MXFP4 | 89.76% ± 0.83% | 88.70% ± 0.87% |
| Model | Size | Reduction |
|---|---|---|
| Qwen3-Coder-30B-A3B-Instruct (BF16) | 57 GB | - |
| Qwen3-Coder-30B-A3B-Instruct_MXFP4 | 18 GB | 68% smaller |