Views
No views yet
reasoning_effort parameter| Component | Bits | Group Size | Format |
|---|---|---|---|
| MLP Experts | 4 | 32 | MXFP4 |
| Attention | - | - | Full precision (bfloat16) |
| Routers | - | - | Full precision (bfloat16) |
| Embeddings | - | - | Full precision (bfloat16) |
| LM Head | - | - | Full precision (bfloat16) |
mlx_lm.chat --model txgsync/gpt-oss-120b-Derestricted-mxfp4-mlxlms get txgsync/gpt-oss-120b-derestrictedmodules_to_not_convert scheme for optimal quality