Views
No views yet
[!NOTE] Status (2026-07-07): no weights published yet. This repository currently contains only the model card — it marks a planned variant that has not been released. Follow the repo to be notified when files land.
majentik/* family navigation.| Device | VRAM | Recommendation |
|---|---|---|
| H100 / H200 | 80–141 GB | native |
| RTX 4090 | 24 GB | does not fit full precision — use 4-bit |
| RTX 5090 | 32 GB | native |
1# No re-quantization needed — use the upstream weights directly.
2huggingface-cli download Qwen/Qwen3.6-35B-A3B-FP8mlx_lm, so it is not
covered by the local MLX eval harness. For measured scores on this family, see the benchmark
tables on the sibling MLX variants (e.g.
qwen3.6-35b-a3b-mlx-mxfp4 and
qwen3.6-35b-a3b-mlx-nvfp4),
which were evaluated with mlx_lm.evaluate (arc_easy, hellaswag; limit 200).apache-2.0. Upstream license of the base model applies.