Views
No views yet
| Feature | Details |
|---|---|
| Architecture | Gemma 3 (12B) |
| Format | MLX (4-bit) |
| Parameters | 12.0B |
| Context Window | 131k (Optimized: 8192) |
| Metric | Speed |
|---|---|
| Prompt Processing | ~140 t/s |
| Token Generation | ~25-45 t/s |
| Model | Type | Note |
|---|---|---|
| Gemma 3 12B IT | SOTA Reasoning | Official Base |
| Council Ultima | Creative / RP | Prioritizes creative freedom & character adherence. |
Disclaimer: This model is minimally filtered and may produce unsafe, offensive, or illegal content. Use at your own risk and always follow local laws and platform terms of service.
1pip install mlx-lm
2python -m mlx_lm.generate --model MtnMCG/Council-Ultima-Gemma-3-12B-MLX-4bit --prompt "Hello Council."