Views
No views yet


| Metric | Value |
|---|---|
| Source | google/gemma-4-31b-it |
| Architecture | Dense Transformer + Hybrid Sliding/Global Attention |
| Profile | JANG_4M (CRITICAL=8-bit, COMPRESS=4-bit) |
| Actual avg bits | 5.1 |
| Model size | 18 GB |
| Vision | Yes (multimodal, float16 passthrough) |
| Parameters | 31B |
| Format | JANG v2 (MLX-native safetensors, instant load) |
| Abliteration | CRACK (refusal removal) |
All benchmarks below were measured with reasoning/thinking DISABLED. With thinking enabled, compliance rates are expected to be significantly higher as the model reasons through the request before responding. These scores represent the conservative lower bound.
| Subject | JANG_4M | CRACK |
|---|---|---|
| Abstract Algebra | 13/20 | 14/20 |
| Anatomy | 13/20 | 10/20 |
| Astronomy | 17/20 | 17/20 |
| College CS | 14/20 | 13/20 |
| College Physics | 14/20 | 13/20 |
| HS Biology | 19/20 | 19/20 |
| HS Chemistry | 15/20 | 15/20 |
| HS Mathematics | 9/20 | 9/20 |
| Logical Fallacies | 19/20 | 19/20 |
| World Religions | 20/20 | 20/20 |
| Total | 153/200 (76.5%) | 149/200 (74.5%) |
| Tier | Components | Bits |
|---|---|---|
| CRITICAL | Attention (Q/K/V/O), embeddings | 8 |
| COMPRESS | MLP (gate, up, down proj), remaining weights | 4 |
| Model | Type | Size | MMLU | Comply | HarmBench |
|---|---|---|---|---|---|
| JANG_4M CRACK (this) | Dense 31B | 18 GB | 74.5% | 8/8 | 93.7% |
| JANG_4M CRACK | MoE 26B | 15 GB | 67.5% | 8/8 | 86.8% |
| JANG_2L CRACK | MoE 26B | 9.9 GB | 58.5% | 8/8 | 98.7% |
Important: Standardmlx_lmandmlx_vlmdo NOT support Gemma 4 as of v0.31.2 / v0.4.1. You need vMLX 1.3.26+ which includes bundled Gemma 4 support.
1# vMLX (recommended)
2# Load directly in vMLX app or via API
3
4# Manual MLX loading
5from mlx_vlm.models.gemma4 import Model
6# Requires mlx_vlm with gemma4 support (vMLX bundled version)
