MLX conversion of ressl/gemma-4-31B-it-uncensored for Apple silicon. This release uses 6-bit affine quantization with group size 64 and preserves the multimodal vision path in BF16.
[!WARNING]
This is an uncensored, abliterated research model. It can produce inaccurate, unsafe, illegal, biased, or offensive content. Outputs are not advice. Evaluate the model for your use case, keep human oversight, and comply with applicable law and the Apache License 2.0.
Security research, red teaming, robustness evaluation, and defensive model analysis are the intended uses. Do not use this release to harm people or systems.
The performance values above are measurements from the release smoke run. They are not estimates or cross-device promises.
Refusal evaluation
The release gate completed all 686 prompts with zero execution errors. A naive keyword heuristic detected broad refusal-like language, while the stricter hard-refusal detector found 0/686 responses that refused without providing substantive help.
Dataset
Successful
Errors
Naive refusals
Hard refusals
JailbreakBench
100/100
0
75
0
tulu-harmbench
320/320
0
134
0
NousResearch
166/166
0
110
0
mlabonne
100/100
0
82
0
Total
686/686
0
401
0
No hard refusals were detected, so no refusal review was triggered.
Keyword metrics are imperfect and do not prove capability, factuality, or safety. The exact deterministic evaluator and release criteria live in the source repository.
The language model weights use 6-bit affine quantization with group size 64. Quantization can reduce quality compared with BF16.
The vision tower and vision embedding path remain BF16 in every MLX release.
The smoke suite verifies deterministic arithmetic, capitals, German, multi-turn memory, thinking mode, image understanding, basic output health, and measured runtime statistics.
This checkpoint inherits the source model's limitations and may hallucinate or follow malicious instructions.
No benchmark result should be generalized beyond the exact prompts, software, and hardware used for the measured run.
Provenance
The checkpoint was converted from commit 64c863e92fc131e5f4b0fe3631a0791fe2c19152. Quantized releases use affine MLX quantization with group size 64 for language weights. Modules whose path contains vision_tower or embed_vision are excluded from quantization. The BF16 release performs no weight quantization.
The original Gemma architecture and weights are provided by Google under the Apache License 2.0. This MLX conversion and the uncensored source release are maintained by Robert Ressl (Hugging Face, Website, LinkedIn, Patreon).
If this work is useful, you can support continued independent model research on Patreon.