Views
No views yet
| Variant | NVFP4 Layers | What stays in BF16 | Size | Quality |
|---|---|---|---|---|
| Ultra | 60 | Attention + layers 0-4 & 25-29 | ~8.0 GB | ⭐⭐⭐⭐⭐ |
| Quality | 90 | All attention (qkv, out) | ~6.5 GB | ⭐⭐⭐ |
| Mixed | 180 | Refiners, embedders, final layer | ~4.5 GB | ⭐ |
| Full | 204 | Only critical embedders | ~3.5 GB | ⭐ |
Original BF16 model size: 12.3 GB
{layer}.weight: uint8 (2 FP4 values packed per byte){layer}.weight_scale: float8_e4m3fn, 2D (per-block scale, 16-element blocks){layer}.weight_scale_2: float32, scalar (per-tensor scale){layer}.input_scale: float32, scalar (activation scale)⚠️ NVFP4 requires specific hardware and software!
cu130) - older versions do not support NVFP4Model: z-image-base-nvfp4_[variant].safetensors
Steps: 28-50
CFG Scale: 3.0-5.0⚠️ Note: This is Z-Image Base, not Turbo. Use 28-50 steps with CFG guidance, not 8 steps like Turbo.