Views
No views yet
| Property | Value |
|---|---|
| Base model | google/gemma-4-12B-it |
| Quant method | NVIDIA ModelOpt (NVFP4) |
| Weight scheme | 4-bit float, block size 16 |
| Input activation | 4-bit float, block size 16 |
| Calibration dataset | CNN DailyMail (512 samples, max_seq_len 1024) |
| Size | ~11 GB (vs ~23 GB BF16) |
lm_headmodel.embed_vision*model.embed_audio*self_attn layers (layers 0–47)