[!Note]
This repository corresponds to the 12B instruction-tuned version of the Gemma 3 model using Quantization Aware Training (QAT).
The checkpoint in this repository is unquantized, please make sure to quantize with Q4_0 with your favorite tool
Thanks to QAT, the model is able to preserve similar quality as bfloat16 while significantly reducing the memory requirements
to load the model.
[Gemma 3… See the full description on the dataset page:
https://huggingface.co/datasets/honmamon/g3.