This repo offers quantized versions of
google/gemma-3-4b-it for use with llama.cpp.
Quantization was done using
an unofficial Docker image
and calibrated on 100 rows from the
agentlans/LinguaNova dataset to maintain coherence and multilingual support.
The importance matrix file is included.