Views
No views yet
google/gemma-4-31B-it.google/gemma-4-31B-itllama.cpp, then quantized
with llama-quantize.| File | Bits/param | Use case |
|---|---|---|
gemma-4-31B-it-Q4_K_M.gguf | ~4.5 | Smallest, modest VRAM |
gemma-4-31B-it-Q5_K_M.gguf | ~5.5 | Sweet spot |
gemma-4-31B-it-Q8_0.gguf | ~8 | Near-lossless |