Gemma 2B IT GGUF Model
This repository contains a quantized version of the Gemma 2B IT model in GGUF format.
Model Details
- Base Model: Gemma 2B IT
- Quantization: Q4_K_M
- Format: GGUF
- Usage: Optimized for inference with llama.cpp
Usage
This model can be used with llama.cpp or its Python bindings for efficient inference on CPU.