Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
gemma-4-E2B-it-Q4_K_M-gguf – AI Model by zanish-labs | AlphaNeural AI
You can deploy this model and start earning money today!
zanish-labs
/
gemma-4-E2B-it-Q4_K_M-gguf
like
0
gguf
gemma
quantized
voco
text-generation
google/gemma-4-E2B
quantized
apache-2.0
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
gemma-4-E2B-it-Q4_K_M-GGUF
GGUF-converted and quantized derivative of Google's Gemma 4 E2B Instruct for on-device inference.
Base Model
Model
: Gemma 4 E2B Instruct (2B parameters)
Provider
: Google
Source
:
https://huggingface.co/google/gemma-4-E2B-it
License
: Apache 2.0
Conversion Details
Converter
: llama.cpp convert_hf_to_gguf.py (upstream master)
Quantization
: Q4_K_M
GGUF Size
: ~3.2 GB
Converted by
: Zanish Labs / Voco
Runtime
Compatible with llama.cpp (CPU/NEON)
Tested on Voco iOS app (STQ1_0 backend, PR #22836)
Attribution
This is a converted/quantized derivative. The original model was created by Google. See LICENSE for full license text.
Links
Original model:
https://huggingface.co/google/gemma-4-E2B-it
llama.cpp:
https://github.com/ggml-org/llama.cpp