Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Llama-3.2-3B-Instruct-IQ3_M-gguf – AI Model by zanish-labs | AlphaNeural AI
You can deploy this model and start earning money today!
zanish-labs
/
Llama-3.2-3B-Instruct-IQ3_M-gguf
like
0
gguf
llama
quantized
voco
text-generation
meta-llama/Llama-3.2-3B-Instruct
quantized
llama3.2
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Llama-3.2-3B-Instruct-IQ3_M-GGUF
GGUF-converted and quantized derivative of Meta's Llama 3.2 3B Instruct for on-device inference.
Base Model
Model
: Llama 3.2 3B Instruct
Provider
: Meta Platforms, Inc.
Source
:
https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct
Model Card
:
https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct
License
: Llama 3.2 Community License
Conversion Details
Converter
: llama.cpp convert_hf_to_gguf.py (upstream master)
Quantization
: IQ3_M
GGUF Size
: ~1.5 GB
Converted by
: Zanish Labs / Voco
Runtime
Compatible with llama.cpp (CPU/NEON)
Tested on Voco iOS app (STQ1_0 backend, PR #22836)
Attribution
This is a converted/quantized derivative. The original model was created by Meta Platforms, Inc. and is licensed under the Llama 3.2 Community License. See LICENSE and NOTICE files.
Links
Original model:
https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct
Llama 3.2 Community License:
https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/LICENSE
Acceptable Use Policy:
https://www.llama.com/llama-downloads/
llama.cpp:
https://github.com/ggml-org/llama.cpp