Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Llama-3.2-1B-ao-int8wo-gs128 – AI Model by medmekk | AlphaNeural AI
You can deploy this model and start earning money today!
medmekk
/
Llama-3.2-1B-ao-int8wo-gs128
like
0
pytorch
llama
medmekk/Llama-3.2-1B-ao-int8wo-gs128
quantized
torchao
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
medmekk/Llama-3.2-1B-ao-int8wo-gs128 (Quantized)
Description
This model is a quantized version of the original model
medmekk/Llama-3.2-1B-ao-int8wo-gs128
.
It's quantized using the TorchAO library using the
torchao-my-repo
space.
Quantization Details
Quantization Type
: int8_weight_only
Group Size
: 128