Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
L3-8B-Lunar-Stheno-bnb-4bit – AI Model by huggingkot | AlphaNeural AI
You can deploy this model and start earning money today!
huggingkot
/
L3-8B-Lunar-Stheno-bnb-4bit
like
0
safetensors
HiroseKoichi/L3-8B-Lunar-Stheno
finetune
8-bit
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
This is a converted weight from
L3-8B-Lunar-Stheno
model in
unsloth 4-bit dynamic quant
using this
collab notebook
.
About this Conversion
This conversion uses
Unsloth
to load the model in
4-bit
format and force-save it in the same
4-bit
format.
How 4-bit Quantization Works
The actual
4-bit quantization
is handled by
BitsAndBytes (bnb)
, which works under
Torch
via
AutoGPTQ
or
BitsAndBytes
.
Unsloth
acts as a wrapper, simplifying and optimizing the process for better efficiency.
This allows for reduced memory usage and faster inference while keeping the model compact.