Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Apertus-8B-Instruct-2509-bnb-8bit – AI Model by jgerster0 | AlphaNeural AI
You can deploy this model and start earning money today!
jgerster0
/
Apertus-8B-Instruct-2509-bnb-8bit
like
0
safetensors
apertus
quantization
llm
swissai
8-bit
bitsandbytes
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Apertus-8B-Instruct-2509-bnb-8bit
This is an INT8 dynamically quantized version of
swiss-ai/Apertus-8B-Instruct-2509
using
llm-compressor
.
This version used the
fineweb-edu-score-2
dataset for calibration.
Quantization Details
Quantization Scheme
: W8A8
Method
: Dynamic quantization of weights and activations to INT8 (W8A8) format
Targets
: All Linear layers
Ignored Layers
:
lm_head
(kept in higher precision for better output quality)
Tool
: llm-compressor