Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Web-Petals-SmolLM2-1.7B-q8 – AI Model by powerpudu | AlphaNeural AI
You can deploy this model and start earning money today!
powerpudu
/
Web-Petals-SmolLM2-1.7B-q8
like
0
onnx
web-petals
distributed-inference
smollm2
int8
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Web-Petals SmolLM2-1.7B ONNX Layers (QInt8)
SmolLM2-1.7B-Instruct split into individual ONNX layers and
dynamically quantized to INT8
for distributed P2P inference in WebGPU/WASM.
Total Size
: ~1793.4 MB
Layer Size
: ~64 MB
Precision
: INT8 (weights only)