Views
No views yet
SicariusSicariiStuff/X-Ray_Alpha using llama.cpp via the ggml.ai's all-gguf-same-where space.
Refer to the original model card for more details on the model.Q4_K_M (Best balance of speed/quality)Q4_0 (Optimized for ARM CPUs)Q8_0 (Near-original quality)| 🚀 Download | 🔢 Type | 📝 Notes |
|---|---|---|
| Download | Basic quantization | |
| Download | Small size | |
| Download | Balanced quality | |
| Download | Better quality | |
| Download | Fast on ARM | |
| Download | Fast, recommended | |
| Download | Best balance | |
| Download | Good quality | |
| Download | Balanced | |
| Download | High quality | |
| Download | Very good quality | |
| Download | Fast, best quality | |
| Download | Maximum accuracy | |
| Download | Multimodal projection file for image processing |
Q4_K_M for most use cases, only use F16 if you need maximum precision.