Tokenizer and metadata preserved during conversion
Choose BF16 for best fidelity, F16 for GPU does not support BF16, Q8_0 for balance, Q4_K_M for lowest memory
⚖️ License
Weights inherit the upstream model’s license.
This repository redistributes format-converted copies only.
Please review and comply with the upstream terms before use.
📝 Acknowledgments
Original model by SCB10X.
GGUF conversion with quantization performed via llama.cpp tooling.