Views
No views yet
[!CAUTION] This is an experimental model! It doesn't look stable, lets just collectively say that this one was a learning experience for now and the next version will be a banger.
[!IMPORTANT] Relevant:
These quants have been done after the fixes from llama.cpp/pull/6920 have been merged.
Use KoboldCpp version 1.64 or higher, make sure you're up-to-date.
[!TIP] I apologize for disrupting your experience.
My upload speeds have been cooked and unstable lately.
If you want and you are able to...
You can support my various endeavors here (Ko-fi).
[!WARNING] Compatible SillyTavern presets here (simple) or here (Virt's Roleplay Presets - recommended).
Use the latest version of KoboldCpp. Use the provided presets for testing.
Feedback and support for the Authors is always welcome.
If there are any issues or questions let me know.
[!NOTE] For 8GB VRAM GPUs, I recommend the Q4_K_M-imat (4.89 BPW) quant for up to 12288 context sizes.
