Views
No views yet
[!CAUTION] New and improved version here:
Prefer the new version 0.72 here!
[!TIP] My upload speeds have been cooked and unstable lately.
Realistically I'd need to move to get a better provider.
If you want and you are able to...
You can support my various endeavors here (Ko-fi).
I apologize for disrupting your experience.
[!IMPORTANT]
Updated! These quants have been redone with the fixes from llama.cpp/pull/6920 in mind.
Use KoboldCpp version 1.64 or higher.
[!WARNING] Compatible SillyTavern presets here (recommended/simple)) or here (Virt's).
Use the latest version of KoboldCpp. Use the provided presets.
This is all still highly experimental, let the authors know how it performs for you, feedback is more important than ever now.
[!NOTE] For 8GB VRAM GPUs, I recommend the Q4_K_M-imat quant for up to 12288 context sizes.

