The files are GGUF quants of (kukulemon-7B)[
https://huggingface.co/grimjim/kukulemon-7B].
A merger of two similar Kunoichi models with strong reasoning, hopefully resulting in "dense" encoding of said reasoning, was merged with a model targeting roleplay.
I've tested with ChatML prompts with temperature=1.1 and minP=0.03. The model itself supports Alpaca format prompts. The model claims a context length of 32K, it seemed to lose coherence after 8K in my informal testing.
This is a merge of pre-trained language models created using
mergekit.
This model was merged using the SLERP merge method.
1slices:
2 - sources:
3 - model: grimjim/kuno-kunoichi-v1-DPO-v2-SLERP-7B
4 layer_range: [0, 32]
5 - model: KatyTheCutie/LemonadeRP-4.5.3
6 layer_range: [0, 32]
7# or, the equivalent models: syntax:
8# models:
9merge_method: slerp
10base_model: KatyTheCutie/LemonadeRP-4.5.3
11parameters:
12 t:
13 - filter: self_attn
14 value: [0, 0.5, 0.3, 0.7, 1]
15 - filter: mlp
16 value: [1, 0.5, 0.7, 0.3, 0]
17 - value: 0.5 # fallback for rest of tensors
18dtype: float16
19