This quant was made using exllamav2-0.0.21 with default dataset.
I tested this quant shortly in some random RPs (including one over 8k context - with RoPE scaling as recommended in webui, maybe with alpha_value a bit higher) and it seems to work fine.
Seems to use llama3 prompt template.
This is a merge of pre-trained language models created using
mergekit.
This model was merged using the
Model Stock merge method using
Sao10K/L3-8B-Stheno-v3.2 as a base.
1
2models:
3 - model: Sao10K/L3-8B-Stheno-v3.2
4 - model: NeverSleep/Llama-3-Lumimaid-8B-v0.1-OAS
5 - model: Hastagaras/Jamet-8B-L3-MK.V-Blackroot
6merge_method: model_stock
7base_model: Sao10K/L3-8B-Stheno-v3.2
8dtype: float16
9