I trained my previous merge with Purpura DPO via Unsloth in an attempt to get rid of impersonation issues, seems to not have worked, i'll return to the drawing board and probably do another merge.
This is a merge of pre-trained language models created using
mergekit.
This model was merged using the Passthrough merge method using
SanXM1/Driftwood-12B +
SanXM1/Purpura-DPO-ckpt-4epochs as a base.
1base_model: SanXM1/Driftwood-12B+SanXM1/Purpura-DPO-ckpt-4epochs
2dtype: bfloat16
3merge_method: passthrough
4models:
5 - model: SanXM1/Driftwood-12B+SanXM1/Purpura-DPO-ckpt-4epochs
6