Finetune of Llama-3-70b-Instruct using unalignment/toxic-dpo-v0.2, using MonsterAPI
Trained for 1 epoch, use convert-lora-to-ggml.py in this repo to merge with the Llama-3-70b GGUF.
Use at your own risk; I am not responsible for what you do with this model.
{system}
USER: {user} ASSISTANT: {assistant}</s>