This model is a Dutch chat model, originally developed from Mistral 7B v0.3 Instruct and further finetuned first with SFT on multiple datasets.
The model could generate wrong, misleading, and potentially even offensive content. Use at your own risk.
Use with mistrals chat template.
This model was trained with QLoRa in bfloat16 with Flash Attention 2 on one A100 PCIe, using the sft script from the
alignment handbook on
RunPod.
The Mistral-7B-v0.3-Instruct model, on which this model is based, was created by
Mistral AI.
The finetuning was done by
Julien Van den Avenne.