Borealis-10.7B-DPO is a 10.7B model made of 48 Mistral 7B layers, finetuned for +70h on 2xA6000 on a big RP and Conversational dataset with llama2 configuration of Axolotl, like SOLAR.
This variant had a DPO train on top of it.
Description
This repo contains GGUF files of Borealis-10.7B-DPO, a conversational model.
The goal of this model isn't to break all benchmark, but to have a better RP/ERP/Conversational model.
It was trained on multiple basic dataset to make it intelligent, but majority of the dataset was basic conversations.