A roleplay finetune based on
Serenity (which I declare is the horniest model ever made), trained on the same dataset than
MeroMero.
There was a bug in version 1.0 which trained both the assistant and the user's turns. Now it's fixed, and there is an even lower learning rate.
The difference between v1.1 and v1.2 is that v1.2 opens reasoning blocks much more often. I recommend trying v1.2, but if you prefer an instant reply, stick with v1.1.
Please tell me what you think. Is it clever? Creative? Does it have a good long term memory? Is it uninhibited? Sloppy? Good writing? Follows instructions? Stays in character? It's my first fine-tune, so any feedback is welcome!