Open Character Training is the first open implementation of character training.
For more information, read our paper!
This repository includes all training data generated for our paper for the misalignment persona only. See Section 2 for details of DPO, self-reflection, and self-interaction.
It is intended for further research / replication.
Please get in touch for further details.
This dataset follows the same license used in LIMA… See the full description on the dataset page:
https://huggingface.co/datasets/maius/OpenCharacterTraining-data-misalignment.