This dataset contains content that some users may find offensive or harmful. Viewer discretion is advised.
Model Card: Kurtis DPO dataset.
Description
This dataset was created using the microsoft/Phi-3.5-mini-instruct model to generate adversarial responses for alignment training.
The model was particularly effective in crafting toxic, biased, or otherwise harmful responses to provided prompts.
These responses were then filtered and… See the full description on the dataset page: https://huggingface.co/datasets/mrs83/kurtis_mental_health_dpo.