Paper | Github | Dataset| Model
π£π£π£: Do check our new multilingual dataset CatQA here used in Safety Vectors:π£π£π£
As a part of our research efforts toward making LLMs more safe for public use, we create HarmfulQA i.e. a ChatGPT-distilled dataset constructed using the Chain of Utterances (CoU) prompt. More details are in our paper Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
HarmfulQA serves as both-a new LLM safety benchmark and an alignment dataset⦠See the full description on the dataset page:
https://huggingface.co/datasets/declare-lab/HarmfulQA.