This is a small dataset to support training and evaluation of conversational AI in emotionally sensitive contexts.
Each sample contains:
a user input
two assistant responses
a human preference
optional rubric scoring
metadata such as tone, formality, and topic
supervised fine-tuning (SFT)
preference modeling (for RLHF or DPO)
safe response generation
tone- or style-controlled generation
Apache 2.0 — free for… See the full description on the dataset page:
https://huggingface.co/datasets/hoanghai2110/EmotionAlignQA.