Public input on model behavior
Summary. CoVal (crowd-originated, values-aware preferences and rubrics) is a human-feedback dataset focused on value-sensitive model behavior. It has three components: two conversation-level files and one annotator-level file. The first conversation-level file contains (i) a synthetic prompt represented as a minimal chat transcript, (ii) four candidate assistant responses, and (iii) annotator assessments with rationales. The second… See the full description on the dataset page:
https://huggingface.co/datasets/openai/coval.