PERSONAS (Prism filter) is one of the largest datasets of synthetic preferences, with over 200k preferences over thousands of questions and 1k personas.
Details on the PERSONAS dataset can be found here paper link
Note that this subset is 5% of the training split of PERSONAS. The full dataset is here, strictly available for academic use.
You MUST request access to the full persona dataset here.
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/SynthLabsAI/PERSONA_subset.