We released the dataset used in our EMNLP2024 paper PRS. Our project website is here.
This dataset can be used for personalized model alignment, which means the model is trained to generate personalized outputs, such as the user preference can be "I expect the response to be humorous" or "I prefer the response to provide support evidence".