This repository provides access to data sampled, cleaned, and categorized by dimensions from pair-wise data built from original UltraFeedback dataset, as presented in Hummer: Towards Limited Competitive Preference Dataset.These data are meant to train preference (or reward) models for subsequent and balanced RLHF training. These data are not meant for supervised training of dialogue agents. Training dialogue agents on these data is… See the full description on the dataset page: https://huggingface.co/datasets/sarinw-2024/Hummer.