Flat version of:
https://huggingface.co/datasets/allenai/ultrafeedback_binarized_cleaned
Update 1/12/2023: I've removed examples identified as faulty by Argilla - see their awesome work for more details.
This is a version of the UltraFeedback binarized dataset but with TruthfulQA prompts removed and source annotations added (so you can filter out samples from different sources yourself if you want!).
Please see the binarized dataset… See the full description on the dataset page:
https://huggingface.co/datasets/allenai/ultrafeedback_binarized_cleaned_train.