This is a DPO mix focused on making natural-sounding and context-aware LLMs. Please credit the original dataset authors for use.
winglian/no_robots_rlhf
nbeerbower/human-writing-dpo
nbeerbower/synthetic-fiction-dpo
jondurbin/gutenberg-dpo-v0.1
nbeerbower/gutenberg2-dpo
nbeerbower/gutenberg-moderne-dpo
sam-paech/gutenberg3-generalfiction-scifi-fantasy-romance-adventure-dpo
nbeerbower/GreatFirewall-DPO
nbeerbower/Schule-DPO
nbeerbower/Purpura-DPO
nbeerbower/Arkhaios-DPO… See the full description on the dataset page:
https://huggingface.co/datasets/schneewolflabs/Athanor-DPO.