100,000 DPO preference pairs training LLMs to stay within knowledge bounds. The chosen response is accurate and appropriately uncertain; the rejected response is confident but wrong — fabricated statistics, fake citations, wrong facts, overclaimed certainty.
Hallucination is the #1 reliability concern blocking enterprise LLM adoption. Models fail in predictable patterns:
Inventing specific statistics with false… See the full description on the dataset page:
https://huggingface.co/datasets/stindardlogic/hallucination-reduction-dpo-100k.