A subset of JudgeBias-DPO-RefFree for training LLM judges to evaluate materials science synthesis recipes without bias in a reference-free setting (no ground truth recipe).
This dataset keeps only the 15% perturbation rate for graded perturbations and all 100% directional/individual datasets, removing the 1%, 2%, 5%, and 10% rate variants:
all_error_perturbation_15pct
all_error_perturbation_{1,2,5… See the full description on the dataset page:
https://huggingface.co/datasets/iknow-lab/JudgeBias-DPO-RefFree-subset.