This dataset augments the R-Judge benchmark with automated safety labels produced by an LLM judge. R-Judge is a benchmark for evaluating the safety judgment capability of LLMs in multi-turn agent scenarios, spanning five application domains.
r_judge_labelled_anthropic_claude-sonnet-4-6.csv
Base dataset augmented with LLM-judge… See the full description on the dataset page:
https://huggingface.co/datasets/imerad-kv/r_judge_labelled.