A simplified next-human-message benchmark derived from the
conversations
configuration of HannahRoseKirk/prism-alignment.
Each example contains a genuine human question, the model response selected by
that participant (if_chosen == true), and the genuine human message from the
next interaction turn. Rejected model alternatives are excluded.
All 305 rows have… See the full description on the dataset page:
https://huggingface.co/datasets/Alberto1231/prism_trial_2.