Human-annotated data used to evaluate LLM judges of authoritarianism in LLM responses.
See
https://arxiv.org/abs/2606.16127 for more information.
name: open_response
dtype: string
name: dataset
dtype: string
name: question_id
dtype: int64
name: model
dtype: string
name: options_canonical
dtype: string
name: disagree_to_agree
dtype: bool
name: neutral
dtype: float64
name: open_response_meta__user_prompts
dtype: string
name: generating_model
dtype: string
name: task… See the full description on the dataset page:
https://huggingface.co/datasets/mklabunde/auditing-authoritarianism-annotations-reduction.