This dataset contains model output preferences evaluated by GPT-4 and Claude for each instruction pair.
It enables analysis of model alignment and preference patterns.
instruction: Instruction provided to the models.
output_1: First model's response.
output_2: Second model's response.
gpt4_preferred_output: GPT-4's preferred response.
claude_preferred_output: Claude's preferred response.
This dataset is useful for model… See the full description on the dataset page:
https://huggingface.co/datasets/pratyushmaini/alpaca_model_preference.