This dataset contains new rationales for story pair evaluations from the LitBench dataset, generated using GPT-4 with a structured rubric-based evaluation approach.
Creativity & Originality (25 points): Uniqueness of concept, innovative elements, fresh perspective
Writing Quality & Style (25 points): Prose quality, voice consistency, grammar and… See the full description on the dataset page:
https://huggingface.co/datasets/SAA-Lab/litbench-rationales-gpt4.