Selected Corral traces for manual annotation of epistemic patterns across environments, models, and QA dimensions
This dataset is part of the Corral collection accompanying the paper AI scientists produce results without reasoning scientifically. It contains the evaluation traces selected for manual annotation of epistemic patterns across the Corral benchmark.
The dataset is organized into 68… See the full description on the dataset page:
https://huggingface.co/datasets/jablonkagroup/questions4manual_annotation.