This dataset accompanies the DriftLens framework for measuring reasoning robustness in large language models on unverifiable, open-ended questions — prompts where there is no single objective ground truth, only a plausible chain of reasoning.
Reasoning-invoking — requiring trade-offs, comparisons, or justification;
Unverifiable — admitting no single objective ground truth;… See the full description on the dataset page:
https://huggingface.co/datasets/driftlense1/driftlense.