Training data for a single-token RAG-groundedness guardrail. Each item is a
(question, context, response) triple with a human- or construction-derived
PASS/FAIL label under one Behavior Spec:
FAIL iff the response makes at least one factual claim unsupported by or
contradicting the retrieved context — truth in the real world is irrelevant
(strict grounding). PASS otherwise, including responses that decline to
answer for lack… See the full description on the dataset page:
https://huggingface.co/datasets/aaryand/context-adherence-guardrail-10k.