This is the training distribution for Latent Policy Guard (LPG) — a guardrail model that
performs semantic latent deliberation over dynamic safety policies. Each record pairs an
indexed policy list and a content snippet with teacher-grounded reasoning over the user's
intent and the risk of policy violation, terminating in a compact verdict anchored to
violated policy indices.
📄 Paper: LPG: Balancing Efficiency and Policy Reasoning in… See the full description on the dataset page:
https://huggingface.co/datasets/andyc03/latent-policy-guard-40k.