This repository contains a synthetic preference dataset built around coding tasks, safety-sensitive refusals, honesty checks, and everyday assistant behavior. It is designed for preference modeling, dataset tooling, and RLHF-style experimentation.
Layout
The repository is organized into four top-level subset folders: