A dataset of 564K examples for training lightweight preference extraction models. Each example pairs a conversation input with structured JSON output describing user preferences as condition-action rules.
This dataset was used to train blackhao0426/pref-extractor-qwen3-0.6b-full-sft, a core component of the VARS framework.
The following snippet from the official repository demonstrates how to use the framework… See the full description on the dataset page:
https://huggingface.co/datasets/blackhao0426/user-preference-564k.