4,835 high-quality SFT pairs for training a model to write natural, direct prose.
Part of the humanize-rl project — a two-layer scoring and alignment pipeline for training small models to generate natural, human-sounding text.
Write natural Slack messages and emails from scratch.
Rewrite stiff/formal/corporate text into direct, human-sounding prose.
Fix grammar without making text formal.
Shorten and… See the full description on the dataset page:
https://huggingface.co/datasets/jayshah5696/humanize-rl-sft-dataset.