Python CWE GRPO training dataset with rubric-aware authoring guidelines, built from the post–step-6 pipeline (6_rewritten.jsonl / 6_rewritten_guidelines.jsonl).
7,977 rows (17 CWEs) — the post-rewrite oracle set, not the older 10k HF aggregate.
Reused rubric-aware guidelines from a prior generation when (cwe, function_name, prompt) matched and generated_code was identical (~5.4k rows).
Regenerated… See the full description on the dataset page:
https://huggingface.co/datasets/AetherPrior/py_cwe_GRPO.