The 113 (code diff, haiku) supervised fine-tuning pairs used to train shahfazal/lgtm-575-gemma4-e4b-v0.1. Each row pairs a real PR code diff and its human reviewer comment with a 5-7-5 haiku that distills the comment's issue.
Source diffs and reviewer comments: the ronantakizawa/codereview-bench dataset (Python subset), which is MIT-licensed.
Haiku targets: generated by the teacher model Qwen/Qwen3-30B-A3B-Instruct-2507… See the full description on the dataset page:
https://huggingface.co/datasets/shahfazal/lgtm-575-training-pairs-v0.1.