Part of a three-line scale comparison (0.5B full fine-tune / 1.5B LoRA /
7B QLoRA) documenting VRAM and quality trade-offs on a single 12GB
consumer GPU. This is the 1.5B point — not the recommended model for
actual use (see twscholar-lm-dpo-7b-final), published for the
comparison's reproducibility.