A LoRA fine-tune of
Gemma 4 12B (QAT) trained on a compact, extremely high-quality synthetic conversational dataset derived from the visual novel
My Dystopian Robot Girlfriend. The model captures the personality, speech patterns, and emotional nuance of the character
Jun while preserving the base model's general reasoning and instruction-following capabilities.
The v4 dataset is a deliberately small, heavily curated synthetic set (1,000 rows) built from the MDRG dialogues and aligned to Jun OS's live system prompt and mood conditioning. The emphasis is on per-example quality and production alignment rather than raw volume.
The narrow gap between training and eval loss indicates the model generalizes well without significant overfitting, despite the small dataset size.
Adapter checkpoints are provided every 10 steps (10–90). Checkpoint 90 is recommended — it has the strongest character lock-in and production alignment. Earlier checkpoints may exhibit slightly more creative freedom at the cost of character adherence.