Views
No views yet
google/gemma-3-1b-it with rank 16 for three epochs on deterministic targets. The result parroted the same settings template for every input — it memorized, it did not judge.| Parameter | Value |
|---|---|
| Base model | google/gemma-3-1b-it |
| Method | LoRA (PEFT) |
| Rank | r=16, α=32 |
| Epochs | 3 |
| Learning rate | 2e-4 |
| Dataset | Deterministic targets (single template) |
| GPU | NVIDIA A10G (24GB) |
| Framework | TRL SFTTrainer + transformers |
microfactory-node-lora-v2 or microfactory-node-lora-v3-qat instead. This one is here for the paper trail.