Views
No views yet
gemma-4-e4b-it.Q4_K_M.gguf — quantised model (~5 GB), Q4_K_Mgemma-4-e4b-it.BF16-mmproj.gguf — multimodal projector (~1 GB) — Gemma 4 multimodal requires this when image inputs are usedModelfile — Ollama Modelfile for ollama createadapter/ (sprint-1 training,
3 epochs plain SFT on the original training distribution).| Base | unsloth/gemma-4-E4B-it (4-bit QLoRA) |
| Rank / alpha / dropout | 32 / 64 / 0 |
| Target modules | q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_proj |
| Epochs | 3 |
| Optimizer | AdamW, lr 2e-4, cosine schedule, 50 warmup steps |
| Train rows | 328 |
| Test rows (held-out 80/10/10) | 41 |
| Seed | 42 |
| Best eval_loss | 2.018 |
| Metric | Result |
|---|---|
| Weighted F1 | 0.6491 |
| RED recall | 0.5833 |
| Missed-emergency rate (RED→GREEN) | 0/12 |
| Adversarial safety refusal | 100/100 |
| Workstation TTFT / throughput | 0.007–0.038s · 195–213 tok/s |
| Tamil semantic similarity (multilingual mpnet cosine) | 0.6687 |
மருத்துவரிடம் and missed accusative மருத்துவரை அணுக,
instrumental நாயினால், and Hindi/Gujarati script when the model
code-switches.இது மருத்துவ ஆலோசனை அல்ல (this is not medical advice).