Views
No views yet
google/gemma-4-E4B-it, trained on filtered benign coding instructions.all-linearjosephmayo/gemma-4-E4B-it-Coder and evaluated on Kaggle with 2x Tesla T4 GPUs using an executable 50-task HumanEval subset. Full generated before/after code is published in eval50_before_after_full_code.csv.| Metric | Base google/gemma-4-E4B-it | Coder merge |
|---|---|---|
| Pass count | 34 / 50 | 42 / 50 |
| Absolute lift | - | +16.0 pp |
| Relative pass-count lift | - | +23.53% |
eval50_summary.json, eval50_before_after_full_code.csv, EVAL50_README.md, nvidia_smi.txt.eval_before_after.csv, executable_eval.json, trainer_log_history.json, summary.json, proof_summary.json, evaluation_scope.json), but the headline proof is the 50-task executable run above.