This model was fine tuned using 1 Billion tokens of Alpaca format code feedback (the dataset is linked). This model is the first of many, I plan to run a full epoch of this dataset soon, and update the model along with it, currently ive only done around 10% of an epoch.
Example usage:
For text only LLMs: llama-cli -hf koshuro/Llama-3B-Coder --jinja
For multimodal models: llama-mtmd-cli -hf koshuro/Llama-3B-Coder --jinja
Benchmarks
On basic reasoning, math, and ela benchmarks, this model scored close to its base model, and near the same score as google/gemma-3n-e4b.