The result of fine tuning llama-8b by mistake when I thought I had created a frankenmerge. Testing suggests it works alright though and with the Alpaca prompt format.
Trained for 3 epochs on the AdamCodd/no_robots-alpaca dataset with V100.
This llama model was trained 2x faster with
Unsloth and Huggingface's TRL library.