QLoRA fine-tune (rank=16, alpha=32) of meta-llama/Meta-Llama-3.1-8B-Instruct on yahma/alpaca-cleaned — 1 epochs, final loss 0.9306, 62.1 min on MI300X.
Details
Base model: meta-llama/Meta-Llama-3.1-8B-Instruct
Trained as part of the MI300X 50-hour fine-tuning bootcamp.