This llama model was trained 2x faster with
Unsloth and Huggingface's TRL library.
This model is fine-tuned on 10k sample points from Cendol V2 dataset.
This model is more straightforward and more natural in Indonesian (idk, it's vibe-based fine-tuning, need benchmark)