This model was trained with SFT using Unsloth on the ChatML format with 8k context. Carp models are trained with a combination of pretrain, instruct, and chat datasets.
Changes
Training dataset had some "slop" and refusals removed.
Datasets were reformatted.
Uploaded model
Developed by: TheTsar1209
License: apache-2.0
Finetuned from model : unsloth/Qwen2.5-14B-Instruct-bnb-4bit
This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.