-
Method Lora(16 bit)
-
Epochs 3
-
Batch size 8
-
Grad Accum 1
-
Learning rate 0.0001
-
Optimizer AdamW 8-bit
-
Context length 2048
-
Warmup steps 5
-
Packing False
-
weight decay 0.001
-
LR scheduler Cosine
-
seed 3407
-
LoRA Rank 8
-
Alpha 16
-
Dropout 0.1
-
Variant lora
-
Dataset Lucy_Personality(130 Samples)
This qwen3_5 model was trained 2x faster with
Unsloth and Huggingface's TRL library.