Views
No views yet
unsloth/DeepSeek-R1-Distill-Llama-8B yang telah di-fine-tune 1 epoch (Total steps: 297, 9.487 contoh) menggunakan Unsloth. Hanya ~0,12% parameter diubah (9.437.184/8B).unsloth/DeepSeek-R1-Distill-Llama-8BNum GPUs used = 1 │ Num examples = 9.487 │ Num Epochs = 1 │ Total steps = 297
Batch size/device = 2 │ Grad. accumulation = 16 │ Total batch = 32
Trainable params: 9.437.184/8.000.000.000 (0,12% trained)