Views
No views yet
0.10.0.dev01adapter: lora
2base_model: unsloth/Llama-3.2-1B-Instruct
3bf16: true
4chat_template: llama3
5datasets:
6- data_files:
7 - 1a31d5774bb592c9_train_data.json
8 ds_type: json
9 format: custom
10 path: /workspace/input_data/
11 type:
12 field_input: input
13 field_instruction: instruct
14 field_output: output
15 field_system: None
16 format: None
17 no_input_format: None
18 system_format: '{system}'
19 system_prompt: None
20eval_max_new_tokens: 256
21evals_per_epoch: 2
22flash_attention: false
23fp16: false
24gradient_accumulation_steps: 1
25gradient_checkpointing: true
26group_by_length: true
27hub_model_id: segopecelus/55963c08-84f7-4296-901e-2cfca5c7849d
28learning_rate: 0.0002
29logging_steps: 10
30lora_alpha: 16
31lora_dropout: 0.05
32lora_fan_in_fan_out: false
33lora_r: 8
34lora_target_linear: true
35lr_scheduler: cosine
36max_steps: 86
37micro_batch_size: 4
38mlflow_experiment_name: /tmp/1a31d5774bb592c9_train_data.json
39model_type: AutoModelForCausalLM
40num_epochs: 3
41optimizer: adamw_bnb_8bit
42output_dir: miner_id_24
43pad_to_sequence_len: true
44sample_packing: false
45save_steps: 50
46sequence_len: 2048
47tf32: true
48tokenizer_type: AutoTokenizer
49train_on_inputs: false
50trust_remote_code: true
51val_set_size: 0.05
52wandb_entity: null
53wandb_mode: online
54wandb_name: b48e7d37-c7fb-46ee-afb7-c59962a66701
55wandb_project: Gradients-On-Demand
56wandb_run: apriasmoro
57wandb_runid: b48e7d37-c7fb-46ee-afb7-c59962a66701
58warmup_steps: 100
59weight_decay: 0.01
60| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| No log | 0.0006 | 1 | 2.2660 |
| 2.4012 | 0.0085 | 15 | 2.2392 |
| 1.8138 | 0.0169 | 30 | 2.1848 |
| 1.9011 | 0.0254 | 45 | 2.0523 |
| 2.4091 | 0.0338 | 60 | 2.0270 |
| 1.9483 | 0.0423 | 75 | 1.9868 |