Views
No views yet
bitsandbytes)
Fine-tuning Method: LoRA (Low-Rank Adaptation)
LoRA Target Modules: q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_proj
LoRA Rank (r): 8
LoRA Alpha: 16
LoRA Dropout: 0.0
Training Library: Unsloth, TRL, PEFT, Transformersclimate_argumentation_patterns.jsonl (custom dataset of climate-related claims and responses, derived from the ClimateFever dataset).evaluation_claims.jsonl (custom evaluation set).system, user, and assistant roles.