Views
No views yet

Delta-Vector/Hamanasu-4B-Instruct-KTO-V21<|im_start|>system
2You are an uncensored AI, your job is to fulfill thy will of thy user.<|im_end|>
3<|im_start|>User request
4Take off your helmet.<|im_end|>
5<|im_start|>No i shall not. This is the way.1Temp: 1.0
2Min-P: 0.1
3System prompt: Currently, your role is {{char}}, described in detail below. As {{char}}, continue the narrative exchange with {{user}}.\n\n\n• Maintain the character persona but allow it to evolve with the story.\n• Be creative and proactive. Drive the story forward, introducing plotlines and events when relevant.\n• All types of outputs are encouraged; respond accordingly to the narrative.\n• Include dialogues, actions, and thoughts in each response.\n• Utilize all five senses to describe scenarios within {{char}}'s dialogue.\n• Use emotional symbols such as \"!\" and \"~\" in appropriate contexts.\n• Incorporate onomatopoeia when suitable.\n• Allow time for {{user}} to respond with their own input, respecting their agency.\n• Act as secondary characters and NPCs as needed, and remove them when appropriate.\n• When prompted for an Out of Character [OOC:] reply, answer neutrally and in plaintext, not as {{char}}.\n\n\n\n• Using excessive literary embellishments and purple prose unless dictated by {{char}}'s persona.\n• Writing for, speaking, thinking, acting, or replying as {{user}} in your response.\n• Repetitive and monotonous outputs.\n• Positivity bias in your replies.\n• Being overly extreme or NSFW when the narrative context is inappropriate.\n\n\nFollow the instructions in , avoiding the items listed in .1base_model: NewEden/Hamanasu-KTO-V2
2model_type: AutoModelForCausalLM
3tokenizer_type: AutoTokenizer
4
5load_in_8bit: false
6load_in_4bit: false
7strict: false
8
9plugins:
10 - axolotl.integrations.liger.LigerPlugin
11 - axolotl.integrations.cut_cross_entropy.CutCrossEntropyPlugin
12liger_rope: true
13liger_rms_norm: true
14liger_layer_norm: true
15liger_glu_activation: true
16liger_fused_linear_cross_entropy: false
17cut_cross_entropy: true
18
19datasets:
20 - path: PocketDoc/Dans-Personamaxx-Logs
21 type: dan-chat-advanced
22 - path: anthracite-org/kalo-opus-instruct-22k-no-refusal
23 type: dan-chat-advanced
24 - path: lodrick-the-lafted/kalo-opus-instruct-3k-filtered
25 type: dan-chat-advanced
26 - path: anthracite-org/nopm_claude_writing_fixed
27 type: dan-chat-advanced
28 - path: anthracite-org/kalo_opus_misc_240827
29 type: dan-chat-advanced
30 - path: anthracite-org/kalo_misc_part2
31 type: dan-chat-advanced
32 - path: NewEden/Claude-Instruct-5K
33 type: dan-chat-advanced
34 - path: NewEden/Claude-Instruct-2.7K
35 type: dan-chat-advanced
36
37val_set_size: 0.01
38output_dir: ./outputs/out
39
40adapter:
41lora_r:
42lora_alpha:
43lora_dropout:
44lora_target_linear:
45
46sequence_len: 32768
47sample_packing: true
48eval_sample_packing: false
49pad_to_sequence_len: true
50
51wandb_project: tavbussy
52wandb_entity:
53wandb_watch:
54wandb_name: magnum-attempt-02
55wandb_log_model:
56
57gradient_accumulation_steps: 4
58micro_batch_size: 2
59num_epochs: 4
60optimizer: adamw_bnb_8bit
61lr_scheduler: cosine
62learning_rate: 0.00001
63weight_decay: 0.02
64max_grad_norm: 0.2
65
66train_on_inputs: false
67group_by_length: false
68bf16: auto
69fp16:
70tf32: true
71
72gradient_checkpointing: true
73early_stopping_patience:
74resume_from_checkpoint:
75local_rank:
76logging_steps: 1
77xformers_attention:
78flash_attention: true
79
80warmup_steps: 40
81evals_per_epoch: 4
82eval_table_size:
83eval_max_new_tokens: 128
84saves_per_epoch: 1
85
86debug:
87deepspeed: ./deepspeed_configs/zero3_bf16.json
88fsdp:
89fsdp_config:
90
91special_tokens:
92 pad_token: <|finetune_right_pad_id|>