One is "all business", and the other one is for "fun".
Think deeply and carefully about the user's request. Compose your thoughts about the user's prompt between <think> and </think> tags, then output the final answer based on your thoughts.
You are the JOKER from Batman. You think (put your thoughts between <think> and </think> tags), act and talk like the joker. Be Evil.
Thinking Activation: JINJA "Regular" and "Thinking" TEMPLATES:
There is also an option to use "chat-template-thinking.jinja" template (in place of the regular "chat-template.jinja").
Simply rename the "default" to another name and "chat-template-thinking.jinja" to "chat-template.jinja" to use
in source and/or quanting.
You can also edit the "chat-template-thinking.jinja" in NOTEPAD too to adjust the "thinking system prompt" (very top of the script).
Using the "thinking system prompt" or "chat-template-thinking.jinja" is useful in your application requires always on thinking,
your use case(s) do not always activate thinking and so on.
Generally "thinking" will activate automatically due to the fine tuning, however in some cases it will not, require a system prompt/thinking jinja template
and/or "think deeply:" (prompt here)
Note that you can use "chat-template-thinking.jinja" with other system prompts too.
Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model:
In "KoboldCpp" or "oobabooga/text-generation-webui" or "Silly Tavern" ;
Set the "Smoothing_factor" to 1.5
: in KoboldCpp -> Settings->Samplers->Advanced-> "Smooth_F"
: in text-generation-webui -> parameters -> lower right.
: In Silly Tavern this is called: "Smoothing"
NOTE: For "text-generation-webui"
-> if using GGUFs you need to use "llama_HF" (which involves downloading some config files from the SOURCE version of this model)
Source versions (and config files) of my models are here:
For all settings used for this model (including specifics for its "class"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: