Views
No views yet

[MODEL_SETTINGS] block between the system prompt and the first user turn. This block controls reasoning effort. Ensure your client is configured correctly to avoid degraded performance.[SYSTEM_PROMPT]Your system prompt here[/SYSTEM_PROMPT][MODEL_SETTINGS]{"reasoning_effort": "high"}[/MODEL_SETTINGS][INST]user message[/INST][THINK]...prefill reasoning here...[/THINK]assistant reply</s>[SYSTEM_PROMPT]Your system prompt here[/SYSTEM_PROMPT][MODEL_SETTINGS]{"reasoning_effort": "none"}[/MODEL_SETTINGS][INST]user message[/INST]assistant reply</s>[MODEL_SETTINGS] block is injected after the system prompt and before [INST], passing a JSON object to control reasoning."reasoning_effort": "high" to enable chain-of-thought reasoning, or "none" to skip it entirely for faster responses.[THINK] block to guide the model's internal reasoning before generating its reply.[MODEL_SETTINGS] is injected at the correct position in the context. The [THINK] prefill is optional but recommended when reasoning is active.[SYSTEM_PROMPT]Prompt[/SYSTEM_PROMPT][MODEL_SETTINGS]{"reasoning_effort": "high"}[/MODEL_SETTINGS][INST]user message[/INST][THINK]I must write in first person perspective the next reply only from {{user}} perspective, using I/My for {{user}}. {{input}}, I will make a short description and dialogue block writing now:[/THINK]assistant reply</s>.json file below and import it into SillyTavern's sampler presets menu.*He walked across the room and stared out the window.**-I wonder what she's thinking.-*Alex (Curious): "What do you see out there?"**scene transitions**. The model should now produce cleaner prose.