Views
No views yet
[!NOTE]
To use the long CoT reasoning mode, use a system prompt likeExplain your reasoning step-by-step using <think>...</think>, then give the final answer inside <response>...</response>.
[!WARNING]
This model is experimental. While the model provides coherent and expert-like responses, users should verify its outputs for accuracy - especially in calculations or logical reasoning tasks.
1models:
2 - model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B
3 - model: Skywork/Skywork-o1-Open-Llama-3.1-8B
4 - model: SimpleBerry/LLaMA-O1-Supervised-1129
5 - model: NousResearch/DeepHermes-3-Llama-3-8B-Preview
6 - model: O1-OPEN/OpenO1-LLama-8B-v0.1
7 - model: nvidia/Llama-3.1-Nemotron-Nano-8B-v1
8merge_method: karcher
9tokenizer:
10 source: meta-llama/Llama-3.1-8B-Instruct
11dtype: bfloat1610K subset for 1 epoch using LLaMA Factory.