Views
No views yet
seed-oss-think for think modeseed-oss-no-think for no think modeThinking mode1<seed:bos>system
2You are Doubao, a helpful AI assistant.
3<seed:eos>
4
5<seed:bos>user
6{user_message_1}
7<seed:eos>
8
9<seed:bos>assistant
10<seed:think>{thinking_content}</seed:think>
11{assistant_message_1}
12<seed:eos>
13
14<seed:bos>user
15{user_message_2}
16<seed:eos>
17
18<seed:bos>assistantNo-thinking mode1<seed:bos>system
2You are Doubao, a helpful AI assistant.
3<seed:eos>
4
5<seed:bos>system
6You are an intelligent assistant that can answer questions in one step without the need for reasoning and thinking, that is, your thinking budget is 0. Next, please skip the thinking process and directly start answering the user's questions.
7<seed:eos>
8
9<seed:bos>user
10{user_message_1}
11<seed:eos>
12
13<seed:bos>assistant
14{assistant_message_1}
15<seed:eos>
16
17<seed:bos>user
18{user_message_2}
19<seed:eos>
20
21<seed:bos>assistant5120001wasmedge --dir .:. \
2 --nn-preload default:GGML:AUTO:Seed-OSS-36B-Instruct-Q5_K_M.gguf \
3 llama-api-server.wasm \
4 --prompt-template seed-oss-no-think \
5 --ctx-size 512000 \
6 --model-name seed-oss| Name | Quant method | Bits | Size | Use case |
|---|---|---|---|---|
| Seed-OSS-36B-Instruct-Q2_K.gguf | Q2_K | 2 | 13.6 GB | smallest, significant quality loss - not recommended for most purposes |
| Seed-OSS-36B-Instruct-Q3_K_L.gguf | Q3_K_L | 3 | 19.1 GB | small, substantial quality loss |
| Seed-OSS-36B-Instruct-Q3_K_M.gguf | Q3_K_M | 3 | 17.6 GB | very small, high quality loss |
| Seed-OSS-36B-Instruct-Q3_K_S.gguf | Q3_K_S | 3 | 15.9 GB | very small, high quality loss |
| Seed-OSS-36B-Instruct-Q4_0.gguf | Q4_0 | 4 | 20.6 GB | legacy; small, very high quality loss - prefer using Q3_K_M |
| Seed-OSS-36B-Instruct-Q4_K_M.gguf | Q4_K_M | 4 | 21.8 GB | medium, balanced quality - recommended |
| Seed-OSS-36B-Instruct-Q4_K_S.gguf | Q4_K_S | 4 | 20.7 GB | small, greater quality loss |
| Seed-OSS-36B-Instruct-Q5_0.gguf | Q5_0 | 5 | 25.0 GB | legacy; medium, balanced quality - prefer using Q4_K_M |
| Seed-OSS-36B-Instruct-Q5_K_M.gguf | Q5_K_M | 5 | 25.6 GB | large, very low quality loss - recommended |
| Seed-OSS-36B-Instruct-Q5_K_S.gguf | Q5_K_S | 5 | 25.0 GB | large, low quality loss - recommended |
| Seed-OSS-36B-Instruct-Q6_K.gguf | Q6_K | 6 | 29.7 GB | very large, extremely low quality loss |
| Seed-OSS-36B-Instruct-Q8_0.gguf | Q8_0 | 8 | 38.4 GB | very large, extremely low quality loss - not recommended |
| Seed-OSS-36B-Instruct-f16-00001-of-00003.gguf | f16 | 16 | 30.0 GB | |
| Seed-OSS-36B-Instruct-f16-00002-of-00003.gguf | f16 | 16 | 30.0 GB | |
| Seed-OSS-36B-Instruct-f16-00003-of-00003.gguf | f16 | 16 | 12.4 GB |