Views
No views yet
llama.cpp to make them runnable on consumer hardware via CPU and GPU offloading.| File Name | Bit Depth | Description |
|---|---|---|
| Qwen3-30B-A3B-Instruct-2507-REAM-heretic-Q4_K_M.gguf | 4-bit | ⭐ Recommended. Best balance of speed, RAM usage, and intelligence. |
| Qwen3-30B-A3B-Instruct-2507-REAM-heretic-Q5_K_M.gguf | 5-bit | Higher quality, slightly larger RAM requirement. |
| Qwen3-30B-A3B-Instruct-2507-REAM-heretic-Q6_K.gguf | 6-bit | Very high quality, near unquantized performance. |
| Qwen3-30B-A3B-Instruct-2507-REAM-heretic-Q8_0.gguf | 8-bit | Extremely high quality, large file size. |
| Qwen3-30B-A3B-Instruct-2507-REAM-heretic-Q3_K_M.gguf | 3-bit | Smallest file size, noticeable intelligence loss. Use only if heavily RAM constrained. |
1<|im_start|>system
2You are a helpful assistant.<|im_end|>
3<|im_start|>user
4Hello!<|im_end|>
5<|im_start|>assistant