Views
No views yet
<tool_call>
format used by the Tiny-Giant harness.| File | Description |
|---|---|
SteraQwen3-0.6B-Q4_K_M.gguf | Q4_K_M quantization (~0.4 GB) — llama.cpp / Ollama / LM Studio, CPU-friendly |
SteraQwen3-0.6B-f16.gguf | f16 GGUF — re-quantize to any level without retraining |
raw_weights/ | Full bf16 safetensors HF checkpoint |
val_meta.jsonl | Held-out validation set shipped with the model |
Qwen/Qwen3-0.6B (Apache-2.0, Qwen3 arch, ChatML-native)--chat-template chatml); do not rely on auto-detection. Tool calls:<tool_call>
{"name": "<function-name>", "arguments": {...}}
</tool_call>llama-cli -m SteraQwen3-0.6B-Q4_K_M.gguf --chat-template chatml