Views
No views yet
| File | Size | Description |
|---|---|---|
Qwen2.5-Omni-3B-decoder-F16.gguf | 6.4 GB | Full precision (FP16) |
llama-cli -m Qwen2.5-Omni-3B-decoder-F16.gguf -p "Hello" -n 100 -no-cnvconvert_hf_to_gguf.py from llama.cpp. The converter automatically strips thinker. prefix and drops vision/audio/talker/token2wav components, keeping only the text decoder (435 tensors).