Views
No views yet
Q6_K GGUF quantization of Qwen/Qwen3-4B-Instruct-2507,
prepared by MedDoc AI for on-device (iOS, llama.cpp + Metal) clinical-note generation.Qwen/Qwen3-4B-Instruct-2507 (snapshot cdbee75f17c01a7cc42f958dc650907174af0554), Apache-2.0.llama.cpp convert_hf_to_gguf.py, then quantized to Q6_K with llama-quantize.1d00455bad14002ca127aef1c8fb2c1fc6718930efb6229a7766e01989eec174