Views
No views yet
| Parameter | Value |
|---|---|
| Base model | Qwen/Qwen3-4B-Instruct-2507 |
| Context length | 2048 |
| Batch size (prefill) | 64 |
| FFN / LM-head quantization | LUT6, per-channel 4 |
| Chunks | 4 |
| ANEMLL version | 0.3.5 |
meta.yaml for the full conversion record.qwen_embeddings.mlmodelc — embeddingsqwen_FFN_PF_lut6_chunk_0Xof04.mlmodelc — FFN + prefill chunksqwen_lm_head_lut6.mlmodelc — LM headmeta.yaml — model/conversion metadata consumed by the runtimetokenizer.json, tokenizer_config.json, vocab.json, merges.txt)