Views
No views yet
| File | Quant | Size | Source Weights | MTP Source |
|---|---|---|---|---|
Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-BF16.gguf | BF16 | 71.07 GB (66.19 GiB) | AEON-7/...-BF16 | unsloth/Qwen3.6-35B-A3B-MTP-GGUF |
Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-NVFP4.gguf | NVFP4 | 23.40 GB (21.80 GiB) | AEON-7/...-NVFP4 | unsloth/Qwen3.6-35B-A3B-MTP-GGUF via s-batman |
Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-Q8_0.gguf | Q8_0 | 37.80 GB (35.21 GiB) | Quantized from the BF16 GGUF | unsloth/Qwen3.6-35B-A3B-MTP-GGUF |
1a13df4cce8a32b2065d8aea51dcc80d7056fea6c3277266d9b040923a2641840 Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-BF16.gguf
2d78f62f6c112de9721390ce8f75b22cf753b3766a640257cc13ca85f16030292 Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-NVFP4.gguf
316696fb2e19b5b3faa316b198524be3dff3652555c67c3f3ea11e811147b219a Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-Q8_0.gguf1deepreinforce-ai/Ornith-1.0-35B
2 -> AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16
3 -> AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-NVFP4blk.40.* MTP tensors grafted directly from Unsloth's BF16 split GGUF.blk.40.* MTP tensors grafted via s-batman/Ornith-1.0-35B-NVFP4-MTP-GGUF. Byte-level verification confirms this block is identical to Unsloth's Qwen3.6-35B-A3B-MXFP4_MOE.gguf MTP block: 20 tensors, 512,079,872 tensor payload bytes, combined tensor-name-plus-payload SHA-256 8b8ba06cf776d2cdbf4d4db6714cf69b8a455105fc848bc02c4e5acb62f585f1.blk.40.* MTP tensors grafted directly from Unsloth's Q8_0 GGUF.draft-mtp speculative decoding support.1llama-cli \
2 -m Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-Q8_0.gguf \
3 --spec-type draft-mtp \
4 --spec-draft-n-max 3 \
5 -p "Explain gradient descent in 3 sentences."1llama-server \
2 -m Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-NVFP4.gguf \
3 --host 0.0.0.0 --port 8080 \
4 -ngl all \
5 -c 65536 \
6 --spec-type draft-mtp \
7 --spec-draft-n-max 3--spec-type draft-mtp flag.1mrexodia/Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-GGUF:BF16
2mrexodia/Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-GGUF:NVFP4
3mrexodia/Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-GGUF:Q8_0--no-mtp.1python convert_hf_to_gguf.py \
2 AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16 \
3 --outtype bf16 \
4 --no-mtp \
5 --outfile Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16.ggufblk.40. from:unsloth/Qwen3.6-35B-A3B-MTP-GGUF/BF16/Qwen3.6-35B-A3B-BF16-00002-of-00002.gguf1qwen35moe.block_count = 41
2qwen35moe.nextn_predict_layers = 1Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-BF16.gguf1python convert_hf_to_gguf.py \
2 AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-NVFP4 \
3 --outtype bf16 \
4 --no-mtp \
5 --outfile Ornith-1.0-35B-AEON-Ultimate-Uncensored-NVFP4.gguf--outtype bf16 only affects non-NVFP4 tensors that remain floating point.blk.40. from:s-batman/Ornith-1.0-35B-NVFP4-MTP-GGUF/ornith-1.0-35b-NVFP4_MOE-MTP.ggufunsloth/Qwen3.6-35B-A3B-MTP-GGUF/Qwen3.6-35B-A3B-MXFP4_MOE.gguf1qwen35moe.block_count = 41
2qwen35moe.nextn_predict_layers = 1Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-NVFP4.gguf1llama-quantize \
2 Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16.gguf \
3 Ornith-1.0-35B-AEON-Ultimate-Uncensored-Q8_0-body.gguf \
4 q8_0blk.40. from:unsloth/Qwen3.6-35B-A3B-MTP-GGUF/Qwen3.6-35B-A3B-Q8_0.gguf1qwen35moe.block_count = 41
2qwen35moe.nextn_predict_layers = 1Ornith-1.0-35B-AEON-Ultimate-Uncensored-MTP-Q8_0.ggufgeneral.namegeneral.author = mrexodiageneral.quantized_by = mrexodiageneral.license = mitgeneral.license.name = MIT Licensegeneral.license.link = https://huggingface.co/deepreinforce-ai/Ornith-1.0-35B/blob/main/LICENSEgeneral.source.huggingface.repositorygeneral.descriptiongeneral.base_model.count = 2general.base_model.0.* for the AEON sourcegeneral.base_model.1.* for the Unsloth MTP donorgeneral.tagsgeneral.base_model was removed in favor of the interoperable general.base_model.{id}.name mapping used by HuggingFace GGUF metadata.qwen35moe.block_count = 41qwen35moe.nextn_predict_layers = 1blk.40. are presentblk.40.nextn.* are present--spec-type draft-mtpWhat is 2+2? Answer with just the number. produced 4