Views
No views yet
| File | Quant | Size | Quality Loss |
|---|---|---|---|
| twentyq-f32.gguf | F32 | 762 KB | 0% |
| twentyq-f16.gguf | F16 | 397 KB | 0% |
| twentyq-q8_0.gguf | Q8_0 | 228 KB | 0% |
| twentyq-q4_0.gguf | Q4_0 | 135 KB | 0% |
| twentyq-q2_k.gguf | Q2_K | 95 KB | 0% |
general.architecture: twentyq
twentyq.block_count: 0
twentyq.embedding_length: 156
twentyq.attention.head_count: 156
twentyq.context_length: 20
twentyq.vocab_size: 1200output.weight) contains the entire model.twentyq architecture support, which does not currently exist in llama.cpp, ollama, or any other GGUF runtime. For inference, use the original model via the transformers library, or the live demo.