Views
No views yet
[!WARNING] A significantly improved version of this model is available.This repo quantizes huihui-ai/Huihui-Qwen3.5-35B-A3B-abliterated — the standard abliteration. A newer fine-tune of the same architecture, trained in the style of Claude 4.6 Opus, has since been released and produces noticeably richer, more expressive outputs.➡️ Recommended upgrade: Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M-GGUFSame architecture, same Q4_K_M quantization, same VRAM footprint — just a better fine-tune. This repo will remain available for reference.
| Property | Value |
|---|---|
| Source model | huihui-ai/Huihui-Qwen3.5-35B-A3B-abliterated |
| Architecture | qwen35moe (35B total params, ~3B active; 256 experts, 8 active per token) |
| Quantization | Q4_K_M (~4.8 BPW) |
| File size | ~21 GB |
| Quantized with | llama.cpp |
1llama-cli \
2 --hf-repo cesarsal1nas/Huihui-Qwen3.5-35B-A3B-abliterated-Q4_K_M-GGUF \
3 --hf-file huihui-qwen3.5-35b-a3b-abliterated-Q4_K_M.gguf \
4 -p "Tell me about the universe"1llama-server \
2 --hf-repo cesarsal1nas/Huihui-Qwen3.5-35B-A3B-abliterated-Q4_K_M-GGUF \
3 --hf-file huihui-qwen3.5-35b-a3b-abliterated-Q4_K_M.gguf \
4 -c 8192ollama run hf.co/cesarsal1nas/Huihui-Qwen3.5-35B-A3B-abliterated-Q4_K_M-GGUF