Based on Unsloth BF16 GGUF and imatrix file. The quantization is not programatically selected.
I carefully checked every detail of the imatrix statistics and obtain quantization suggestions from Qwen3-235B-A22B/DeepSeek V3.1/Gemini 2.5 Pro/ChatGPT.
Full protection of first 0-2 dense layers.
Full protection of output tensor and embedding layer.
Further compression is possible by llama.cpp.