Views
No views yet
Agents-A1-4B-bf16enhanced=True)oq_imatrix_report.json records the resulting quantization allocation.python run_oqe_conversion.py Agents-A1-4B-bf16 --models-dir /path/to/models --oq-level 6.0 --imatrix-samples 128 --imatrix-seq-length 512 --mtp-donor /path/to/Qwen3.5-4B-MTP-donor --mtp-source-repo https://huggingface.co/guru87/Qwen3.5-4B-MTP --source-repo https://huggingface.co/InternScience/Agents-A1-4BQwen3.5-4B-MTP-donormodel.safetensors: quantized MLX weightsmodel-mtp.safetensors: shipped-precision Lightning MTP headconfig.json: model configurationoq_imatrix_report.json: imatrix calibration and quantization report