A BF16 Qwen2 GGUF model for higher-quality FastVLM language decoding. This
model provides improved numerical fidelity compared to aggressive quantization
and is suitable for BF16-capable inference.
Model file
fastvlm_qwen2_bf16.gguf
Usage
Use this model for higher-quality multimodal reasoning when BF16 performance
is available.