Qwen3-4B-Base Blind Spots Dataset: A Ghana-Focused Evaluation
Model Tested
Qwen/Qwen3-4B-Base — A 4B parameter pretrained causal language model from the Qwen3 family, released in May 2025 by the Qwen Team (Alibaba). The model was pre-trained on 36 trillion tokens across 119 languages, with a context length of 32,768 tokens. It features 36 layers, 32 attention heads for Q and 8 for KV (GQA), and 3.6B non-embedding parameters.