Qwen3-4B-Base Blind Spots Dataset
Model Tested
Model: Qwen/Qwen3-4B-Base
Type: Causal Language Model — pretrained base model (NOT instruction-tuned)
Parameters: 4.0 billion (3.6B non-embedding)
Architecture: 36 layers, 32 attention heads (GQA: 32 Q / 8 KV)
Context Length: 32,768 tokens
Training: 36 trillion tokens across 119 languages in a 3-stage pretraining pipeline
This dataset documents 10 confirmed blind spots of Qwen3-4B-Base identified… See the full description on the dataset page:
https://huggingface.co/datasets/abdulmatinomotoso/Fatimah_Fellowship_Blind_Spot.