Qwen3.5-4B trained with Vision-OPD using random 5% visual-token retention for the student and full visual tokens for the teacher. This is the 30-step proof-of-concept checkpoint.
The checkpoint can run with full visual tokens using standard Transformers. For 5% visual-token inference, use the pruning-aware serving code in
prune-opd.