Views
No views yet
Qwen3_VL_8B_Instruct_DPE_v3 is an enhanced version of the Qwen3-VL-8B-Instruct base model, evolved through the DPE (Diagnostic-driven Progressive Evolution) framework after three complete iterations.Qwen3_VL_8B_Instruct_DPE_v3 compared to the base model across 11 benchmarks:| Category | Benchmark | Base Model | DPE_v3 (Ours) | Improvement |
|---|---|---|---|---|
| STEM | MMMU | 65.44 | 69.11 | +3.67 |
| MMVet | 67.29 | 72.80 | +5.51 | |
| MMStar | 61.27 | 72.13 | +10.86 | |
| Visual Math | MathVerse | 53.22 | 57.18 | +3.96 |
| MathVision | 51.97 | 53.88 | +1.91 | |
| OCR | CharXiv (RQ) | 47.20 | 48.10 | +0.90 |
| Specialized | BLINK | 68.54 | 69.22 | +0.68 |
| Overall | Average | 65.64 | 68.04 | +2.40 |
1@misc{jia2026blindspotsgainsdiagnosticdriven,
2 title={From Blind Spots to Gains: Diagnostic-Driven Iterative Training for Large Multimodal Models},
3 author={Hongrui Jia and Chaoya Jiang and Shikun Zhang and Wei Ye},
4 year={2026},
5 eprint={2602.22859},
6 archivePrefix={arXiv},
7 primaryClass={cs.CV},
8 url={https://arxiv.org/abs/2602.22859 },
9}