Refer to
Qwen3-VL-4B-Instruct for guidance on model inference acceleration and PDF processing, etc.
For high-performance inference and deployment, we recommend using
vLLM. We also provide a standalone script for efficiently processing multi-page PDF documents. This script operates independently and does not require the official olmOCR toolkit, offering a lightweight and fast way to perform OCR on entire documents.
We express our gratitude to the teams that developed
olmOCR and
Qwen3-VL, which were instrumental in our research.