Unlimited-OCR MLX is a high-precision OCR solution that fully migrates the Baidu PaddlePaddle team's Unlimited-OCR model to the Apple MLX framework.
Based on the DeepSeek-V2 architecture, combined with SAM-ViT-B + CLIP-L dual vision encoders, it can parse documents of any length in a single pass, implementing end-to-end text recognition and structured extraction.
✨ Core Features
Feature
Description
📄 Document Parsing
Supports full-page OCR for PDFs and single/multi-page images
🌍 Multilingual Recognition
Precise recognition of Chinese, English, and other multilingual text
📊 Table Extraction
Automatically recognizes and structures table content
🎯 Layout Analysis
Preserves original layout structure (paragraphs, headings, lists, etc.)
🔄 Unlimited Length
Dynamic image tiling, no document length restrictions