The model is distilled from Qwen2-VL-72B-Instruct and trained on a corpus of 17.6 million pages / 45.5 billion tokens.
🟢 Why it matters ?
- 1B parameters
- allows processing 5.7 pages/s on a single H100 (which is approximately ≈ 493,000 pages per day)
- Recognizes tables, forms, equations, and complex layouts
- 6.5× faster than dots.ocr, 1.7× faster than DeepSeekOCR
- Costs < $0.01 per 1000 A4 pages
- Outperforms DeepSeekOCR
- Comparable to dots.ocr (while the model is 3 times smaller in size)
- +16 points over Qwen3-VL-2B-Instruct
This model is an excellent balance of quality, speed, and cost.
Model 1B, Model 0.9B (32k), LightOn Blog, Demo
#OCR #ML
🤖 Data Science, ML & Big Data with @DataXplore
