SOTA for multilingual document parsing, supports almost any writing system.
➡️ Features:
☞ Elo 1089 on olmOCR-Bench and 1157 on XDocParse: higher than GLM-OCR and PaddleOCR-VL-1.5
☞ Outperforms Qwen3-VL-235B (0.069) and Gemini 2.5 Pro (0.075) On OmniDocBench (text edit 0.031),
☞ Can generate SVG code for graphs, diagrams, and chemical formulas
☞ Supports web page parsing, text recognition in scenes, and object counting
☞ Works through vLLM and runs on a single GPU
Model, GitHub, Demo | #Utility
••••••••••••••••••••••••••••••••••••••
🤖 Data & ML | @DataXplore
