DeepSeek has released a powerful OCR model for text recognition, capable of converting document images directly into Markdown or text.
🟢 Features:
- Recognizes text in images and PDFs
- Works with documents, tables, and complex layouts
- Supports different modes: Tiny, Small, Base, Large
- Optimized for GPU (PyTorch + CUDA 11.8)
- MIT license — free to use and modify
DeepSeek-OCR achieves high accuracy and efficiency through visual token compression. On Omnidocbench, it has the best accuracy with minimal visual tokens, outperforming other OCR models in efficiency and speed.
Can find HF, Github, Paper
#OCR #DeepSeek
🤖 Data Science, ML & Big Data with @DataXplore