deepseek-ai/DeepSeek-OCR is an open-source model for optical character recognition, enabling extraction of text from images. It supports integration with PyTorch and Docker, making it suitable for document digitization and data extraction workflows. Designed for data scientists and developers working with scanned documents and images.
In the Foundation models & chat space, DeepSeek OCR takes a focused approach. It automates extraction of text from images for document processing and digitization. DeepSeek OCR is an open-source project aimed at data scientists. The project is open source (Open Source). The product ships for the command line, and it can be self-hosted.
It is developed by deepseek-ai, and the product first shipped in 2025. Among its 5 catalogued features are optical character recognition, image-to-text conversion, and pyTorch support.
Latest indexed changes and source events
deepseek-ai/DeepSeek-OCR verified by the PulseGate indexer
Other apps tracked under the same category.