dots.ocr is a multilingual vision-language model designed for document layout parsing and optical character recognition. Hosted on Hugging Face, it processes document images to understand layout elements and extract text in many languages using a single model. It is intended for developers integrating advanced document intelligence into applications via pip, Docker, or the Inference API.
Dots.ocr sits in PulseGate's Foundation models & chat category. Accurately parsing complex document layouts and extracting text across multiple languages in a single model. Dots.ocr is an open-source project aimed at developers. The project is open source (MIT). Dots.ocr is available on the web, the command line, and API, and it can be self-hosted.
It is developed by dots-studio, and the product first shipped in 2025. The project is developed in the open on GitHub with 9k stars. Among its 4 catalogued features are Document Layout Analysis, Multilingual OCR, and Vision-Language Model. It exposes integrations via a public API.
Latest indexed changes and source events
dots-studio/dots.ocr verified by the PulseGate indexer
Other apps tracked under the same category.