multi-ocr-py is an open-source toolkit that enables users to extract text from PDFs and images and convert it into Markdown format. It supports multiple OCR engines and can be used both as a command-line tool and as a Python SDK, making it suitable for developers and technical users who require automated OCR workflows.
multi-ocr-py sits in PulseGate's CLI tools & terminal category. It focuses on extracting and converting text from PDFs and images to Markdown using multiple OCR engines via CLI or SDK. multi-ocr-py is an open-source project aimed at developers and technical users needing OCR automation. The project is open source (MIT). The product ships for the command line and API, and it can be self-hosted.
It is developed by BlackBoxRecorder, and the product first shipped in 2026. The project is developed in the open on GitHub with 21 commits in the last 90 days. Among its 7 catalogued features are Multi-engine OCR, PDF to Markdown, and CLI usage.
Latest indexed changes and source events
Other apps tracked under the same category.