PulseGateLive intelligence on AI-era software
Coverage
175,634

Software tracked

 

Freshness
21 min ago

Last update

 

Cadence
488/day

7-day average

Indexed today: 709

PulseGate

Live intelligence on the software shipping in the AI era — apps, models, agents, and infra.

Software is shipping faster than ever, and a growing share of it lives outside the official app stores. PulseGate tracks it live — free, for builders, analysts, and everyone keeping up.

Follow
GitHubX (Twitter)LinkedIn

Platform

  • All Apps
  • Full Index
  • Categories
  • Industry Updates
  • Data Sources
  • Coverage Rules
  • Glossary
  • Embed Widget

Support

  • Help Center
  • Suggest a URL
  • Report an Issue

Company

  • About
  • Dimaxia
  • Press & Data
  • Contact
  • Platform Status

Legal

  • Privacy
  • Terms
  • Disclaimer

© 2026 PulseGate. Operated by Dimaxia, a brand of Dymaxio s.r.o., Prague, Czech Republic.·

All systems operational
PulseGate
MU
multi-ocr-py
Visit ↗
  1. Home/
  2. CLI tools & terminal/
  3. multi-ocr-py
←Back to results
M

multi-ocr-py

PyPI·Infrastructure

multi-ocr-py is an open-source toolkit that enables users to extract text from PDFs and images and convert it into Markdown format. It supports multiple OCR engines and can be used both as a command-line tool and as a Python SDK, making it suitable for developers and technical users who require automated OCR workflows.

Open SourceMIT
CLIAPISelf-hosted
M
Visit PyPI↗

Overview

7 features

multi-ocr-py sits in PulseGate's CLI tools & terminal category. It focuses on extracting and converting text from PDFs and images to Markdown using multiple OCR engines via CLI or SDK. multi-ocr-py is an open-source project aimed at developers and technical users needing OCR automation. The project is open source (MIT). The product ships for the command line and API, and it can be self-hosted.

It is developed by BlackBoxRecorder, and the product first shipped in 2026. The project is developed in the open on GitHub with 21 commits in the last 90 days. Among its 7 catalogued features are Multi-engine OCR, PDF to Markdown, and CLI usage.

  • ✓Multi-engine OCR
  • ✓PDF to Markdown
  • ✓CLI usage
  • ✓SDK integration
  • ✓Image text extraction
  • ✓Markdown output
  • ✓MIT license

Tags

ocr-toolkitpdf-to-markdownvision-language-modelcli-ocrmulti-engine-ocr

AI capabilities

TextImageWeights: Open

Built with & integrations

AI providers
openai
Runs on
CLIAPI-onlySelf-hosted

Trust & compliance

LicenseMIT
Verified signals
✓ HTTPS✓ Open Source✓ Free tier✓ GitHub✓ Active maintenance

Recent events

Latest indexed changes and source events

  1. IndexedJul 16, 1:35 AM

    multi-ocr-py verified by the PulseGate indexer

    Source: PulseGate indexerOpen ↗

Frequently asked questions about multi-ocr-py

What is multi-ocr-py?
Multi-ocr-py focuses on extracting and converting text from PDFs and images to Markdown using multiple OCR engines via CLI or SDK. It is catalogued under CLI tools & terminal on PulseGate.
Who should use multi-ocr-py?
multi-ocr-py is an open-source project built for developers and technical users needing OCR automation.
Does multi-ocr-py have a free plan?
Yes — multi-ocr-py is open source under the MIT license and free to use.
What platforms does multi-ocr-py run on?
multi-ocr-py runs on the command line and API. It can also be self-hosted.
Is multi-ocr-py still active?
PulseGate's liveness checks currently classify multi-ocr-py as active. The GitHub repository shows 21 commits in the last 90 days.
What tools are similar to multi-ocr-py?
Similar tools tracked by PulseGate include pagewise-pdf-extractor, pdf2md-tool, and omniocr.pagewise-pdf-extractorpdf2md-toolomniocr
Who develops multi-ocr-py?
multi-ocr-py is developed by BlackBoxRecorder.
How long has multi-ocr-py been around?
multi-ocr-py first shipped in 2026.

At a glance

Platforms
Cli
Languages
Chinese
Open source
Yes (GitHub)
License
MIT
First seen
Jul 16, 2026
Activity
🟢 Active
Status
🟢 Active
Built for
developers and technical users needing OCR automation
Model
Open source
Solves
Extracting and converting text from PDFs and images to Markdown using multiple OCR engines via CLI or SDK.

Developer

BlackBoxRecorder
Solo developer
↗ GitHub

Open source

View on GitHub →
⭐ Stars
0
🍴 Forks
0
Open issues
0
Last commit
6d ago
Commits 90d
21
Contributors
1
Authorship
Solo
Default branch
main

Live coverage

Confidence
Low · 64
Indexed
Jul 16, 2026
Lifecycle
Alive
Activity
Active
First seen
Jul 2026
Last seen
6d ago
Identity audit (9)
Entity ID
cmrmu7cmq03bu10g35e60a3r6
Slug
multi-ocr-py-pypi-org
Verification state
Indexed for public listing
Claim / listing state
Unclaimed · listed: yes
Index status
Included in index
Latest evidence snapshot
Jul 16, 2026
Timeline basis
Indexed-at chronology (no inferred launch/funding milestones).
Last updated
Jul 16, 2026
Canonical URL
https://pypi.org/project/multi-ocr-py

Similar apps

Other apps tracked under the same category.

  • pagewise-pdf-extractor
    github.com
  • pdf2md-tool
    pypi.org
  • omniocr
    omniocr.ai
  • Multimodal OCR3
    huggingface.co
  • pdf2md-uv
    github.com
  • upspawn-ocr-cli
    github.com