GLM-OCR is a model from the GLM family optimized for optical character recognition and document parsing tasks. It accepts image inputs and outputs extracted text or structured data. The model is hosted on Hugging Face and can be used with standard inference libraries.
GLM OCR sits in PulseGate's Multimodal & vision category. It focuses on extracting structured text and information from document images. It is built as an open-source project for developers. GLM OCR is open source under the Apache-2.0 license. It ships for the web, the command line, and API.
It is developed by unsloth, and it first shipped in 2026. Development happens publicly on GitHub with 7.2k stars. Among its 3 catalogued features are OCR, Document Understanding, and Multimodal Input.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do