LayoutLMv3 Large is a pre-trained multimodal model developed by Microsoft that integrates text, layout, and image information for document understanding tasks. It excels at layout analysis, form understanding, receipt parsing, and visual question answering on scanned or digital documents. The model is distributed via Hugging Face and is used by developers building document AI pipelines, often fine-tuned for specific enterprise document processing needs.
Layoutlmv3 Large is an Other AI product. It focuses on understanding and extracting information from complex document images that combine text, layout, and visual elements. Layoutlmv3 Large is an open-source project aimed at developers. The project is open source (Apache-2.0). Layoutlmv3 Large is available on the web and API, and it can be self-hosted.
Microsoft builds and maintains Layoutlmv3 Large, and the product first shipped in 2018. The project is developed in the open on GitHub with 162.7k stars and 758 commits in the last 90 days. Among its 4 catalogued features are document layout analysis, visual question answering, and information extraction. It exposes integrations via a public API.
Latest indexed changes and source events
microsoft/layoutlmv3-large verified by the PulseGate indexer
Other apps tracked under the same category.