layoutlm-document-qa is a LayoutLM-based model fine-tuned for document question answering. It processes both the textual content and spatial layout of documents (PDFs, scanned forms, invoices) to answer questions that require understanding of visual structure and reading order. The model is available on Hugging Face with full Transformers integration and has been widely downloaded for enterprise document intelligence applications.
Layoutlm Document Qa is a Foundation models & chat project. It focuses on answering natural language questions about the content and structure of scanned or digital documents that contain both text and layout information. It is built as an open-source project for developers. The project is open source (Apache-2.0). It runs on the web and the command line, and it can be self-hosted.
Behind Layoutlm Document Qa is Impira, and it first shipped in 2018. Development happens publicly on GitHub with 163.3k stars and 777 commits in the last 90 days. Among its 4 catalogued features are Document QA, Layout Understanding, and Visual Question Answering. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do