Skip to content
AI
Computer vision, OCR & document AI
Computer vision, OCR & document AI inside AI.
Newest through the gate
- OCRStackocrstack.comOCRStack extracts editable text from scanned PDFs, images, and documents for online use.
- Instruct X-Decoderhuggingface.coInstruct X-Decoder is a web-based computer vision demo for interactive image understanding and segmentation.
- ProteinMPNNhuggingface.coProteinMPNN designs protein sequences from uploaded structures or PDB codes.
- Face Recognition SDKhuggingface.coFace Recognition SDK compares two uploaded or captured images to estimate facial similarity.
- Convertisseur IAgithub.comConvertisseur IA provides local MCP connectors for document conversion to Markdown and Windows OCR for scanned files and photos.
- CORAatlassian.comCORA automatically labels and tags Confluence images with AI for easier content search and organization.
- SnapMeasureAIsnapmeasureai.comSnapMeasureAI creates 3D body scans and measurements from two smartphone photos.
- Tables Extractoratlassian.comTables Extractor extracts tables from PDFs and images and converts them into editable Confluence tables.
- Face Detectionwordpress.orgFace Detection is a WordPress plugin that centers detected faces in automatically cropped image thumbnails.
- vlm-toolkitpypi.orgvlm-toolkit provides batch image processing with vision-language models for captioning, watermark detection, and logo removal.
- Face recognition WPwordpress.orgFace recognition WP identifies people in group photos for WordPress site owners.
- SAM3D Body with Rerunhuggingface.coSAM3D Body with Rerun is a web app for analyzing and visualizing 3D human body data.
- ScouterAIhuggingface.coScouterAI analyzes uploaded images, detects objects, and annotates them with bounding boxes and labels.
- ChronoDepthhuggingface.coChronoDepth generates colored depth maps from uploaded videos for visualizing near and far regions.
- VisionScope-R2huggingface.coVisionScope-R2 analyzes uploaded images and answers questions with captions, OCR, and spatial descriptions.
- AI Image to Textatlassian.comAI Image to Text extracts printed and handwritten text from image attachments for Jira Cloud users.
- TextLens Proatlassian.comTextLens Pro extracts, highlights, and copies text from Jira image attachments using OCR.
- google-lens-propypi.orggoogle-lens-pro is a Python package for Google Lens visual search, OCR, and multimodal product intelligence.
- xberggithub.comxberg extracts text, metadata, images, tables, and structured data from documents and source code.
- Extract Text from Image – ETFIwordpress.orgETFI is a WordPress plugin that extracts text, handwriting, and structured data from images using Amazon Textract.
- Captionizeratlassian.comCaptionizer analyzes Confluence images and generates captions for teams using Atlassian Rovo.
- AI Screenshot Insights for Jiraatlassian.comAI Screenshot Insights for Jira extracts editable text and information from screenshot attachments using OCR.
- Irisnet API Clientwordpress.orgIrisnet API Client is a WordPress plugin that blocks or blurs unwanted images using AI.
- STAR.VISIONstar.visionSTAR.VISION provides space-based computing and AI satellite remote-sensing services for real-time geospatial data processing.
This is the newest 24 of 446. Open Computer vision, OCR & document AI in the live index →
Subscribe to Computer vision, OCR & document AI by RSS — new listings in this category, in your reader, no account.
Elsewhere in AI
Foundation models & chatCoding AI & assistantsImage generationVideo generationVoice, TTS & speechAutonomous agents & workflowsRAG, search & retrievalData science & ML workbenchFine-tuning & trainingLLM eval & observabilityInference & model servingOther AIWriting & editingAI security & guardrails3D generation