opencode-vision is an open-source MCP server that empowers text-only AI models with vision capabilities. It integrates PaddleOCR for state-of-the-art image text extraction and uses Google Gemini as a fallback for model-agnostic image analysis. Designed for AI developers and researchers needing multimodal support.
opencode-vision sits in PulseGate's Frameworks & runtimes category. It focuses on enabling AI models to analyze images and extract text using a vision-empowered MCP server. It is built as an open-source project for AI developers and researchers. The project is open source (MIT). It ships for the web, the command line, and API, and it can be self-hosted.
NickRivers1983 builds and maintains opencode-vision, and it first shipped in 2026. The project is developed in the open on GitHub with 4 commits in the last 90 days. Key capabilities include image analysis, OCR integration, and MCP server. It exposes integrations via an MCP server and a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do