InternVL3_5-1B is a 1-billion-parameter multimodal model developed by OpenGVLab. It processes both images and text, supporting vision-language tasks such as image captioning, visual question answering, and multimodal chat. The model is available on Hugging Face and optimized for local or self-hosted inference.
InternVL3 5 1B sits in PulseGate's Multimodal & vision category. It focuses on running efficient multimodal vision-language inference locally with a small 1B parameter model. It is built as an open-source project for developers. InternVL3 5 1B is open source under the MIT license. InternVL3 5 1B is available on the web, the command line, and API.
Behind InternVL3 5 1B is OpenGVLab, and it first shipped in 2023. The project is developed in the open on GitHub with 10.1k stars. Among its 4 catalogued features are vision-Language, Image Understanding, and Chat Templates.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do