Qwen3-VL-2B-Instruct-GGUF is a GGUF-quantized version of Alibaba's Qwen vision-language model. It enables local multimodal inference combining text and vision inputs for instruction following and chat. The model is designed for developers who want to run compact vision-language models on consumer hardware using standard GGUF runtimes.
Qwen3 VL 2B Instruct sits in PulseGate's Multimodal & vision category. It focuses on running efficient vision-language models locally without cloud dependency. It is built as an open-source project for developers. The project is open source (MIT). It ships for the web, the command line, and API.
Qwen builds and maintains Qwen3 VL 2B Instruct, and it first shipped in 2023. Development happens publicly on GitHub with 121k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable projects. Among its 4 catalogued features are vision-language, Quantized GGUF, and instruction tuned.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do