This is a 4B parameter quantized version of the Qwen3-VL vision-language model optimized for MLX. It supports multimodal inputs combining text and images and can be used for visual question answering, document understanding, and tool-calling tasks. The model is distributed on Hugging Face for local inference via libraries such as Transformers or MLX.
Qwen3 VL 4B Instruct sits in PulseGate's Multimodal & vision category. It focuses on running advanced multimodal AI models locally with reduced memory requirements. It is built as an open-source project for developers. Qwen3 VL 4B Instruct is open source under the MIT license. It runs on the web, the command line, and API.
lmstudio-community builds and maintains Qwen3 VL 4B Instruct, and it first shipped in 2024. The project is developed in the open on GitHub with 5.2k stars and 419 commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable projects. Among its 4 catalogued features are Vision-Language Model, Quantized Weights, and MLX Format.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do