This is a quantized version of the Qwen3-VL-4B vision-language model optimized with AWQ 4-bit precision. It enables efficient multimodal inference combining vision and text understanding. The model supports instruction following and tool calling, making it suitable for developers building local multimodal applications with lower hardware requirements.
Qwen3 VL 4B Instruct is a Multimodal & vision project. It focuses on running large vision-language models with reduced memory and compute requirements on local hardware. It is built as an open-source project for developers. The project is open source (Open Source). Qwen3 VL 4B Instruct is available on the web, the command line, and API.
cyankiwi builds and maintains Qwen3 VL 4B Instruct. The category is crowded — PulseGate's index counts 25 comparable projects. Key capabilities include vision-language understanding, 4-bit quantization, and instruction following.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do