This repository contains GGUF quantized versions of the Qwen3-VL-8B-Instruct model, enabling efficient local inference on CPUs and GPUs. It supports vision-language tasks including image understanding, visual question answering, and document analysis. The GGUF format makes it compatible with popular local LLM runners like llama.cpp.
Qwen3 VL 8B Instruct is a Multimodal & vision project. It focuses on running powerful multimodal vision-language models efficiently on consumer hardware. It is built as an open-source project for developers. Qwen3 VL 8B Instruct is open source under the MIT license. It ships for the web, the command line, and API.
It is developed by Qwen, and it first shipped in 2023. Development happens publicly on GitHub with 121.2k stars and 1.2k commits in the last 90 days. It competes in a saturated segment with 25 similar projects in PulseGate's index. Among its 3 catalogued features are Vision-Language Model, GGUF Quantization, and Multimodal Chat.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do