This repository contains GGUF quantized versions of the Qwen3-VL-8B-Instruct model, enabling efficient local inference on CPUs and GPUs. It supports vision-language tasks including image understanding, visual question answering, and document analysis. The GGUF format makes it compatible with popular local LLM runners like llama.cpp.
In the Foundation models & chat space, Qwen3 VL 8B Instruct takes a focused approach. It focuses on running powerful multimodal vision-language models efficiently on consumer hardware. Qwen3 VL 8B Instruct is an open-source project aimed at developers. The project is open source (MIT). Qwen3 VL 8B Instruct is available on the web, the command line, and API.
Qwen builds and maintains Qwen3 VL 8B Instruct, and the product first shipped in 2023. The project is developed in the open on GitHub with 121.2k stars and 1.2k commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are Vision-Language Model, GGUF Quantization, and Multimodal Chat.
Latest indexed changes and source events
Qwen/Qwen3-VL-8B-Instruct-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.