This repository contains GGUF quantized files for the Qwen3-4B model, enabling efficient local inference with tools such as llama.cpp. It includes support for tool calling and follows the Qwen chat template. The models are intended for developers seeking lightweight, locally runnable large language models.
In the Foundation models & chat space, Qwen3 4B takes a focused approach. It focuses on running the Qwen3 4B model efficiently on consumer hardware using GGUF quantization. It is built as an open-source project for developers. Qwen3 4B is open source under the Open Source license. Qwen3 4B is available on the web, the command line, and API.
It is developed by MaziyarPanahi. It operates in a well-populated space: PulseGate tracks 13 similar tools. Key capabilities include GGUF format, quantized model, and tool calling.
Latest indexed changes and source events
MaziyarPanahi/Qwen3-4B-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.