This repository contains GGUF quantized files for the Qwen3-4B model, enabling efficient local inference with tools such as llama.cpp. It includes support for tool calling and follows the Qwen chat template. The models are intended for developers seeking lightweight, locally runnable large language models.
Qwen3 4B is a Foundation models & chat project. It focuses on running the Qwen3 4B model efficiently on consumer hardware using GGUF quantization. Qwen3 4B is an open-source project aimed at developers. Qwen3 4B is open source under the Open Source license. It runs on the web, the command line, and API.
It is developed by MaziyarPanahi. It operates in a well-populated space: PulseGate tracks 13 similar projects. Among its 3 catalogued features are GGUF format, quantized model, and tool calling.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do