Qwen3-4B-AWQ is a 4-billion parameter quantized version of the Qwen3 large language model optimized for efficient inference. It supports text generation, tool calling, and can be used with the Hugging Face Transformers library or vLLM. The model is designed for developers who need a capable yet memory-efficient open-weights LLM that can run locally or be self-hosted.
Qwen3 4B is a Foundation models & chat product. It focuses on running large language models efficiently on consumer hardware with reduced memory requirements. Qwen3 4B is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by Qwen (China), and the product first shipped in 2024. The project is developed in the open on GitHub with 27.4k stars. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 4 catalogued features are Text Generation, Quantized Model, and Tool Calling.
Latest indexed changes and source events
Qwen/Qwen3-4B-AWQ verified by the PulseGate indexer
Other apps tracked under the same category.