A GPTQ 4-bit quantized version of the Qwen3-14B model, optimized for lower memory usage while maintaining strong performance. It includes support for tool calling and follows the Qwen chat template. Suitable for local inference on GPUs with limited VRAM using libraries such as transformers or vLLM.
In the Foundation models & chat space, Qwen3 14B takes a focused approach. It focuses on running a powerful 14B parameter language model efficiently on consumer hardware with reduced memory requirements. It is built as an open-source project for Local LLM users and developers. Qwen3 14B is open source under the Open Source license. It ships for the web, the command line, and API.
JunHowie builds and maintains Qwen3 14B, and it first shipped in 2024. Development happens publicly on GitHub with 27.5k stars. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Quantized LLM, GPTQ Int4, and tool calling support.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do