Qwen3-30B-A3B-GPTQ-Int4 is a GPTQ 4-bit quantized variant of the Qwen3 30B-A3B model. It supports advanced capabilities such as tool calling and follows a detailed chat template for conversational use. The model is optimized for local or server-based inference using tools like llama.cpp or Hugging Face Text Generation Inference while maintaining strong performance.
Qwen3 30B A3B sits in PulseGate's Quantised & converted weights category. It focuses on deploying a powerful 30B-parameter language model with significantly reduced memory footprint using GPTQ quantization. It is built as an open-source project for developers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by Qwen, and it first shipped in 2024. Development happens publicly on GitHub with 27.4k stars. It competes in a saturated segment with 23 similar projects in PulseGate's index. Among its 3 catalogued features are Quantized LLM, Tool Calling, and 4-bit Quantization.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do