Qwen3-30B-A3B-GPTQ-Int4 is a GPTQ 4-bit quantized variant of the Qwen3 30B-A3B model. It supports advanced capabilities such as tool calling and follows a detailed chat template for conversational use. The model is optimized for local or server-based inference using tools like llama.cpp or Hugging Face Text Generation Inference while maintaining strong performance.
Qwen3 30B A3B sits in PulseGate's Foundation models & chat category. It focuses on deploying a powerful 30B-parameter language model with significantly reduced memory footprint using GPTQ quantization. Qwen3 30B A3B is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web, the command line, and API.
Behind Qwen3 30B A3B is Qwen, and the product first shipped in 2024. The project is developed in the open on GitHub with 27.4k stars. It competes in a saturated segment with 23 similar apps in PulseGate's index. Among its 3 catalogued features are Quantized LLM, Tool Calling, and 4-bit Quantization.
Latest indexed changes and source events
Qwen/Qwen3-30B-A3B-GPTQ-Int4 verified by the PulseGate indexer
Other apps tracked under the same category.