Qwen3-30B-A3B-GGUF contains GGUF quantized files for the 30B-A3B Mixture-of-Experts version of the Qwen3 model. These quantized weights enable efficient local inference on a wide range of hardware. The model supports advanced capabilities including tool calling and complex reasoning while maintaining a manageable memory footprint.
In the Foundation models & chat space, Qwen3 30B A3B takes a focused approach. Efficiently running the Qwen3 30B-A3B MoE model locally using GGUF quantization. It is built as an open-source project for developers. The project is open source (Open Source). It ships for the web, the command line, and API.
It is developed by Maziyar Panahi, and it first shipped in 2025. It operates in a well-populated space: PulseGate tracks 16 similar projects. Among its 4 catalogued features are GGUF format, mixture-of-Experts, and quantized weights.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do