Qwen3-30B-A3B-Instruct-2507-FP8 is an FP8 quantized variant of Alibaba's Qwen3 30B-A3B Mixture-of-Experts instruct model. It supports advanced features such as tool calling and is optimized for efficient inference. The model is hosted on Hugging Face and is suitable for developers needing high-performance language capabilities with lower memory footprint.
In the Foundation models & chat space, Qwen3 30B A3B Instruct 2507 takes a focused approach. It focuses on deploying a large-scale Mixture-of-Experts language model efficiently with reduced precision for faster inference. Qwen3 30B A3B Instruct 2507 is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by Qwen, and it first shipped in 2024. The project is developed in the open on GitHub with 27.4k stars. It operates in a well-populated space: PulseGate tracks 12 similar projects. Key capabilities include mixture-of-Experts, instruct-tuned, and FP8 quantized.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do