Qwen3-30B-A3B-NVFP4 is an NVIDIA-optimized version of the Qwen3 model using NVFP4 quantization. It supports advanced tool calling and function calling through a detailed chat template. The model is designed for efficient inference on NVIDIA GPUs while preserving the capabilities of the original Qwen3 architecture.
In the Foundation models & chat space, Qwen3 30B A3B takes a focused approach. It focuses on running high-performance quantized versions of Qwen3 models on NVIDIA hardware. It is built as an open-source project for developers. Qwen3 30B A3B is open source under the Apache-2.0 license. Qwen3 30B A3B is available on the web and API.
It is developed by NVIDIA (United States), and the product first shipped in 2024. Development happens publicly on GitHub with 3.3k stars and 354 commits in the last 90 days. The category is crowded — PulseGate's index counts 23 comparable apps. Key capabilities include Tool Calling, Function Calling, and NVFP4 Quantization.
Latest indexed changes and source events
nvidia/Qwen3-30B-A3B-NVFP4 verified by the PulseGate indexer
Other apps tracked under the same category.