Qwen3-32B-NVFP4 is an FP4 quantized version of the Qwen3-32B model created by NVIDIA. It includes support for tool calling and is optimized for high-performance inference on NVIDIA GPUs. The model is distributed on Hugging Face for developers seeking state-of-the-art performance with reduced memory and compute requirements.
Qwen3 32B sits in PulseGate's Foundation models & chat category. It focuses on enabling efficient inference of large 32B parameter models on NVIDIA hardware using advanced low-precision formats. It is built as an open-source project for developers. Qwen3 32B is open source under the Open Source license. The product ships for the web and API.
Behind Qwen3 32B is NVIDIA, based in the United States, and the product first shipped in 2025. It operates in a well-populated space: PulseGate tracks 18 similar tools. Key capabilities include FP4 quantization, tool calling, and high performance.
Latest indexed changes and source events
nvidia/Qwen3-32B-NVFP4 verified by the PulseGate indexer
Other apps tracked under the same category.