nvidia/Qwen3-8B-NVFP4 provides a quantized version of the Qwen3 8B model optimized for NVIDIA GPUs. It supports advanced features such as tool calling and is designed for high-performance local or server-side inference. The model is distributed via Hugging Face.
Qwen3 8B sits in PulseGate's Foundation models & chat category. It focuses on running the Qwen3 8B model efficiently on NVIDIA hardware using low-precision FP4 quantization. Qwen3 8B is an open-source project aimed at developers. The project is open source (Apache-2.0). Qwen3 8B is available on the web, API, and the command line.
Behind Qwen3 8B is NVIDIA, and it first shipped in 2024. The project is developed in the open on GitHub with 3.3k stars and 355 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 8 similar projects. Among its 3 catalogued features are Text Generation, FP4 Quantization, and Tool Calling Support.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do