Qwen3-235B-A22B-NVFP4 is a massive open-weights Mixture-of-Experts (MoE) model from the Qwen3 family, quantized to NVFP4 precision for NVIDIA hardware. It supports advanced capabilities including tool calling and multi-step reasoning. The model is hosted on Hugging Face and is intended for local or self-hosted deployment by developers building applications that require high-capability language understanding and agentic behavior.
In the Foundation models & chat space, Qwen3 235B A22B takes a focused approach. It focuses on running very large language models efficiently on NVIDIA GPUs with reduced precision for lower memory usage and faster inference. Qwen3 235B A22B is an open-source project aimed at developers. The project is open source (Apache-2.0). It runs on the web, the command line, and API, and it can be self-hosted.
It is developed by NVIDIA, and it first shipped in 2024. Development happens publicly on GitHub with 3.3k stars and 350 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 18 similar projects. Key capabilities include Tool Calling, Function Calling, and Multi-step Reasoning.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do