This Hugging Face repository provides an NVFP4-quantized Qwen3.6 model checkpoint intended for efficient inference. It is aimed at developers and ML practitioners deploying language and multimodal model workloads locally or through inference infrastructure.
In the Foundation models & chat space, Qwen3.6 35B A3B takes a focused approach. It focuses on running a capable Qwen model with reduced memory requirements for local or hosted inference. It is built as an open-source project for developers and machine-learning practitioners. The project is open source (Apache-2.0). Qwen3.6 35B A3B is available on the web, the command line, and API, and it can be self-hosted.
It is developed by sakamakismile, and it first shipped in 2019. The project is developed in the open on GitHub with 3.8k stars and 180 commits in the last 90 days. It competes in a saturated segment with 25 similar projects in PulseGate's index. Among its 6 catalogued features are quantized weights, text generation, and vision input.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do