This Hugging Face repository provides a quantized Qwen language-model checkpoint using NVFP4 and BF16 components, including a language-model head. It is intended for developers and researchers deploying or experimenting with local model inference.
In the Quantised & converted weights space, Qwen3.8 27B NVFP4 BF16 LMHead takes a focused approach. It focuses on running a large language model locally with reduced-precision weights for lower memory and hardware requirements. It is built as an open-source project for machine-learning developers and researchers. The project is open source (Apache-2.0). It runs on the command line, and it can be self-hosted.
Behind Qwen3.8 27B NVFP4 BF16 LMHead is RadixArk, and it first shipped in 2024. Development happens publicly on GitHub with 3.7k stars and 321 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 17 similar projects. Key capabilities include quantized weights, local inference, and chat template.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do