This is a 4-bit quantized (w4a16) version of the Gemma 3 27B instruction-tuned model, optimized for efficient inference. It uses the same chat template and capabilities as the original but with reduced memory requirements. The model is published by RedHatAI on Hugging Face.
Gemma 3 27b It Quantized.w4a16 sits in PulseGate's Text generation category. It focuses on running large language models efficiently on hardware with limited memory using 4-bit quantization. It is built as an open-source project for developers. The project is open source (Apache-2.0). Gemma 3 27b It Quantized.w4a16 is available on the web, the command line, and API.
RedHatAI builds and maintains Gemma 3 27b It Quantized.w4a16, and it first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 171 commits in the last 90 days.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do