This is a 4-bit quantized (w4a16) version of the Gemma 3 27B instruction-tuned model, optimized for efficient inference. It uses the same chat template and capabilities as the original but with reduced memory requirements. The model is published by RedHatAI on Hugging Face.
In the Foundation models & chat space, Gemma 3 27b It Quantized.w4a16 takes a focused approach. It focuses on running large language models efficiently on hardware with limited memory using 4-bit quantization. Gemma 3 27b It Quantized.w4a16 is an open-source project aimed at developers. The project is open source (Apache-2.0). Gemma 3 27b It Quantized.w4a16 is available on the web, the command line, and API.
RedHatAI builds and maintains Gemma 3 27b It Quantized.w4a16, and the product first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 171 commits in the last 90 days.
Latest indexed changes and source events
RedHatAI/gemma-3-27b-it-quantized.w4a16 verified by the PulseGate indexer
Other apps tracked under the same category.