This is a post-training quantized (QAT) version of Google's Gemma 4 model with 4-bit weights and 16-bit activations (w4a16). It is designed for efficient inference while preserving model quality. The model is published on Hugging Face and intended for developers seeking to deploy Gemma 4 with reduced memory footprint.
Gemma 4 E4B It Qat W4a16 Ct sits in PulseGate's Foundation models & chat category. It focuses on running large Gemma 4 models efficiently on hardware with limited memory using quantization. It is built as an open-source project for developers. Gemma 4 E4B It Qat W4a16 Ct is open source under the Open Source license. It runs on the web, the command line, and API.
Behind Gemma 4 E4B It Qat W4a16 Ct is Google, and it first shipped in 2025. PulseGate's similarity index places it among 12 comparable projects.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do