A community-quantized (W4A16) version of a Gemma 4 model tuned for instruction following. It provides efficient inference while maintaining most of the original model's capabilities. The model is hosted on Hugging Face and includes a custom chat template for structured interactions.
In the Quantised & converted weights space, Gemma 4 E4B It W4A16 takes a focused approach. It focuses on running Gemma-based instruction models with reduced memory footprint through quantization. It is built as an open-source project for developers and researchers. The project is open source (Open Source). It runs on the web and API.
Abhishek Chohan builds and maintains Gemma 4 E4B It W4A16, and it first shipped in 2025. The category is crowded — PulseGate's index counts 25 comparable projects. Key capabilities include Instruction Tuned, Quantized Inference, and Text Generation.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do