This is a 4-bit AWQ quantized version of the Gemma-4-E4B instruction-tuned model. It enables efficient local inference of a capable open-weight LLM using significantly less VRAM than the original. The model supports chat-based interactions via a provided Jinja chat template and is intended for developers integrating language capabilities into applications or running models on consumer hardware.
Gemma 4 E4B It sits in PulseGate's Foundation models & chat category. It focuses on running large language models with reduced memory requirements on local hardware. Gemma 4 E4B It is an open-source project aimed at developers. The project is open source (Open Source). The product ships for the web, the command line, and API.
Chunity builds and maintains Gemma 4 E4B It, and the product first shipped in 2025. Among its 3 catalogued features are quantized weights, chat template, and token classification.
Latest indexed changes and source events
Chunity/gemma-4-E4B-it-AWQ-4bit verified by the PulseGate indexer
Other apps tracked under the same category.