A 4-bit quantized GGUF version of Google's Gemma-4 4B instruction-tuned (it) model. It enables efficient local inference on consumer hardware using tools like llama.cpp or LM Studio. The model supports standard chat templates and is suitable for developers seeking lightweight, high-performance open models.
In the Foundation models & chat space, Gemma 4 E4B It Qat Q4 0 takes a focused approach. It focuses on running large language models locally with reduced memory and compute requirements. Gemma 4 E4B It Qat Q4 0 is an open-source project aimed at developers. The project is open source (Open Source). The product ships for the web, the command line, and API.
Behind Gemma 4 E4B It Qat Q4 0 is Google, and the product first shipped in 2025. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are GGUF format, 4-bit quantized, and instruction tuned.
Latest indexed changes and source events
google/gemma-4-E4B-it-qat-q4_0-gguf verified by the PulseGate indexer