This is a 4-bit quantized GGUF version of Google's Gemma-4 2B (E2B) instruction-tuned model. The QAT (Quantization Aware Training) and GGUF format allow efficient local inference using tools like llama.cpp. It is designed for on-device or local LLM applications.
Gemma 4 E2B It Qat Q4 0 sits in PulseGate's Foundation models & chat category. It focuses on running powerful instruction-tuned language models locally with reduced memory requirements. Gemma 4 E2B It Qat Q4 0 is an open-source project aimed at developers. The project is open source (Open Source). The product ships for the web, the command line, and API.
Google builds and maintains Gemma 4 E2B It Qat Q4 0, and the product first shipped in 2025. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are Quantized Model, Instruction Tuned, and GGUF Format.
Latest indexed changes and source events
google/gemma-4-E2B-it-qat-q4_0-gguf verified by the PulseGate indexer