This is a GGUF-quantized version of Gemma-4-E2B-it optimized by Unsloth for local CPU/GPU inference. It supports instruction following and is compatible with llama.cpp and similar runtimes. The model is intended for developers seeking efficient deployment of Gemma 4 without cloud dependency.
Gemma 4 E2B It Qat is a Foundation models & chat project. It focuses on running large language models efficiently on consumer hardware with reduced memory requirements. It is built as an open-source project for developers. Gemma 4 E2B It Qat is open source under the Apache-2.0 license. Gemma 4 E2B It Qat is available on the web, the command line, and API.
Unsloth builds and maintains Gemma 4 E2B It Qat, and it first shipped in 2023. Development happens publicly on GitHub with 68.7k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable projects. Among its 3 catalogued features are Quantized GGUF, Instruction Tuned, and Local Inference. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do