This is a GGUF-quantized version of Gemma-4-E2B-it optimized by Unsloth for local CPU/GPU inference. It supports instruction following and is compatible with llama.cpp and similar runtimes. The model is intended for developers seeking efficient deployment of Gemma 4 without cloud dependency.
In the Foundation models & chat space, Gemma 4 E2B It Qat takes a focused approach. It focuses on running large language models efficiently on consumer hardware with reduced memory requirements. It is built as an open-source project for developers. Gemma 4 E2B It Qat is open source under the Apache-2.0 license. It runs on the web, the command line, and API.
Unsloth builds and maintains Gemma 4 E2B It Qat, and the product first shipped in 2023. Development happens publicly on GitHub with 68.7k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Quantized GGUF, Instruction Tuned, and Local Inference. It exposes integrations via a public API.
Latest indexed changes and source events
unsloth/gemma-4-E2B-it-qat-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.