This is a GGUF quantized version of the Gemma-4-E4B instruct model published by Unsloth. It enables efficient CPU and GPU inference of a capable open-weight language model using tools like llama.cpp or Ollama. The model is designed for developers and researchers who want to run high-performance instruction-tuned LLMs locally with reduced memory requirements.
Gemma 4 E4B It is a Foundation models & chat product. It focuses on running large language models efficiently on consumer hardware without high-end GPUs. It is built as an open-source project for developers. Gemma 4 E4B It is open source under the Apache-2.0 license. Gemma 4 E4B It is available on the web, the command line, and API.
Unsloth builds and maintains Gemma 4 E4B It, and the product first shipped in 2023. Development happens publicly on GitHub with 68.7k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include GGUF Quantization, Local Inference, and Instruct Model.
Latest indexed changes and source events
unsloth/gemma-4-E4B-it-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.