This is a community-quantized GGUF version of Google's Gemma 4 12B instruction-tuned (it) model. It enables efficient local inference using tools such as llama.cpp, Ollama, and LM Studio. The model provides strong performance across general language tasks while being runnable on mid-range GPUs or high-end CPUs.
In the Text generation space, Gemma 4 12B It takes a focused approach. It focuses on running a capable 12B instruction-tuned model locally on consumer hardware. Gemma 4 12B It is an open-source project aimed at developers. The project is open source (MIT). Gemma 4 12B It is available on the web and the command line, and it can be self-hosted.
Behind Gemma 4 12B It is Google, based in the United States, and it first shipped in 2023. Development happens publicly on GitHub with 121.2k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable projects. Among its 3 catalogued features are Quantized GGUF, instruction tuned, and local inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do