This is a 4-bit quantized GGUF variant of Google's Gemma-4 12B instruction-tuned (it) model. It enables efficient local inference using tools like llama.cpp or Ollama. The model supports chat and instruction following while fitting on more modest GPUs or CPUs.
In the Foundation models & chat space, Gemma 4 12B It Qat Q4 0 takes a focused approach. It focuses on running a capable 12B instruction model efficiently on consumer hardware with reduced memory usage. Gemma 4 12B It Qat Q4 0 is an open-source project aimed at developers. The project is open source (Open Source). Gemma 4 12B It Qat Q4 0 is available on the web, the command line, and API.
Behind Gemma 4 12B It Qat Q4 0 is google, and the product first shipped in 2025. It competes in a saturated segment with 25 similar apps in PulseGate's index.
Latest indexed changes and source events
google/gemma-4-12B-it-qat-q4_0-gguf verified by the PulseGate indexer
Other apps tracked under the same category.