This is a community or official variant of Google's Gemma 4 12B instruction-tuned (it) model with QAT (Quantization-Aware Training) and Q4_0 quantization options. It includes an advanced chat template supporting tool calling and structured output. The model can be used with Transformers for local inference on consumer hardware.
Gemma 4 12B It Qat Q4 0 Unquantized Assistant is a Foundation models & chat product. It focuses on running large instruction-tuned language models locally with reduced memory footprint via quantization. It is built as an open-source project for developers. Gemma 4 12B It Qat Q4 0 Unquantized Assistant is open source under the Open Source license. Gemma 4 12B It Qat Q4 0 Unquantized Assistant is available on the web, the command line, and API.
Google builds and maintains Gemma 4 12B It Qat Q4 0 Unquantized Assistant, and the product first shipped in 2026. The category is crowded — PulseGate's index counts 25 comparable apps. Key capabilities include Instruction Tuned, Tool Calling, and Chat Template.
Latest indexed changes and source events
google/gemma-4-12B-it-qat-q4_0-unquantized-assistant verified by the PulseGate indexer
Other apps tracked under the same category.