Gemma 4 12B It Qat W4a16 Ct is a quantized instruction-tuned language model hosted on Hugging Face. The model belongs to the Gemma 4 family and uses a specific quantization configuration indicated by its name, w4a16 with ct, along with a provided chat template for structured interactions.
Its page supplies tokenizer configuration that defines special tokens including bos_token, eos_token, mask_token, pad_token, and unk_token. A chat_template_jinja is included with a macro for formatting parameters that handles properties such as description, type, and enum values for tool-calling and conversation formatting. The template notes updates for fixed tool-calling loops, turn closures, and thinking content-ordering, and is attributed to the Google Gemma Engineering Team with a listed publication date.
The model is delivered as a repository on the Hugging Face platform, where users can access the model files, tokenizer, and associated configuration for local or hosted inference. It forms one implementation within the class of foundation models focused on text-based generation and instruction following.
Gemma 4 12B It Qat W4a16 Ct is a Foundation models & chat product. It focuses on enabling efficient, instruction-following text generation with a quantized open-source model. Gemma 4 12B It Qat W4a16 Ct is an open-source project aimed at AI developers and researchers. The project is open source (Open Source). The product ships for the web, the command line, and API, and it can be self-hosted.
It is developed by Google, and the product first shipped in 2024. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 5 catalogued features are instruction tuning, quantized model, and text generation.
Latest indexed changes and source events
google/gemma-4-12B-it-qat-w4a16-ct verified by the PulseGate indexer
Other apps tracked under the same category.