gemma-4-12B-AWQ is an AWQ-quantized variant of Google's Gemma 4 12B parameter model. It enables efficient local inference while preserving most of the original model's instruction-following and reasoning capabilities. The model is published on Hugging Face and is intended for developers building applications that require strong language understanding without relying on cloud APIs.
Gemma 4 12B is a Foundation models & chat product. It focuses on running large language models efficiently on consumer or edge hardware using quantization. Gemma 4 12B is an open-source project aimed at developers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by Google (United States), and the product first shipped in 2026. The project is developed in the open on GitHub with 763 commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are Quantized Inference, Instruction Following, and Tool Use.
Latest indexed changes and source events
mattbucci/gemma-4-12B-AWQ verified by the PulseGate indexer
Other apps tracked under the same category.