A 4-bit AWQ quantized version of the Gemma 4 12B instruction-tuned model. It enables efficient local inference for chat and text generation tasks. Suitable for developers who want to run powerful LLMs on GPUs with limited VRAM.
Gemma 4 12B It is a Quantised & converted weights project. It focuses on running large language models efficiently on consumer hardware with reduced memory requirements. Gemma 4 12B It is an open-source project aimed at developers. Gemma 4 12B It is open source under the Open Source license. It ships for the web, the command line, and API.
It is developed by cyankiwi. It competes in a saturated segment with 25 similar projects in PulseGate's index. Key capabilities include Quantized Model, Text Generation, and AWQ INT4. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do