This is a community-quantized version of Google's Gemma 4 31B instruction-tuned (it) model using AWQ 4-bit quantization. It is hosted on Hugging Face and can be used with popular transformer libraries via pip or Docker. The model is optimized for lower memory usage while maintaining performance for text generation and chat applications. It is intended for developers and researchers who want to run a capable large language model locally or on modest hardware.
In the Foundation models & chat space, Gemma 4 31B It takes a focused approach. It focuses on running large instruction-tuned language models locally with reduced memory requirements. Gemma 4 31B It is an open-source project aimed at developers. Gemma 4 31B It is open source under the Open Source license. It runs on the web, the command line, and API.
cyankiwi builds and maintains Gemma 4 31B It, and it first shipped in 2025. It competes in a saturated segment with 25 similar apps in PulseGate's index. Key capabilities include quantized weights, instruction tuned, and text generation.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Same category — not a similarity match