QuantTrio/gemma-4-31B-it-AWQ provides an Activation-aware Weight Quantized (AWQ) version of Google's Gemma 4 31B instruction-tuned model. The quantization enables efficient inference on consumer or enterprise hardware with lower VRAM requirements. It includes a complete chat template and tokenizer configuration for seamless integration with existing LLM serving frameworks.
Gemma 4 31B It is a Text generation project. It focuses on deploying the large Gemma 4 31B model with significantly reduced memory footprint while retaining instruction-following performance. Gemma 4 31B It is an open-source project aimed at AI developers and researchers. Gemma 4 31B It is open source under the Open Source license. Gemma 4 31B It is available on the web, the command line, and API.
It is developed by QuantTrio, and it first shipped in 2025. It competes in a saturated segment with 22 similar projects in PulseGate's index. Key capabilities include Instruction Tuned, AWQ Quantization, and Chat Template.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do