QuantTrio/gemma-4-31B-it-AWQ provides an Activation-aware Weight Quantized (AWQ) version of Google's Gemma 4 31B instruction-tuned model. The quantization enables efficient inference on consumer or enterprise hardware with lower VRAM requirements. It includes a complete chat template and tokenizer configuration for seamless integration with existing LLM serving frameworks.
Gemma 4 31B It sits in PulseGate's Foundation models & chat category. It focuses on deploying the large Gemma 4 31B model with significantly reduced memory footprint while retaining instruction-following performance. Gemma 4 31B It is an open-source project aimed at AI developers and researchers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by QuantTrio, and the product first shipped in 2025. It competes in a saturated segment with 22 similar apps in PulseGate's index. Among its 3 catalogued features are Instruction Tuned, AWQ Quantization, and Chat Template.
Latest indexed changes and source events
QuantTrio/gemma-4-31B-it-AWQ verified by the PulseGate indexer
Other apps tracked under the same category.