This repository hosts a 4-bit AWQ quantized variant of the Gemma-4-31B-it model. The W4A16 quantization scheme enables efficient inference on consumer GPUs while preserving most of the model's capabilities.
Gemma 4 31B It 4bit W4A16 sits in PulseGate's Foundation models & chat category. It focuses on running the large Gemma 4 31B model on hardware with limited VRAM using 4-bit AWQ quantization. Gemma 4 31B It 4bit W4A16 is an open-source project aimed at developers. Gemma 4 31B It 4bit W4A16 is open source under the Apache-2.0 license. It ships for the web and API.
ebircak builds and maintains Gemma 4 31B It 4bit W4A16, and it first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 168 commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are 4-bit Quantization, AWQ Format, and Instruction Tuning.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do