This repository contains a 4-bit quantized GGUF version of Google's Gemma-4 26B model (A4B-IT variant). It is optimized for local inference using tools such as llama.cpp. The model supports instruction following and is suitable for developers who want to run a capable LLM offline with reduced memory requirements.
Gemma 4 26B A4B It Qat Q4 0 sits in PulseGate's Foundation models & chat category. It focuses on running a large instruction-tuned Gemma model efficiently on consumer hardware using quantization. Gemma 4 26B A4B It Qat Q4 0 is an open-source project aimed at AI developers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by Google, and the product first shipped in 2025. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are quantized model, GGUF format, and instruction tuned.
Latest indexed changes and source events
google/gemma-4-26B-A4B-it-qat-q4_0-gguf verified by the PulseGate indexer