This is a GGUF quantized version of the Gemma-4-31B instruct model published by Unsloth. It enables efficient CPU and GPU inference of a capable open-weight language model using tools like llama.cpp or Ollama. The model is designed for developers and researchers who want to run high-performance instruction-tuned LLMs locally with reduced memory requirements.
Gemma 4 31B It is a Text generation project. It focuses on running large language models efficiently on consumer hardware without high-end GPUs. It is built as an open-source project for developers. Gemma 4 31B It is open source under the Apache-2.0 license. It ships for the web, the command line, and API.
It is developed by Unsloth, and it first shipped in 2023. The project is developed in the open on GitHub with 68.7k stars and 1.2k commits in the last 90 days. The category is crowded — PulseGate's index counts 25 comparable projects. Among its 3 catalogued features are GGUF Quantization, Local Inference, and Instruct Model.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do