This is a GGUF quantized version of the Gemma-4-31B instruct model published by Unsloth. It enables efficient CPU and GPU inference of a capable open-weight language model using tools like llama.cpp or Ollama. The model is designed for developers and researchers who want to run high-performance instruction-tuned LLMs locally with reduced memory requirements.
Gemma 4 31B It is a Foundation models & chat product. It focuses on running large language models efficiently on consumer hardware without high-end GPUs. Gemma 4 31B It is an open-source project aimed at developers. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
Behind Gemma 4 31B It is Unsloth, and the product first shipped in 2023. The project is developed in the open on GitHub with 68.7k stars and 1.2k commits in the last 90 days. It competes in a saturated segment with 25 similar apps in PulseGate's index. Among its 3 catalogued features are GGUF Quantization, Local Inference, and Instruct Model.
Latest indexed changes and source events
unsloth/gemma-4-31B-it-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.