This is a community-quantized (NVFP4) version of Google's Gemma 4 31B instruction-tuned (it) model hosted on Hugging Face. It enables efficient local or hosted inference of a powerful open LLM using libraries such as Transformers or vLLM. The model includes a chat template and tokenizer configuration optimized for conversational use.
Gemma 4 31B It NVFP4 Turbo sits in PulseGate's Text generation category. It focuses on running large language models efficiently on consumer or edge hardware without high compute costs. Gemma 4 31B It NVFP4 Turbo is an open-source project aimed at developers. Gemma 4 31B It NVFP4 Turbo is open source under the Open Source license. Gemma 4 31B It NVFP4 Turbo is available on the web, the command line, and API.
Behind Gemma 4 31B It NVFP4 Turbo is LilaRest, and it first shipped in 2025. It competes in a saturated segment with 25 similar projects in PulseGate's index. Among its 3 catalogued features are quantized weights, instruction tuned, and chat template.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do