This is an FP8-dynamic quantized version of the 31-billion parameter Gemma 4 instruction-tuned (it) model, published by RedHatAI on Hugging Face. It enables efficient inference of a large language model while preserving most of the original performance. The model includes a chat template and is designed for developers seeking optimized open-weight LLMs.
Gemma 4 31B It FP8 Dynamic sits in PulseGate's Text generation category. It focuses on running large instruction-tuned language models efficiently on hardware with reduced memory and compute requirements. Gemma 4 31B It FP8 Dynamic is an open-source project aimed at developers. Gemma 4 31B It FP8 Dynamic is open source under the Apache-2.0 license. It ships for the web, the command line, and API.
It is developed by RedHatAI, and it first shipped in 2019. Development happens publicly on GitHub with 3.6k stars and 171 commits in the last 90 days. Key capabilities include instruction tuned, FP8 quantization, and dynamic quantization.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do