gemma-4-31B-it-FP8-block is an FP8-quantized variant of the Gemma 4 31B language model, designed for efficient AI inference and deployment. It supports text generation and instruction tuning, making it suitable for AI engineers and researchers seeking resource-efficient LLMs.
Gemma 4 31B It FP8 Block is a Foundation models & chat project. It focuses on reducing computational costs for deploying large language models in production environments. Gemma 4 31B It FP8 Block is an open-source project aimed at AI engineers and researchers. The project is open source (Apache-2.0). It ships for the web, the command line, and API.
It is developed by RedHatAI, and it first shipped in 2019. Development happens publicly on GitHub with 3.6k stars and 171 commits in the last 90 days. Among its 5 catalogued features are FP8 quantization, efficient inference, and text generation.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do