gemma-4-31B-it-FP8-block is an FP8-quantized variant of the Gemma 4 31B language model, designed for efficient AI inference and deployment. It supports text generation and instruction tuning, making it suitable for AI engineers and researchers seeking resource-efficient LLMs.
In the Foundation models & chat space, Gemma 4 31B It FP8 Block takes a focused approach. It focuses on reducing computational costs for deploying large language models in production environments. Gemma 4 31B It FP8 Block is an open-source project aimed at AI engineers and researchers. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
It is developed by RedHatAI, and the product first shipped in 2019. The project is developed in the open on GitHub with 3.6k stars and 171 commits in the last 90 days. Among its 5 catalogued features are FP8 quantization, efficient inference, and text generation.
Latest indexed changes and source events
RedHatAI/gemma-4-31B-it-FP8-block verified by the PulseGate indexer
Other apps tracked under the same category.