GLM-5.2-Int4-Int8Mix is a quantized variant of the GLM-5.2 model using a mix of 4-bit and 8-bit precision. It includes support for reasoning effort levels, tool calling, and advanced system prompting. The model is designed for efficient local or server-based inference while maintaining strong language understanding and generation capabilities.
In the Foundation models & chat space, GLM 5.2 Int4 Int8Mix takes a focused approach. It focuses on deploying large GLM models with balanced memory usage and performance through selective quantization. It is built as an open-source project for developers. GLM 5.2 Int4 Int8Mix is open source under the Apache-2.0 license. It runs on the web, the command line, and API.
QuantTrio builds and maintains GLM 5.2 Int4 Int8Mix, and it first shipped in 2026. The project is developed in the open on GitHub with 6.7k stars and 12 commits in the last 90 days.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do