GLM-5.2-Int4-Int8Mix is a quantized variant of the GLM-5.2 model using a mix of 4-bit and 8-bit precision. It includes support for reasoning effort levels, tool calling, and advanced system prompting. The model is designed for efficient local or server-based inference while maintaining strong language understanding and generation capabilities.
GLM 5.2 Int4 Int8Mix sits in PulseGate's Foundation models & chat category. It focuses on deploying large GLM models with balanced memory usage and performance through selective quantization. It is built as an open-source project for developers. GLM 5.2 Int4 Int8Mix is open source under the Apache-2.0 license. It runs on the web, the command line, and API.
QuantTrio builds and maintains GLM 5.2 Int4 Int8Mix, and the product first shipped in 2026. Development happens publicly on GitHub with 6.7k stars and 12 commits in the last 90 days.
Latest indexed changes and source events
QuantTrio/GLM-5.2-Int4-Int8Mix verified by the PulseGate indexer
Other apps tracked under the same category.