This is an FP8 quantized release of GLM-5.1 with support for advanced tool calling and function calling via a specialized chat template. The model uses a custom prompting format for tool use and is designed for efficient inference while maintaining strong performance on reasoning and agentic tasks.
GLM sits in PulseGate's Foundation models & chat category. It focuses on running large GLM models with reduced precision for faster inference on compatible hardware. It is built as an open-source project for developers. GLM is open source under the Apache-2.0 license. GLM is available on the web, the command line, and API.
Behind GLM is zai-org, and the product first shipped in 2026. Development happens publicly on GitHub with 6.8k stars and 12 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 6 similar tools. Key capabilities include Tool Calling, Function Calling, and Quantized Inference.
Latest indexed changes and source events
zai-org/GLM-5.1-FP8 verified by the PulseGate indexer
Other apps tracked under the same category.