GLM-5.2-NVFP4 is an NVIDIA-published, quantized version of the GLM-5.2 model optimized for NVFP4 precision. It includes built-in support for tool calling, function execution, and configurable reasoning effort. The model is distributed on Hugging Face for use with Transformers and compatible inference engines.
In the Foundation models & chat space, GLM takes a focused approach. It focuses on deploying a high-performance, quantized language model with native tool use and reasoning on NVIDIA hardware. It is built as an open-source project for developers. GLM is open source under the Apache-2.0 license. The product ships for the web, the command line, and API.
NVIDIA builds and maintains GLM, and the product first shipped in 2024. Development happens publicly on GitHub with 3.3k stars and 350 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 5 similar tools. Key capabilities include Tool Calling, Reasoning Engine, and Function Calling.
Latest indexed changes and source events
nvidia/GLM-5.2-NVFP4 verified by the PulseGate indexer
Other apps tracked under the same category.