An NVIDIA-optimized version of the GLM-5 large language model using NVFP4 precision for accelerated inference on compatible GPUs. It includes support for tool calling and follows standard chat templates, making it suitable for deployment in performance-sensitive local or on-premise environments.
GLM 5 sits in PulseGate's Foundation models & chat category. It focuses on running large GLM models efficiently on NVIDIA hardware with optimized precision formats. GLM 5 is an open-source project aimed at developers. GLM 5 is open source under the Apache-2.0 license. It ships for the web and the command line, and it can be self-hosted.
Behind GLM 5 is nvidia, and it first shipped in 2024. Development happens publicly on GitHub with 3.3k stars and 354 commits in the last 90 days.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do