unsloth/GLM-4.7-Flash-GGUF contains GGUF quantized files for the GLM-4.7 Flash model. It includes support for tool calling using XML-style function call formatting. The model is suitable for local deployment using llama.cpp or other GGUF-compatible runtimes and is provided to enable efficient on-premise or desktop AI applications.
GLM 4.7 Flash is a Tool calling project. It focuses on running the GLM-4.7 Flash model locally with GGUF quantization for efficient inference. GLM 4.7 Flash is an open-source project aimed at developers and researchers. The project is open source (Apache-2.0). It runs on the web, the command line, and API.
It is developed by Unsloth, and it first shipped in 2023. Development happens publicly on GitHub with 68.4k stars and 1.2k commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 5 similar projects. Key capabilities include GGUF Format, Tool Calling Support, and XML Function Calling.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do