unsloth/GLM-4.7-Flash-GGUF contains GGUF quantized files for the GLM-4.7 Flash model. It includes support for tool calling using XML-style function call formatting. The model is suitable for local deployment using llama.cpp or other GGUF-compatible runtimes and is provided to enable efficient on-premise or desktop AI applications.
GLM 4.7 Flash is a Foundation models & chat product. It focuses on running the GLM-4.7 Flash model locally with GGUF quantization for efficient inference. GLM 4.7 Flash is an open-source project aimed at developers and researchers. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
Behind GLM 4.7 Flash is Unsloth, and the product first shipped in 2023. The project is developed in the open on GitHub with 68.4k stars and 1.2k commits in the last 90 days. PulseGate's similarity index places it among 5 comparable tools. Among its 4 catalogued features are GGUF Format, Tool Calling Support, and XML Function Calling.
Latest indexed changes and source events
unsloth/GLM-4.7-Flash-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.