GLM-4.6V-Flash is an open-weights multimodal foundation model from zai-org that processes both text and images. It supports advanced features such as tool calling and is distributed on Hugging Face for local or self-hosted inference. The model is intended for developers building vision-language applications or integrating multimodal capabilities into their own systems.
GLM 4.6V Flash sits in PulseGate's Multimodal & vision category. It focuses on accessing and running a capable open-source vision-language model locally or via self-hosted inference. GLM 4.6V Flash is an open-source project aimed at developers. GLM 4.6V Flash is open source under the Apache-2.0 license. GLM 4.6V Flash is available on the web, the command line, and API, and it can be self-hosted.
zai-org builds and maintains GLM 4.6V Flash, and it first shipped in 2025. The project is developed in the open on GitHub with 2.4k stars and 2 commits in the last 90 days. Key capabilities include Multimodal Input, Tool Calling, and Chat Template.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do