GLM-4.6V-Flash is an open-weights multimodal foundation model from zai-org that processes both text and images. It supports advanced features such as tool calling and is distributed on Hugging Face for local or self-hosted inference. The model is intended for developers building vision-language applications or integrating multimodal capabilities into their own systems.
GLM 4.6V Flash sits in PulseGate's Foundation models & chat category. It focuses on accessing and running a capable open-source vision-language model locally or via self-hosted inference. GLM 4.6V Flash is an open-source project aimed at developers. The project is open source (Apache-2.0). The product ships for the web, the command line, and API, and it can be self-hosted.
zai-org builds and maintains GLM 4.6V Flash, and the product first shipped in 2025. The project is developed in the open on GitHub with 2.4k stars and 2 commits in the last 90 days. Among its 4 catalogued features are Multimodal Input, Tool Calling, and Chat Template.
Latest indexed changes and source events
zai-org/GLM-4.6V-Flash verified by the PulseGate indexer
Other apps tracked under the same category.