PulseGateIndexCategoriesUpdatesLive intelligence on AI-era software— indexed in the last hour
Coverage—in the index
Freshness—newest listing
Cadence—last week average · — today
Index9 markets86 categories
PulseGate

Live intelligence on the AI-era software market — apps, models, agents and infrastructure.

Most of it never reaches an official store. PulseGate maps the whole market — every category, worldwide — and measures how it moves: what’s launching, what’s gaining, what’s going quiet.

FollowGitHubX (Twitter)LinkedIn
Platform
All AppsFull IndexCategoriesIndustry UpdatesData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateG4GLM 4.7 Flash
Visit↗
Skip to content
  1. Index›
  2. Foundation models & chat›
  3. GLM 4.7 Flash
← Back to the index
G4

GLM 4.7 Flash

huggingface.co·Infrastructure

GLM-4.7-Flash-AWQ is an AWQ-quantized version of the GLM-4.7-Flash model optimized for fast and memory-efficient inference. It supports advanced features such as function calling and custom chat templates. The model is distributed on Hugging Face and can be run locally using Transformers or Docker.

Open SourceApache-2.0WebCLIAPI
GGLM 4.7 Flash preview
Visit huggingface.co↗
4.4kstars
468forks
4alternatives
3features
2025since

Overview

3 features

In the Foundation models & chat space, GLM 4.7 Flash takes a focused approach. It focuses on deploying large language models with significantly reduced memory footprint while preserving tool-use and conversational capabilities. GLM 4.7 Flash is an open-source project aimed at developers. The project is open source (Apache-2.0). GLM 4.7 Flash is available on the web, the command line, and API.

QuantTrio builds and maintains GLM 4.7 Flash, and it first shipped in 2025. The project is developed in the open on GitHub with 4.4k stars. Key capabilities include AWQ Quantization, Function Calling, and Chat Templates.

Summary written by a language model from the project’s public pages.

  • ✓AWQ Quantization
  • ✓Function Calling
  • ✓Chat Templates
Tags
awq-quantizedglm-4flash-modeltool-calling-llm
AI capabilities
Text
Inference: LocalWeights: Open

Built with & integrations

Hosting
AWS
AI providers
local_oss
Runs on
BrowserCLIAPI-only
Detected from
AWS
x-amz-cf-id header · x-amz-cf-pop header · via header
local_oss
bvllm in the HTML

Trust & compliance

License
Apache-2.0
Verified signals
✓HTTPS✓Privacy Policy✓Terms of Service✓Open Source✓Free tier
Legal
Privacy Policy →Terms of Service →

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed28 Jul · 10:15 UTC
    QuantTrio/GLM-4.7-Flash-AWQ verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about GLM 4.7 Flash

What does GLM 4.7 Flash do?
GLM 4.7 Flash focuses on deploying large language models with significantly reduced memory footprint while preserving tool-use and conversational capabilities. It is catalogued under Foundation models & chat on PulseGate.
Who should use GLM 4.7 Flash?
GLM 4.7 Flash is an open-source project built for developers.
Is GLM 4.7 Flash free?
Yes — GLM 4.7 Flash is open source under the Apache-2.0 license and free to use.
What platforms does GLM 4.7 Flash run on?
GLM 4.7 Flash runs on the web, the command line, and API.
Is GLM 4.7 Flash still active?
Unverified. GLM 4.7 Flash has not been re-checked since it entered the index, so there is no finding either way — and only a positive finding would say otherwise.
What projects are similar to GLM 4.7 Flash?
Similar projects tracked by PulseGate include GLM, GLM 4.7 Flash, and GLM 4.7 Flash.GLMGLM 4.7 FlashGLM 4.7 Flash
Who makes GLM 4.7 Flash?
GLM 4.7 Flash is developed by QuantTrio.
When did GLM 4.7 Flash launch?
GLM 4.7 Flash first shipped in 2025.

At a glance

Pricing
Open Source
Platforms
Cli · Web
Languages
English
License
Apache-2.0
Built for
developers
Model
Open source
Solves
Deploying large language models with significantly reduced memory footprint while preserving tool-use and conversational capabilities.

Registered as

Hugging Face
QuantTrio/GLM-4.7-Flash-AWQ

Developer

QuantTrio
Small team

Open source

Stars
4,406
Forks
468
Open issues
27
Last commit
1 Feb 2026
Commits 90d
0
Contributors
4
Authorship
Small team
Default branch
main

Live coverage

Identity confidence
Medium · 70.4
Indexed
28 Jul 2026
Lifecycle
Alive
Last seen
28 Jul 2026
Identity audit (13)
Slug
quanttrio-glm-4-7-flash-awq-huggingface-co
Lifecycle state recorded
28 Jul 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
28 Jul 2026
Timeline basis
Indexed-at chronology. This listing's first-seen date was written by the catalog backfill, not observed here, so it is not treated as a sighting.
Name from
Derived from the project's own page and URL.
Category from
Derived from the model's own Hugging Face metadata.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model from page content.
Last updated
5 Aug 2026
Canonical URL
https://huggingface.co/QuantTrio/GLM-4.7-Flash-AWQ

Ship this? Send a correction — no account, and you get a link to follow it.

Similar apps

Closest matches by what these projects do

  • GLGLMhuggingface.co
  • G4GLM 4.7 Flashhuggingface.co
  • G4GLM 4.7 Flashhuggingface.co
  • G4GLM 4.7 Flashhuggingface.co
  • G4GLM 4.7 Flashhuggingface.co
  • G4GLM 4.7 Flashhuggingface.co