GLM-4.7-Flash
Budget GLM lane for fast coding help, routing, and everyday automation.
- Capability
- —No ECI score yet
- Input
- $0.06per 1M tokens
- Output
- $0.40per 1M tokens
- Context
- 200K131K max output
What it costs
- Input
- $0.06per million tokens
- Output
- $0.40per million tokens
- Blended (3:1)
- $0.145cheaper than 87% of priced models
Typical monthly bills
| Monthly usage | Estimated cost |
|---|---|
| Side project2M input + 0.5M output tokens | $0.321 |
| Team assistant25M input + 5M output tokens | $3.51 |
| Production app250M input + 50M output tokens | $35.13 |
Median of 13 paid API providers on models.dev; Z.ai (Zhipu) also offers it free on Z.AI.
Independent benchmarks
Epoch AI has not published a Capabilities Index score for GLM-4.7-Flash yet.
Scores from Epoch AI, run independently of Z.ai (Zhipu).
The details
- Lab
- Z.ai (Zhipu)
- Released
- Jan 19, 2026
- Knowledge cutoff
- Apr 2025
- Context window
- 200,000 tokens
- Max output
- 131,072 tokens
- Inputs
- Text
- Output
- Text
- Reasoning
- Can be switched on or off
- Tool calling
- Yes
- Structured output
- No
- Weights
- Open weights
- API model ID
glm-4.7-flashon Z.AI- Availability
- 19 API providerslisted on models.dev
What Z.ai (Zhipu) claims
Published by the lab at launch. Settings vary, so compare these only with care.
| Benchmark | Score | Setting | Source |
|---|---|---|---|
| SWE-Bench Verified | 59.2 resolved | — | Source |
More from Z.ai (Zhipu)
About GLM-4.7-Flash
How much does GLM-4.7-Flash cost?
GLM-4.7-Flash costs $0.06 input / $0.40 output per million tokens (median across 13 API providers; free on Z.AI). At a 3:1 input-to-output mix that is $0.145 per million tokens, cheaper than 87% of the 360 priced models we track.
What is the context window of GLM-4.7-Flash?
GLM-4.7-Flash accepts up to 200,000 tokens per request and can write up to 131,072 tokens in one response.
How good is GLM-4.7-Flash?
Epoch AI has not published a Capabilities Index score for GLM-4.7-Flash yet. Epoch AI benchmark results: GPQA Diamond 45.1%, OTIS Mock AIME 2024–2025 25.0%.
Is GLM-4.7-Flash open source?
Yes. Z.ai (Zhipu) publishes the weights, so it can be downloaded and self-hosted.
What inputs does GLM-4.7-Flash support?
GLM-4.7-Flash accepts text and replies in text. It is a reasoning model, supports tool calling.
When was GLM-4.7-Flash released?
Z.ai (Zhipu) released GLM-4.7-Flash on Jan 19, 2026. Its training data runs to Apr 2025.
What are the best alternatives to GLM-4.7-Flash?
The closest current models from other labs on capability, price and release date are Mistral Small 3.2 (Mistral AI, ECI 131.7, $0.10 / $0.30), Gemma 3 27B IT (Google, ECI 130.0, $0.08 / $0.20), Qwen Turbo (Alibaba (Qwen), $0.05 / $0.20) and Step 3.5 Flash (StepFun, $0.10 / $0.30).