ZMIME
Z.ai (Zhipu) · Released Jan 19, 2026

GLM-4.7-Flash

Budget GLM lane for fast coding help, routing, and everyday automation.

  • Open weights
  • Reasoning
  • Tool calling
Capability
—No ECI score yet
Input
$0.06per 1M tokens
Output
$0.40per 1M tokens
Context
200K131K max output
Pricing

What it costs

Input
$0.06per million tokens
Output
$0.40per million tokens
Blended (3:1)
$0.145cheaper than 87% of priced models

Typical monthly bills

Monthly usageEstimated cost
Side project2M input + 0.5M output tokens$0.321
Team assistant25M input + 5M output tokens$3.51
Production app250M input + 50M output tokens$35.13

Median of 13 paid API providers on models.dev; Z.ai (Zhipu) also offers it free on Z.AI.

Capability

Independent benchmarks

Epoch AI has not published a Capabilities Index score for GLM-4.7-Flash yet.

  • GPQA DiamondGraduate-level science questions45.1%
  • OTIS Mock AIME 2024–2025Competition mathematics25.0%

Scores from Epoch AI, run independently of Z.ai (Zhipu).

Specifications

The details

Lab
Z.ai (Zhipu)
Released
Jan 19, 2026
Knowledge cutoff
Apr 2025
Context window
200,000 tokens
Max output
131,072 tokens
Inputs
Text
Output
Text
Reasoning
Can be switched on or off
Tool calling
Yes
Structured output
No
API model ID
glm-4.7-flashon Z.AI
Availability
19 API providerslisted on models.dev

Z.AI model documentation

Lab-reported

What Z.ai (Zhipu) claims

Published by the lab at launch. Settings vary, so compare these only with care.

BenchmarkScoreSettingSource
SWE-Bench Verified59.2 resolved—Source
Same lab

More from Z.ai (Zhipu)

Questions

About GLM-4.7-Flash

How much does GLM-4.7-Flash cost?

GLM-4.7-Flash costs $0.06 input / $0.40 output per million tokens (median across 13 API providers; free on Z.AI). At a 3:1 input-to-output mix that is $0.145 per million tokens, cheaper than 87% of the 360 priced models we track.

What is the context window of GLM-4.7-Flash?

GLM-4.7-Flash accepts up to 200,000 tokens per request and can write up to 131,072 tokens in one response.

How good is GLM-4.7-Flash?

Epoch AI has not published a Capabilities Index score for GLM-4.7-Flash yet. Epoch AI benchmark results: GPQA Diamond 45.1%, OTIS Mock AIME 2024–2025 25.0%.

Is GLM-4.7-Flash open source?

Yes. Z.ai (Zhipu) publishes the weights, so it can be downloaded and self-hosted.

What inputs does GLM-4.7-Flash support?

GLM-4.7-Flash accepts text and replies in text. It is a reasoning model, supports tool calling.

When was GLM-4.7-Flash released?

Z.ai (Zhipu) released GLM-4.7-Flash on Jan 19, 2026. Its training data runs to Apr 2025.

What are the best alternatives to GLM-4.7-Flash?

The closest current models from other labs on capability, price and release date are Mistral Small 3.2 (Mistral AI, ECI 131.7, $0.10 / $0.30), Gemma 3 27B IT (Google, ECI 130.0, $0.08 / $0.20), Qwen Turbo (Alibaba (Qwen), $0.05 / $0.20) and Step 3.5 Flash (StepFun, $0.10 / $0.30).