ZMIME
Z.ai (Zhipu) · Released Aug 14, 2026

GLM-5.3

Flagship GLM model for long-horizon coding, agents, and complex project delivery.

  • Open weights
  • Reasoning
  • Tool calling
Capability
155.8ECI · #24 of 148
Input
$1.40per 1M tokens
Output
$4.40per 1M tokens
Context
1M131K max output
Pricing

What it costs

Input
$1.40per million tokens
Output
$4.40per million tokens
Cached input
$0.26per million tokens
Blended (3:1)
$2.15pricier than 72% of priced models

Typical monthly bills

Monthly usageEstimated cost
Side project2M input + 0.5M output tokens$5.00
Team assistant25M input + 5M output tokens$57.00
Production app250M input + 50M output tokens$570.00

Official Z.AI API price, as listed on models.dev.

Capability

Independent benchmarks

155.8Epoch Capabilities Index
#24 of 148 scored models
88Median 146167

The shaded band is Epoch AI’s confidence range (153.7–158.3); the tick marks the median scored model.

  • GPQA DiamondGraduate-level science questions · max setting90.9%
  • FrontierMath Tiers 1–3Research-level mathematics · max setting68.8%
  • OTIS Mock AIME 2024–2025Competition mathematics · max setting91.1%
  • SimpleQA VerifiedShort factual questions · max setting41.0%

Scores from Epoch AI, run independently of Z.ai (Zhipu).

Specifications

The details

Lab
Z.ai (Zhipu)
Released
Aug 14, 2026
Context window
1,000,000 tokens
Max output
131,072 tokens
Inputs
Text
Output
Text
Reasoning
Adjustable effort low · high · max
Tool calling
Yes
Structured output
Yes
Weights
Open weights
API model ID
glm-5.3on Z.AI
Availability
62 API providerslisted on models.dev

Z.AI model documentation

Lab-reported

What Z.ai (Zhipu) claims

Published by the lab at launch. Settings vary, so compare these only with care.

BenchmarkScoreSettingSource
Terminal-Bench v2.188.2max effortSource
Terminal-Bench v3.028.3 avg@3max effortSource
DeepSWE v1.166.9 resolvedmax effort; 6h timeout; 400K contextSource
NL2Repo58max effortSource
ProgramBench19 almost solvedmax effortSource
FrontierSWE78.1max effortSource
SWE-Marathon v1.142.5max effort; modified anti-cheat checksSource
PostTrainBench39.8 weighted averagemax effort; 3 runs; modified external-API checksSource
CyberGym84.5max effortSource
ExploitGym105 tasks solvedmax effort; 2h rescaled timeoutSource
ExploitGym130 tasks solvedmax effort; 6h rescaled timeoutSource
ExploitBench54.4 average coveragemax effort; union over 3 revisionsSource
Same lab

More from Z.ai (Zhipu)

Questions

About GLM-5.3

How much does GLM-5.3 cost?

GLM-5.3 costs $1.40 input / $4.40 output per million tokens (official Z.AI API price). Cached input is $0.26 per million tokens. At a 3:1 input-to-output mix that is $2.15 per million tokens, more expensive than 72% of the 360 priced models we track.

What is the context window of GLM-5.3?

GLM-5.3 accepts up to 1,000,000 tokens per request and can write up to 131,072 tokens in one response.

How good is GLM-5.3?

Epoch AI gives GLM-5.3 a Capabilities Index score of 155.8 (likely range 153.7–158.3), ranking it #24 of 148 models Epoch has scored. Epoch AI benchmark results: GPQA Diamond 90.9%, FrontierMath Tiers 1–3 68.8%, OTIS Mock AIME 2024–2025 91.1%, SimpleQA Verified 41.0%.

Is GLM-5.3 open source?

Yes. Z.ai (Zhipu) publishes the weights, so it can be downloaded and self-hosted.

What inputs does GLM-5.3 support?

GLM-5.3 accepts text and replies in text. It is a reasoning model with low, high and max effort settings, supports tool calling and can return structured JSON output.

When was GLM-5.3 released?

Z.ai (Zhipu) released GLM-5.3 on Aug 14, 2026.

What are the best alternatives to GLM-5.3?

The closest current models from other labs on capability, price and release date are Muse Spark 1.2 (Meta, ECI 155.0, $1.25 / $4.25), Grok 4.6 (xAI, ECI 156.6, $2.00 / $6.00), Qwen3.8 Max 0902 (Alibaba (Qwen), ECI 155.2, $2.00 / $6.00) and Gemini 3.8 Flash (Google, ECI 156.9, $0.75 / $3.75).