Z.ai

Tier B · Specialist or second-choice models.

GLM 5.3

A B-tier builder with a complete visual benchmark suite that still benefits from close review.

API pricing

USD per million tokens · Z.ai API
Input
$1.4
Output
$4.4
Cached input
$0.26
Context window
1M
Pricing details and specifications

Try the active visual tests

Open the generated scene to test it, or compare that prompt with another model.

10 runs

City Scroll Journey

Run date: 2026-08-17

Specialist test · evaluated separately

Cinematic scroll journey testing AI-generated scene continuity, scroll-scrubbed camera motion, and art-directed landing craft.

Low-Poly Tower Defense

Run date: 2026-08-17

Diorama tower defense testing economy balance, wave design, placement rules, and combat readability.

Mechanical Watch Simulator

Run date: 2026-08-17

Interactive watch movement testing mechanical legibility, accurate relative motion, and real-time 3D controls.

Stormwind Trebuchet Simulator

Run date: 2026-08-17

Counterweight siege simulation testing coupled mechanics, trajectory prediction, projectile cameras, and interactive tuning.

Yingzao Fashi Assembly

Run date: 2026-08-17

Specialist test · evaluated separately

Timber assembly scene testing structure, joinery, construction order, and material clarity.

Retired tests · 6 archived runs

These tests are no longer in the active suite. Their original results remain available for inspection.

Editorial verdict

Where this model fits

Back to all benchmarks →

GLM 5.3 completed all 16 Superbash visual benchmarks in one local session; every artifact validates and runs. The first attempt stalled when the account hit its usage ceiling, but a lower-concurrency rerun completed the full set. B tier reflects its useful build coverage alongside the team's cost and access concerns; validation alone is not a quality score.

Best for

  • Complex software engineering
  • Long-horizon agent tasks
  • Coding Plan workflows

Watch out

  • Game balance and gold times are tuned analytically, not by long playtesting
  • Still expensive in the team's initial use
  • General API access is still coming soon

Why it is ranked here

  1. Z.ai documents GLM 5.3 with a one-million-token context window, 128K maximum output, and multiple reasoning modes.
  2. The completed 16-run visual set covers the full prompt range, from office simulations to a coupled trebuchet model, with honest manifest warnings where things were simplified.
  3. The team saw Kimi-level potential on day one, but the credit-limited experience and unscored output quality keep this a reviewed B-tier choice.

Evidence

2026-09-23

Superbash editorial model ranking

Takeaway: GLM 5.3 is currently placed in Tier B.

The September 2026 editorial roster places GLM 5.3 at rank 12.

Open source →
2026-08-14

Introducing GLM 5.3

Takeaway: Use GLM 5.3 for cost-conscious coding work after reviewing the output.

The model is currently available through Z.ai's Coding Plan; verify general API availability before building a direct integration.

Open source →

How GLM 5.3 compares

2 peers with published comparison scores.

  • GLM 5.3 This model
  • Claude Mythos 5
  • GPT-5.6 Sol

1 benchmarks · 3 models

CyberGym

Cybersecurity · Higher is better

Source
GLM 5.384.5%
Claude Mythos 583.8%
GPT-5.6 Sol83.6%

Official benchmark profile

GLM-5.3 coding, agent, and cybersecurity results.

Z.ai sourceAugust 2026Source report →
Agentic coding

Terminal-Bench 3.0

28.3%
Coding

DeepSWE v1.1

66.9%
Terminal agents

Agents’ Last Exam (CLI)

28.5%
Full official benchmark table6 scores and peer comparisons
BenchmarkAreaScoreComparison
Terminal-Bench 3.0Agentic coding28.3%
DeepSWE v1.1Coding66.9%
Agents’ Last Exam (CLI)Terminal agents28.5%
GDPval-AA v2Professional work1769 Elo
CyberGymCybersecurity84.5%
GLM 5.384.5%
Claude Mythos 583.8%
GPT-5.6 Sol83.6%
ExploitBenchCybersecurity54.4%
Sources, pricing details and methodology

Verified model facts

Model specifications

Pricing source · checked 2026-09-08 ↗
Status
Preview / limited access
API model ID
Not publicly verified
Context
1M
Max output
128K
License
Not publicly verified
Access route
Not publicly verified
Modalities
Not publicly verified
Hardware
Not publicly verified
Z.ai API · USD / 1M tokens
$1.4 input · $0.26 cached input · $4.4 output

Cached-input storage is currently free for a limited time.

Benchmark sources

Z.ai reports large gains over GLM-5.2 across coding, terminal-agent, and cybersecurity evaluations. These are vendor-reported launch results; the general GLM-5.3 API is still marked as coming soon.

Introducing GLM-5.3

GLM-5.3 is available to GLM Coding Plan users. Z.ai says direct API access is coming soon.

Evaluation settings

  • Z.ai launch result; GLM-5.2 scored 4.6% in the same comparison.
  • Z.ai launch result; GLM-5.2 scored 46.2% in the same comparison.
  • Z.ai launch result; GLM-5.2 scored 23.8% in the same comparison.
  • Z.ai-reported Artificial Analysis occupational-work result.
  • Z.ai launch result.
  • Z.ai launch result; GLM-5.2 scored 24.4% in the same comparison.