AI agents · OpenClaw · self-hosting · automation

Quick Answer

What Is GLM-5.3? Z.ai's New Coding Model (Aug 2026)

Published:

The Short Answer

GLM-5.3 is Z.ai’s coding-focused model released August 14, 2026. It is not a new architecture and not a new pretrain — it runs on the same ~743B-parameter base as GLM-5.2, and every reported gain comes from scaled post-training: more task environments, more environment types, longer training runs. The headline result is long-horizon coding: Terminal-Bench 3.0 jumps from 4.6 to 28.3.

Key Facts

GLM-5.3
ReleasedAugust 14, 2026
Base modelSame ~743B base as GLM-5.2 (post-training only)
API price$1.40 in / $4.40 out per MTok
Cached input$0.26 per MTok
Coding PlanFrom $18/month
Off-peak discount50% of standard points outside 14:00-18:00 UTC+8 weekdays
Open weightsStaged — not available at launch
ThinkingMandatory; three effort levels (low, high, max)

What Actually Improved

The gains concentrate on agentic, long-horizon work rather than one-shot Q&A:

  • Terminal-Bench 3.0: 4.6 → 28.3 — the largest single jump, on the hardest terminal-agent benchmark generation.
  • Cyber and automation: Z.ai reports GLM-5.3 leading CyberGym and AutomationBench among open-line models — enough of a step that its emergent cyber capability drew security commentary on launch week.
  • Complex coding: improvements on DeepSWE-style multi-file tasks, per Z.ai’s launch materials.

The lesson labs keep re-learning in 2026: post-training scale is a product lever on its own. Z.ai shipped a meaningfully better coding agent without touching the pretrain.

What Changed Operationally

Two changes matter for anyone deploying it:

  1. Thinking is now mandatory. You pick an effort level (low/high/max) but cannot disable reasoning — budget output tokens accordingly.
  2. Weights are staged, not immediate. GLM built its reputation on day-one open weights; GLM-5.3 launched API-first with weights “to follow.” If your stack depends on self-hosting, GLM-5.2 remains the current self-hostable GLM (see best open-weight coding model 2026).

How It Stacks Up

At $1.40/$4.40 per MTok, GLM-5.3 undercuts Western frontier coding rates (Claude Opus 4.8/Opus 5 at $5/$25; GPT-5.6 Sol at $5/$30) while claiming competitive agentic-coding scores. Against open-weight rivals it competes with Kimi K3 and DeepSeek V4 — with the caveat that, for now, only its rivals give you the weights.

Who Should Care

  • Coding-agent builders wanting near-frontier long-horizon performance at ~3-6× less than Western API rates.
  • GLM Coding Plan users — the $18/month plan now fronts a materially better model.
  • Self-hosters — wait for the staged weights before counting GLM-5.3 as an open model.

Last verified: August 15, 2026.

Sources