AI agents · OpenClaw · self-hosting · automation

Quick Answer

Qwen3.8-Max vs Kimi K3 vs GLM-5.2 (Aug 2026)

Published:

The Short Answer

For open-weight agentic coding in August 2026: Kimi K3 is the proven, deployable leader (weights since Jul 27); GLM-5.2 is the best mid-VRAM long-horizon coder; Qwen3.8-Max claims the top agentic benchmarks but is unverified and GPU-hungry (weights ~Aug 10). Deploy Kimi K3 / GLM-5.2 now, watch Qwen3.8-Max.

Side-by-Side

Qwen3.8-MaxKimi K3GLM-5.2
VendorAlibabaMoonshotZhipu
Open weights~week of Aug 10Yes (Jul 27)Yes
Params2.4T (MoE)Large MoEMid MoE
Terminal-Bench 2.1*86.6
SWE-bench Pro*67.7
API price (in/out)TBD$3 / $15Low
Self-host difficultyVery highModerateApproachable

*Qwen3.8-Max figures are Alibaba-reported, unverified independently as of Aug 5, 2026.

How To Choose

  • Ship an open-weight agent today → Kimi K3. Weights are public and mature; $3/$15 API if you don’t self-host. Best proven agentic coder in this group.
  • Self-host on modest hardware → GLM-5.2. Strong long-horizon coding without data-center GPUs.
  • Chase top claimed benchmarks → Qwen3.8-Max, once weights ship and independent SWE-bench/OSWorld numbers confirm Alibaba’s claims — and if you have the GPUs for 2.4T.

The Trade-Off

  • Proven vs claimed: Kimi K3 and GLM-5.2 are deployed and tested; Qwen3.8-Max’s lead is vendor-reported.
  • Capability vs runnability: Qwen3.8-Max aims highest but is the hardest to serve; GLM-5.2 is the pragmatic pick.
  • API vs self-host: Kimi K3’s $3/$15 API is a fine hedge until you commit GPUs.

Watch Outs

  • Qwen3.8-Max weights + license unconfirmed at launch (verify before planning self-host).
  • Benchmark tables lagged the Qwen3.8-Max launch — some are internal evals.

Verdict

  • Best proven open agentic coder → Kimi K3
  • Best mid-VRAM self-host → GLM-5.2
  • Top claimed / watch → Qwen3.8-Max

Sources