AI agents · OpenClaw · self-hosting · automation

Quick Answer

GLM-5.2 vs Kimi K3: Open-Weight King After K3

Published:

The Short Answer

Kimi K3’s weights drop July 27, 2026 amid claims it’s the new open-weight king. The honest answer: it depends what “king” means.

  • Kimi K3 — the capability king. Artificial Analysis Index ~57, #1 on the Arena.ai Frontend Code Arena (ahead of Claude Fable 5). But 2.8T params, $3/$15 API, “Modified MIT” license.
  • GLM-5.2 — the deployment king. 744B, clean MIT license, self-hostable today, faster serving, $1.40/$4.40 (~1/3 of K3’s output cost).

If you’re paying API rates and want the top quality → Kimi K3. If you want open weights you can actually run, license, and afford → GLM-5.2.

The Comparison Table (July 2026)

GLM-5.2Kimi K3
DeveloperZhipu / Z.aiMoonshot AI
Total params744B2.8T
Context1M1M
LicenseMITModified MIT
WeightsAvailable nowJuly 27, 2026
$/MTok input$1.40$3.00
$/MTok output$4.40$15.00
AA Intelligence Index~51~57
Frontend Code ArenaStrong#1 (ahead of Fable 5)
Self-host difficultyModerateVery high (multi-node)
Serving speedFasterSlower

Where Each Wins

Kimi K3 — the ceiling. The Artificial Analysis Index ~57 and Frontend Code Arena #1 make it the strongest open model on raw capability in July 2026, beating Claude Fable 5 on frontend code. It’s the pick when quality is the only thing that matters and you’re paying API rates anyway. The costs: $3/$15 (5-17x GLM-5.2’s output on some comparisons) and a 2.8T weight footprint that needs ~16x H100 to self-host at frontier throughput.

GLM-5.2 — the practical king. 744B with a true MIT license makes it the model you can self-host, fine-tune, and ship commercially with zero license friction. It tops the open-weights Intelligence Index, serves faster than K3, and costs ~1/3 as much on output. For any team that actually intends to run open weights rather than call an API, GLM-5.2 is the default.

Does K3 Dethrone GLM-5.2?

On benchmarks and frontend coding — yes, K3 takes the capability lead. On the metric most self-hosters care about — cost-to-deployno, GLM-5.2 stays king. K3 is 5-17x the output price, harder to serve, and its “Modified MIT” terms need reading before commercial use. GLM-5.2 remains the sensible default; K3 is the reach model.

How to Choose

  • Want the highest open-weight quality, API budget → Kimi K3
  • Want to self-host on realistic hardware → GLM-5.2
  • Need clean commercial license → GLM-5.2 (MIT)
  • Frontend/agentic coding at the top end → Kimi K3
  • Cost-sensitive, high volume → GLM-5.2 (or DeepSeek V4 Pro for even cheaper)

Sources