AI agents · OpenClaw · self-hosting · automation

Quick Answer

Best Open-Weight Coding Model 2026: Ranked

Published:

The Short Answer

In July 2026 the open-weight coding crown splits three ways:

  • Kimi K3 — highest ceiling. Leads the Arena.ai Frontend Code Arena, Artificial Analysis Index ~57. But $3/$15 API and 2.8T weights make it the priciest and hardest to self-host.
  • GLM-5.2 — best all-round. 744B, MIT license, tops the open-weights Intelligence Index, self-hostable today, $1.40/$4.40.
  • DeepSeek V4 — best value. Leads raw SWE-bench Verified and competitive coding; $0.435/$0.87 (Pro) or $0.14/$0.28 (Flash) off-peak.

Pick GLM-5.2 to self-host, DeepSeek V4 for cheapest API coding, Kimi K3 for the top quality ceiling.

The Ranking Table (July 2026)

Model$/MTok in$/MTok outLicenseSelf-hostCoding strength
DeepSeek V4 Pro$0.435*$0.87*Open weightsData-centerSWE-bench, competitive
DeepSeek V4 Flash$0.14*$0.28*Open weights24GB (quantized)Fast, cheap tasks
GLM-5.2$1.40$4.40MITModerateAgentic, long-context
Kimi K3$3.00$15.00Modified MITMulti-nodeFrontend Code Arena #1
Qwen 3.7 MaxvariesvariesOpen weightsData-centerStrong all-round

DeepSeek off-peak; 2x during peak hours (1-4 & 6-10 UTC). V4 Pro cache-hit ~$0.043.

Best by Use Case

Best value coding → DeepSeek V4 Pro / Flash. Nothing beats $0.435/$0.87 (Pro) for cost-per-task on real coding work, and V4 leads raw SWE-bench Verified plus LiveCodeBench (~93.5%) and Codeforces (~3206). Use Flash ($0.14/$0.28) for high-volume, cheap tasks and Pro for hard reasoning. Watch peak-hour surge and note the July 24 retirement of the deepseek-chat/deepseek-reasoner names.

Best to self-host → GLM-5.2. The 744B MoE with a true MIT license is the only frontier-class open model you can realistically run, fine-tune, and ship commercially without license friction. Tops the open-weights Intelligence Index and serves faster than K3 or V4 Pro. Output cost ($4.40) is ~1/3 of Kimi K3’s.

Best quality ceiling → Kimi K3. Frontend Code Arena leader (ahead of Claude Fable 5), Index ~57. Full weights drop July 27 under Modified MIT, but at 2.8T you need a ~16x H100 cluster to self-host and $3/$15 to use the API. Choose it when quality outranks cost.

Strong all-rounder → Qwen 3.7 Max. Alibaba’s flagship remains a capable open coding model and a common Claude Code drop-in, though it trails the top three on the most-cited July 2026 coding benchmarks.

How to Choose

  • Cheapest API coding → DeepSeek V4 Pro (or Flash for volume)
  • Self-host on your own hardware → GLM-5.2 (MIT)
  • Highest open-weight quality → Kimi K3
  • Commercial license clarity → GLM-5.2 (MIT) or DeepSeek V4
  • Single 24GB card → DeepSeek V4 Flash (quantized)

Sources