AI agents · OpenClaw · self-hosting · automation

Quick Answer

Qwen3.8-Max vs Kimi K3 vs DeepSeek V4 Pro (Aug 2026)

Published:

The Short Answer

Three open-frontier contenders as of August 2026:

  • Qwen3.8-Max (Alibaba, Aug 3) — biggest and newest, top benchmark claims, weights signaled.
  • Kimi K3 (Moonshot, open weights July 27) — the proven, deployable leader at $3/$15.
  • DeepSeek V4 Pro — the cost king at $0.435/$0.87 off-peak.

Side-by-Side

Qwen3.8-MaxKimi K3DeepSeek V4 Pro
LaunchedAug 3, 2026Open weights Jul 27, 2026GA Jul 24, 2026
Params2.4T (MoE)Large MoEMoE
Context1MLargeLong
API price (in/out per MTok)TBD$3 / $15 (flat)$0.435 / $0.87 (off-peak, 2x peak)
Cache-hit inputTBD~$0.043
Open weightsSignaled✅ Yes✅ Yes
30K/5K task costTBD~$0.16~$0.017

How To Choose

  • Want the safest production deploy today → Kimi K3. Open weights are out, pricing is flat (no peak/valley surprises), and it was the open-weight benchmark leader before Qwen3.8-Max’s claim.
  • Optimizing for cost → DeepSeek V4 Pro. At $0.017 per typical task it’s an order of magnitude cheaper than K3. Watch the 2x peak-hour surcharge (1–4 & 6–10 UTC) and cache your prompts ($0.043 cache-hit).
  • Chasing the absolute frontier → Qwen3.8-Max, if the independent benchmarks confirm Alibaba’s claims and the weights ship on schedule. At 2.4T it’s the largest of the three.

Watch Outs

  • Qwen3.8-Max benchmarks are vendor-reported as of Aug 4, 2026. Don’t rebuild your stack around unverified numbers.
  • DeepSeek peak pricing doubles in Beijing-time windows — model your real traffic pattern before assuming the off-peak rate.
  • Scale ≠ deployability. All three are large MoE models; even Qwen3.8-Max’s “open weights” will demand serious GPU memory.

Verdict

  • Best proven open-weight model → Kimi K3
  • Cheapest to run → DeepSeek V4 Pro
  • Biggest / potential new leader → Qwen3.8-Max (pending independent benchmarks)

Sources