AI agents · OpenClaw · self-hosting · automation

Quick Answer

Inkling vs Kimi K3 vs DeepSeek V4 Pro to Self-Host

Published:

The Short Answer

Pick by goal: Kimi K3 for the highest open intelligence, Inkling for a permissive, fine-tunable base, DeepSeek V4 Pro for lowest serving cost. All three are open-weight and released in July 2026.

The Comparison

ModelReleasedParamsLicenseIntelligence IndexHosted API
Kimi K3Jul 27Trillion-scale MoEOpen weights~57 (#3 overall)$3 / $15
InklingJul 15975B total / 41B activeApache 2.0well-rounded base— (self-host)
DeepSeek V4 Pro2026Trillion-scale MoEOpen weights~44 (Max)~$0.435 / $0.87*

DeepSeek off-peak; 2x during peak (1–4 & 6–10 UTC). Cache-hit ~$0.043.

Kimi K3 — highest open intelligence

Moonshot’s K3 ranks #3 overall on the Artificial Analysis Intelligence Index (~57), behind only Claude Fable 5 and GPT-5.6 Sol, and comparable to Opus 4.8. It adds native vision. Demand at launch was so high Moonshot briefly suspended new subscriptions. Flat $3/$15 pricing, open weights July 27.

Inkling — the customizable base

Mira Murati’s Thinking Machines released Inkling on July 15, 2026: a 975B-total / 41B-active MoE, 1M context, trained on 45T multimodal tokens (text/image/audio/video). It’s explicitly not chasing top closed-model scores — it’s a well-rounded base for fine-tuning, licensed Apache 2.0, with an Inkling-Small (12B active) preview and Tinker console support.

DeepSeek V4 Pro — the cost floor

Text-only, trillion-scale MoE, priced near the absolute floor (~$0.435/$0.87 off-peak, 2x peak). Intelligence trails Kimi K3 but the serving economics are unmatched for high-volume text pipelines.

What to Do

  1. Need the smartest open model: Kimi K3.
  2. Building a fine-tuned product: Inkling (Apache 2.0 is the friendliest license here).
  3. Cheapest high-volume text: DeepSeek V4 Pro.

Sources