SAT, AUGUST 01, 2026
Independent · In‑Depth · Practitioner‑Tested
Code Tools

Cheapest Open-Weight Coding Models (August 2026): DeepSeek, Laguna, Kimi, Grok — Full Comparison

After DeepSeek V4 Flash 0731 Official Release — The Full Cost-Performance Ranking

🕐 6 min read 👁 18 views 📅 Aug 1, 2026

COST-PERFORMANCE RANKING — AUGUST 1, 2026

#1 Price (input): Laguna S 2.1 — $0.10/M
#2 Price: DeepSeek V4 Flash 0731 — $0.14/M
#3 Price: Grok 4.5 — $2/M
#4 Price: Kimi K3 API — $3/M (+ sanctions risk)
#1 Terminal-Bench: Grok 4.5 — 83.3%
#2 Terminal-Bench: V4 Flash 0731 — 82.7%
#3 Terminal-Bench: Laguna S 2.1 — 70.2%
#1 SWE Marathon: Kimi K3 — 42.0% (K3 not published on Terminal-Bench)
ModelInput /1MTerminal-BenchBest forRisk
Laguna S 2.1$0.1070.2%Lowest cost, DGX Spark deployNone
DeepSeek V4 Flash 0731$0.1482.7%Best cost-performance, Codex-compatibleNone
Grok 4.5$2.0083.3%Cursor, live X data54% hallucination (AA)
Kimi K3$3.00Not publishedSWE Marathon #1Sanctions threat + 51% hallucination

The new default recommendation for August 2026: DeepSeek V4 Flash 0731 at $0.14/M for API-first agentic coding workloads. It is 14× cheaper than Grok 4.5, 0.6 points behind on Terminal-Bench, open-weight MIT, and Codex-compatible out of the box. Laguna S 2.1 remains the choice where hardware accessibility (DGX Spark) or Western-company provenance are the deciding factors.

Last updated August 1, 2026. Related: V4 Flash 0731 full review → · Three coding agents comparison →

⚖ Our Verdict

DeepSeek V4 Flash 0731 is the new cost-performance leader: $0.14/M, 82.7% Terminal-Bench, 14× cheaper than Grok 4.5 at near-identical benchmark performance, MIT license, Codex-compatible. Laguna S 2.1 wins on absolute lowest price ($0.10/M) and DGX Spark accessibility. Grok 4.5 wins on Terminal-Bench (83.3%) and Cursor/live X data. Kimi K3 leads SWE Marathon but carries sanctions risk.