DeepSeek's MIT-licensed model just undercut Claude Code and Codex on price — by a lot
DeepSeek's retrained V4-Flash-0731 model beats its own flagship on every published agentic coding benchmark while costing a fraction as much per token, adding a credible open-weight option to the Claude Code vs Codex vs Cursor conversation founders are already having.
14 August 2026
DeepSeek’s V4-Flash-0731 checkpoint went live on Hugging Face on 31 July 2026, MIT-licensed, and immediately outscored the company’s own larger V4-Pro-preview flagship on all nine agentic benchmarks it publishes — Terminal Bench 2.1 went from 72.1 to 82.7, and DeepSWE jumped more than sevenfold, from 7.3 to 54.4, from post-training alone. The API prices at $0.14 per million input tokens and $0.28 per million output tokens — roughly 3x cheaper than DeepSeek’s own Pro tier, and a fraction of what Claude Code or OpenAI Codex charge per token. It also ships with native OpenAI Responses API support, meaning it slots straight into Codex CLI without a custom integration layer.
Search interest around “AI coding agent pricing” and “cheapest AI coding tool” has been climbing steadily through 2026, and this release is exactly the kind of story that drives it: a credible, MIT-licensed, benchmark-leading model that costs a fraction of the market leaders. One important caveat — every one of those nine benchmark numbers is vendor-stated, run on DeepSeek’s own harness, with no independent reproduction confirmed as of this writing. Treat the headline figures the way you’d treat any vendor’s own marketing claims: directionally useful, not proof.
So what
Cheap, capable open-weight coding models widen the field beyond “which of the big three do we standardise on” — and that’s good news for anyone budgeting an AI-assisted build, because it puts real competitive pressure on Anthropic and OpenAI’s pricing. But price-per-token was never the hard part of shipping production software with AI agents; integration, code review discipline, and knowing which tasks actually benefit from agentic coding are. If you’re weighing which AI coding stack to commission a build on — and whether the cheapest option is actually the right one for your team — that’s exactly the kind of decision we help clients make. See our AI-assisted development approach or get in touch to talk through your setup.