DeepSeek V4-Flash
A 284B-parameter (13B active) MoE workhorse re-post-trained for agentic and coding tasks — beats the V4-Pro preview on every published agent benchmark at ultra-commodity pricing.
Best open-weights flagship — near-frontier coding at a tenth of the price.
Frontier-level coding and reasoning on a budget
You need vision input, verified agentic performance, or predictable pricing (surge pricing and an announced increase loom).
Compare every model's knowledge cutoff, max output, and context window.
Open-weight preview April 24; GA ~July 20, 2026. Off-peak pricing verified on api-docs.deepseek.com; Beijing-business-hours surge doubles it. Legacy deepseek-chat/reasoner endpoints retired July 24, 2026.
80.6% SWE-bench Verified (self-reported) — reported as tied with Gemini 3.1 Pro
93.5% LiveCodeBench and Codeforces 3206 — elite competitive-coding results
1M context with 384K max output at $0.87/1M output — an order of magnitude cheaper than closed frontier models
Independent harnesses report much lower agentic scores than the self-reported numbers; trails GPT-5.6 and Opus-class on hard agentic evals
Peak-hour surge pricing doubles rates, a price increase is announced, and it's text-only (no vision)
What people actually use DeepSeek V4-Pro for.
Repository-level coding — 80.6% SWE-bench Verified (self-reported), the top open-weights score at release
Competitive-programming-grade reasoning (Codeforces rating 3206)
Self-hosted frontier capability under an MIT license
The nearest models people weigh against it, and what actually separates them.
vs DeepSeek V4-Flash — Against DeepSeek V4-Flash (DeepSeek), DeepSeek V4-Pro costs about 68% more per token and answers slower. DeepSeek V4-Flash is the one to check first if the price difference matters more than the ceiling.
vs Devstral Small 1.1 — Against Devstral Small 1.1 (Mistral), DeepSeek V4-Pro costs about 69% more per token, takes 7.6x the context and answers slower. Devstral Small 1.1 is the one to check first if the price difference matters more than the ceiling.
vs GLM-5.2 — Against GLM-5.2 (Z.ai), DeepSeek V4-Pro runs about 78% cheaper per token. Take DeepSeek V4-Pro unless you specifically need what GLM-5.2 does better.
Price History
→0% since Aug 7
38 data points · tracked daily since Aug 7, 2026
Frontier-level coding and reasoning on a budget. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
A 284B-parameter (13B active) MoE workhorse re-post-trained for agentic and coding tasks — beats the V4-Pro preview on every published agent benchmark at ultra-commodity pricing.
Devstral Small 1.1 is Mistral's code-specialized small model, purpose-built for software engineering tasks including code generation, debugging, and repository-level reasoning. It succeeds Devstral Small 1.0 with improved instruction following and agentic coding capabilities at a fraction of flagship model costs.
Z.ai's MIT-licensed open-weight flagship — the top open-weights coding model of mid-2026, beating GPT-5.5 on agentic coding benchmarks at roughly a sixth of the cost.
DeepSeek V4-Pro costs $0.435 per million input tokens and $0.87 per million output tokens on the API. A month of 10M input and 2M output tokens runs about $6.09 at list price, before any batch or caching discounts.
DeepSeek V4-Pro has a 1M tokens context window, with up to 384k tokens of output per response. That is the total of prompt plus response the model can hold in one request.
DeepSeek V4-Pro's training data runs through May 2025, and the model was released on April 23, 2026. For anything after that date it needs web search or documents in the prompt.
DeepSeek V4-Pro is best for frontier-level coding and reasoning on a budget. It is a strong fit when that workflow matters more than the tradeoffs around budget pricing and balanced speed.
You need vision input, verified agentic performance, or predictable pricing (surge pricing and an announced increase loom).
DeepSeek V4-Flash (DeepSeek) at $0.14/1M/1M input against DeepSeek V4-Pro's $0.43/1M/1M — roughly 68% less per token all in. Best agentic capability per dollar in the directory. Compare it first if DeepSeek V4-Pro's pricing is the thing stopping you.
Devstral Small 1.1 — fast against DeepSeek V4-Pro's balanced, with 131k tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.