DeepSeek V4-Flash
A 284B-parameter (13B active) MoE workhorse re-post-trained for agentic and coding tasks — beats the V4-Pro preview on every published agent benchmark at ultra-commodity pricing.
80.8% SWE-bench Verified at introductory Flash pricing.
High-volume coding and long-context work at introductory Flash pricing
You are planning 2027 spend and need price certainty — the introductory rate expires December 31, 2026.
Compare every model's knowledge cutoff, max output, and context window.
Released August 13, 2026, only three weeks after Gemini 3.6 Flash. Introductory pricing of $0.75/$3.75 runs through December 31, 2026; on January 1, 2027 it doubles to $1.50/$7.50, which is exactly Gemini 3.6 Flash's rate.
80.8% on SWE-bench Verified — frontier-class coding from a Flash-tier model
Large jumps over 3.6 Flash on software engineering: FrontierCode 34.4% to 43.6%, DeepSWE 49.0% to 65.3%
Artificial Analysis Intelligence Index of 56 at high thinking level, with a 1M token context window
The $0.75/$3.75 launch price is introductory — it doubles to $1.50/$7.50 on January 1, 2027
Still short of Claude Opus 5 (96%) and GPT-5.6 Sol (96.2%) on SWE-bench Verified for the hardest coding work
What people actually use Gemini 3.7 Flash for.
Bulk code review and refactoring where 80.8% SWE-bench Verified is enough and volume matters
1M-context document and repository analysis at $0.75/1M input
Agentic loops that need frontier-adjacent coding quality without frontier pricing
The nearest models people weigh against it, and what actually separates them.
vs DeepSeek V4-Flash — Against DeepSeek V4-Flash (DeepSeek), Gemini 3.7 Flash costs about 91% more per token and takes 1x the context. DeepSeek V4-Flash is the one to check first if the price difference matters more than the ceiling.
vs DeepSeek V4-Pro — Against DeepSeek V4-Pro (DeepSeek), Gemini 3.7 Flash costs about 71% more per token, takes 1x the context and answers faster. DeepSeek V4-Pro is the one to check first if the price difference matters more than the ceiling.
vs Devstral Small 1.1 — Against Devstral Small 1.1 (Mistral), Gemini 3.7 Flash costs about 91% more per token and takes 8x the context. Devstral Small 1.1 is the one to check first if the price difference matters more than the ceiling.
Price History
↑300% since Sep 1
8 data points · tracked daily since Sep 1, 2026
High-volume coding and long-context work at introductory Flash pricing. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
A 284B-parameter (13B active) MoE workhorse re-post-trained for agentic and coding tasks — beats the V4-Pro preview on every published agent benchmark at ultra-commodity pricing.
DeepSeek's 1.6T-parameter (49B active) MoE flagship with hybrid sparse attention — near-frontier coding and reasoning at roughly a tenth of closed-rival pricing, MIT-licensed open weights.
Devstral Small 1.1 is Mistral's code-specialized small model, purpose-built for software engineering tasks including code generation, debugging, and repository-level reasoning. It succeeds Devstral Small 1.0 with improved instruction following and agentic coding capabilities at a fraction of flagship model costs.
Pricing moves, ranking shifts, and capability updates.
Gemini 3.7 Flash output pricing changed from $3.75/1M to $0.94/1M (↓ cheaper, 75% cut).
View modelGemini 3.7 Flash input pricing changed from $0.75/1M to $0.19/1M (↓ cheaper, 75% cut).
View modelGemini 3.7 Flash costs $0.75 per million input tokens and $3.75 per million output tokens on the API, with cached input at $0.075 per million. A month of 10M input and 2M output tokens runs about $15.00 at list price, before any batch or caching discounts.
Gemini 3.7 Flash has a 1M tokens context window, with up to 66k tokens of output per response. That is the total of prompt plus response the model can hold in one request.
Gemini 3.7 Flash's training data runs through March 2026, and the model was released on August 13, 2026. For anything after that date it needs web search or documents in the prompt.
Gemini 3.7 Flash is best for high-volume coding and long-context work at introductory flash pricing. It is a strong fit when that workflow matters more than the tradeoffs around balanced pricing and fast speed.
You are planning 2027 spend and need price certainty — the introductory rate expires December 31, 2026.
DeepSeek V4-Pro (DeepSeek) at $0.43/1M/1M input against Gemini 3.7 Flash's $0.75/1M/1M — roughly 71% less per token all in. Best open-weights flagship — near-frontier coding at a tenth of the price. Compare it first if Gemini 3.7 Flash's pricing is the thing stopping you.
DeepSeek V4-Flash — fast against Gemini 3.7 Flash's fast, with 1M tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.