UseRightAI
HomeModelsAsk AIComparePricingWhat's New
UseRightAICut through AI hype. Pick what works.

Independent AI model tracker. Live pricing, real benchmarks, zero vendor bias.

X (Twitter)LinkedInUpdatesContact

Compare

Opus 4.8 vs Opus 4.7Fable 5 vs Opus 4.8New AI Models 2026ChatGPT vs ClaudeGPT-4o vs Claude SonnetClaude vs GeminiDeepSeek vs ChatGPTMistral vs ClaudeGemini Flash vs GPT-4o MiniLlama vs ChatGPTAll comparisons →Build your own →

Best For

CodingWritingDevelopersProduct ManagersDesignersSalesBest Cheap AIBest Free AI

Pricing & Data

API Token PricingCost per TaskPrice HistoryBenchmark ScoresPrivacy & SafetySubscription PlansPlan Usage LimitsCost CalculatorWhich AI is Cheapest?Cheapest AI APIs

Company

About UseRightAIContactWhat ChangedAll ModelsEditorial PolicyDisclosuresPrivacy PolicyTerms of Service

© 2026 UseRightAI. Independent · Free forever · Not affiliated with any AI provider.

Affiliate links are clearly labeled. See disclosures.

HomeModelsGemini 3.7 Flash
GoogleBalanced

Gemini 3.7 Flash

80.8% SWE-bench Verified at introductory Flash pricing.

89
Coding
82
Writing
85
Research
78
Images
82
Value
90
Long Context
Published benchmarks
80.8%
SWE-bench
Use this when

High-volume coding and long-context work at introductory Flash pricing

Skip this if

You are planning 2027 spend and need price certainty — the introductory rate expires December 31, 2026.

Pricing
$0.75/1M in
$3.75/1M out
↑300%since Sep 2026
Context
1.0M tokens
Speed
Fast

Gemini 3.7 Flashspecs & pricing

Verified Sep 4, 2026 against the AI Gateway catalog
Input price
$0.75 / 1M tokens
Output price
$3.75 / 1M tokens
Cached input(prompt-cache read)
$0.075 / 1M tokens
Context window
1M tokens
Max output
66k tokens
Knowledge cutoff
Mar 2026
Released
Aug 13, 2026
Input modalities
Text, Image, PDF, Video
Output modalities
Text
Reasoning mode
Yes
Tool use
Yes
Gateway model ID
google/gemini-3.7-flash

Compare every model's knowledge cutoff, max output, and context window.

Released August 13, 2026, only three weeks after Gemini 3.6 Flash. Introductory pricing of $0.75/$3.75 runs through December 31, 2026; on January 1, 2027 it doubles to $1.50/$7.50, which is exactly Gemini 3.6 Flash's rate.

How to access
Subscription
Google One AI Premium — $19.99/mo
API
$0.75/1M input tokens
Subscription = chat interface. API = build with it. Compare all subscription plans · Google One AI Premium usage limits
Switch to instead if...
Best overall
GPT-6 Astra
Cheaper option
DeepSeek V4-Pro
Faster option
DeepSeek V4-Flash

Strengths

80.8% on SWE-bench Verified — frontier-class coding from a Flash-tier model

Large jumps over 3.6 Flash on software engineering: FrontierCode 34.4% to 43.6%, DeepSWE 49.0% to 65.3%

Artificial Analysis Intelligence Index of 56 at high thinking level, with a 1M token context window

Weaknesses

The $0.75/$3.75 launch price is introductory — it doubles to $1.50/$7.50 on January 1, 2027

Still short of Claude Opus 5 (96%) and GPT-5.6 Sol (96.2%) on SWE-bench Verified for the hardest coding work

Real-world use cases

What people actually use Gemini 3.7 Flash for.

Bulk code review and refactoring where 80.8% SWE-bench Verified is enough and volume matters

1M-context document and repository analysis at $0.75/1M input

Agentic loops that need frontier-adjacent coding quality without frontier pricing

How Gemini 3.7 Flash compares

The nearest models people weigh against it, and what actually separates them.

vs DeepSeek V4-Flash — Against DeepSeek V4-Flash (DeepSeek), Gemini 3.7 Flash costs about 91% more per token and takes 1x the context. DeepSeek V4-Flash is the one to check first if the price difference matters more than the ceiling.

vs DeepSeek V4-Pro — Against DeepSeek V4-Pro (DeepSeek), Gemini 3.7 Flash costs about 71% more per token, takes 1x the context and answers faster. DeepSeek V4-Pro is the one to check first if the price difference matters more than the ceiling.

vs Devstral Small 1.1 — Against Devstral Small 1.1 (Mistral), Gemini 3.7 Flash costs about 91% more per token and takes 8x the context. Devstral Small 1.1 is the one to check first if the price difference matters more than the ceiling.

Price History

Gemini 3.7 Flash pricing over time

↑300% since Sep 1

$0.810$0.651$0.491$0.332$0.173Sep 1Sep 2Sep 4Sep 5Sep 7Sep 8

8 data points · tracked daily since Sep 1, 2026

Ready to try it?

Start using Gemini 3.7 Flash

High-volume coding and long-context work at introductory Flash pricing. Start free — no card required.

Try Gemini 3.7 Flash freeCompare alternatives

Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.

Compare alternatives

Similar models worth checking before you commit.

All Gemini 3.7 Flash alternatives →
DeepSeekBudget

DeepSeek V4-Flash

A 284B-parameter (13B active) MoE workhorse re-post-trained for agentic and coding tasks — beats the V4-Pro preview on every published agent benchmark at ultra-commodity pricing.

Verdict
Best agentic capability per dollar in the directory.
Quality score
76%
Pricing
$0.14/1M in
$0.28/1M out
Speed
Fast
4/5 speed
Context
1M tokens
Official V4-Flash-0731 release July 31, 2026; weights on Hugging Face, API in public beta. Only DeepSeek model supporting the Responses API. DeepSeek has warned of a future price increase.
Open weightsBudgetAgenticUltra cheap1M context
Best for
High-volume agentic coding and tool-use pipelines
View model
DeepSeekBudget

DeepSeek V4-Pro

DeepSeek's 1.6T-parameter (49B active) MoE flagship with hybrid sparse attention — near-frontier coding and reasoning at roughly a tenth of closed-rival pricing, MIT-licensed open weights.

Verdict
Best open-weights flagship — near-frontier coding at a tenth of the price.
Quality score
82%
Pricing
$0.43/1M in
$0.87/1M out
Speed
Balanced
3/5 speed
Context
1M tokens
Open-weight preview April 24; GA ~July 20, 2026. Off-peak pricing verified on api-docs.deepseek.com; Beijing-business-hours surge doubles it. Legacy deepseek-chat/reasoner endpoints retired July 24, 2026.
Open weightsCodingReasoningBudget1M context
Best for
Frontier-level coding and reasoning on a budget
View model
MistralBudget

Devstral Small 1.1

Devstral Small 1.1 is Mistral's code-specialized small model, purpose-built for software engineering tasks including code generation, debugging, and repository-level reasoning. It succeeds Devstral Small 1.0 with improved instruction following and agentic coding capabilities at a fraction of flagship model costs.

Verdict
The best dollar-for-dollar coding model for agentic pipelines that doesn't need to do anything else.
Quality score
54%
Pricing
$0.10/1M in
$0.30/1M out
Speed
Fast
4/5 speed
Context
131k tokens
Available via Mistral API and can be self-hosted via open weights. Pricing is among the lowest available for a code-specialized model. Designed to work within coding agent frameworks like SWE-agent and OpenHands.
code-specialistbudgetagenticopen-source-friendlySWE-bench
Best for
Developers who need a cheap, fast coding assistant for agentic workflows, code review, and multi-file repo tasks without paying flagship prices.
View model

Gemini 3.7 Flash head-to-head

All Gemini 3.7 Flash alternatives →Gemini 3.7 Flash vs Gemini 3.6 Flash →Gemini 3.7 Flash vs Claude Sonnet 5 →Gemini 3.7 Flash vs GPT-5.6 Terra →Gemini 3.7 Flash vs DeepSeek V4-Pro →Gemini 3.7 Flash vs Qwen 3.8 Flash →Grok 4.6 vs Gemini 3.7 Flash →Claude Fable 5.1 vs Gemini 3.7 Flash →GPT-6 Astra vs Gemini 3.7 Flash →View benchmark scores →

Change history

Pricing moves, ranking shifts, and capability updates.

PricingSep 1, 2026

Gemini 3.7 Flash — output price cut

Gemini 3.7 Flash output pricing changed from $3.75/1M to $0.94/1M (↓ cheaper, 75% cut).

View model
PricingSep 1, 2026

Gemini 3.7 Flash — input price cut

Gemini 3.7 Flash input pricing changed from $0.75/1M to $0.19/1M (↓ cheaper, 75% cut).

View model

FAQ

How much does Gemini 3.7 Flash cost?

Gemini 3.7 Flash costs $0.75 per million input tokens and $3.75 per million output tokens on the API, with cached input at $0.075 per million. A month of 10M input and 2M output tokens runs about $15.00 at list price, before any batch or caching discounts.

What is the context window of Gemini 3.7 Flash?

Gemini 3.7 Flash has a 1M tokens context window, with up to 66k tokens of output per response. That is the total of prompt plus response the model can hold in one request.

What is the knowledge cutoff of Gemini 3.7 Flash?

Gemini 3.7 Flash's training data runs through March 2026, and the model was released on August 13, 2026. For anything after that date it needs web search or documents in the prompt.

What is Gemini 3.7 Flash best for?

Gemini 3.7 Flash is best for high-volume coding and long-context work at introductory flash pricing. It is a strong fit when that workflow matters more than the tradeoffs around balanced pricing and fast speed.

When should I avoid Gemini 3.7 Flash?

You are planning 2027 spend and need price certainty — the introductory rate expires December 31, 2026.

What is a cheaper alternative to Gemini 3.7 Flash?

DeepSeek V4-Pro (DeepSeek) at $0.43/1M/1M input against Gemini 3.7 Flash's $0.75/1M/1M — roughly 71% less per token all in. Best open-weights flagship — near-frontier coding at a tenth of the price. Compare it first if Gemini 3.7 Flash's pricing is the thing stopping you.

What is a faster alternative to Gemini 3.7 Flash?

DeepSeek V4-Flash — fast against Gemini 3.7 Flash's fast, with 1M tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.

Newsletter

Get notified when Gemini 3.7 Flash pricing changes

We track pricing daily. When this model drops or spikes, you'll know first.

No spam. Useful updates only. Affiliate disclosures always clearly labeled.

User reviews

No reviews yet — be the first.