A strong open-weight performer for short-context coding and reasoning, hobbled by an outdated 8K context limit.
74
Coding
70
Writing
62
Research
0
Images
65
Value
18
Long Context
Use this when
Teams that need strong open-weight model performance for coding and reasoning tasks without paying flagship prices.
Skip this if
You need long document processing, multimodal inputs, or frontier-level reasoning — the 8K context window alone disqualifies it for most RAG or document analysis workflows.
Pricing
$0.65/1M in
$0.65/1M out
→0%since May 2026
Context
8k tokens
Speed
Fast
Symmetric input/output pricing at $0.65/1M tokens is straightforward but positions it oddly — it's pricier than GPT-4o Mini while lacking its multimodal features. Available via multiple inference providers including Google Vertex AI and third-party APIs.
Exceptionally strong for its parameter count — outperforms many larger open models on benchmarks like MMLU and HumanEval
Competitive instruction-following that rivals Claude Haiku and GPT-4o Mini on structured tasks
Clean, well-formatted outputs with low hallucination rates compared to similarly-sized open models
Cost-effective at $0.65/1M tokens for both input and output — symmetric pricing simplifies budgeting
Weaknesses
Tiny 8K context window is a serious limitation — Gemini 3.1 Pro and Claude Sonnet 4.6 offer 200K+ tokens
No multimodal capabilities; text-only limits applicability in modern pipelines
Reasoning depth falls short of frontier models like GPT-5.4 or Gemini 3.1 Pro on complex multi-step problems
Real-world use cases
What people actually use Gemma 2 27B for.
Generating and reviewing Python or JavaScript code snippets with accurate, well-commented outputs
Summarizing short documents or research abstracts under 6K tokens
Answering structured Q&A or classification tasks in enterprise pipelines requiring open-weight models
How Gemma 2 27B compares
The nearest models people weigh against it, and what actually separates them.
vs Gemini 2.5 Pro Preview 06-05 — Against Gemini 2.5 Pro Preview 06-05 (Google), Gemma 2 27B runs about 88% cheaper per token, gives up 128x on context and answers faster. Take Gemma 2 27B unless you specifically need what Gemini 2.5 Pro Preview 06-05 does better.
vs Claude 3.5 Sonnet — Against Claude 3.5 Sonnet (Anthropic), Gemma 2 27B runs about 96% cheaper per token, gives up 24.4x on context and answers faster. Take Gemma 2 27B unless you specifically need what Claude 3.5 Sonnet does better.
vs Claude Fable 5 — Against Claude Fable 5 (Anthropic), Gemma 2 27B runs about 98% cheaper per token, gives up 122.1x on context and answers faster. Take Gemma 2 27B unless you specifically need what Claude Fable 5 does better.
Price History
Gemma 2 27B pricing over time
→0% since May 31
90 data points · tracked daily since May 31, 2026
Ready to try it?
Start using Gemma 2 27B
Teams that need strong open-weight model performance for coding and reasoning tasks without paying flagship prices.. Start free — no card required.
Gemini 2.5 Pro Preview 06-05 is Google's most capable reasoning-focused model, featuring a massive 1M token context window and strong performance across code, math, and complex analysis tasks. It represents Google's top-tier offering in the Gemini 2.5 generation, optimized for depth over speed.
Verdict
Google's most capable model — a top-tier reasoning and coding powerhouse with an unmatched context window, held back only by its preview status and output cost.
Quality score
83%
Pricing
$1.25/1M in
$10.00/1M out
Speed
Deliberate
2/5 speed
Context
1.0M tokens
This is a preview model (06-05 date suffix indicates a versioned snapshot); Google may deprecate or modify it before a stable GA release. Pricing tiers differ based on prompt length — prompts over 200K tokens are charged at $2.50/1M input and $15/1M output, significantly increasing cost for very long-context use cases.
FlagshipLong ContextReasoningCodingPreview
Best for
Complex multi-step reasoning, large codebase analysis, and tasks requiring deep synthesis across very long documents.
Claude 3.5 Sonnet is Anthropic's mid-cycle flagship model, balancing strong reasoning, coding, and instruction-following with a 200K context window. It sits between Haiku and Opus in Anthropic's lineup, offering near-flagship quality at a lower cost than top-tier models.
Verdict
One of the best models for coding and complex instruction-following, but its premium pricing demands premium use cases.
Quality score
81%
Pricing
$6.00/1M in
$30.00/1M out
Speed
Balanced
3/5 speed
Context
200k tokens
Pricing at $6 input / $30 output per million tokens is significantly higher than GPT-4o ($2.50/$10). Best accessed via Anthropic API or Amazon Bedrock. Claude 3.5 Sonnet (October 2024 version) supersedes the June 2024 release with improved performance.
Anthropic's new Mythos-class flagship and the most capable coding model anyone can use — 80.3% SWE-Bench Pro, an 11-point jump over Opus 4.8. 1M context, 128K output, native parallel subagents. Released June 9, 2026.
Verdict
New global #1 — 80.3% SWE-Bench Pro, the most capable model generally available.
Quality score
98%
Pricing
$10.00/1M in
$50.00/1M out
Speed
Deliberate
2/5 speed
Context
1M tokens
Launched June 9, 2026 as the public, Mythos-class release. Available on the Claude API, Microsoft Foundry, and Google Vertex AI. Free for all users until June 22, 2026. Same underlying model as Claude Mythos 5, with safeguards that block specific high-risk cyber responses.
Coding leaderSWE-Bench Pro #1Mythos-classParallel subagentsAgenticLong contextPremiumNew
Best for
The hardest coding tasks, autonomous multi-step agents, and frontier-grade reasoning
Pricing moves, ranking shifts, and capability updates.
New ModelMar 27, 2026
Google: Gemma 2 27B — added to UseRightAI
Google: Gemma 2 27B (Google) is now indexed. A strong open-weight performer for short-context coding and reasoning, hobbled by an outdated 8K context limit.
Gemma 2 27B costs $0.65 per million input tokens and $0.65 per million output tokens on the API. A month of 10M input and 2M output tokens runs about $7.80 at list price, before any batch or caching discounts.
What is Gemma 2 27B best for?
Gemma 2 27B is best for teams that need strong open-weight model performance for coding and reasoning tasks without paying flagship prices.. It is a strong fit when that workflow matters more than the tradeoffs around balanced pricing and fast speed.
When should I avoid Gemma 2 27B?
You need long document processing, multimodal inputs, or frontier-level reasoning — the 8K context window alone disqualifies it for most RAG or document analysis workflows.
What is a cheaper alternative to Gemma 2 27B?
GPT-5.1-Codex-Max (OpenAI) at $1.25/1M/1M input against Gemma 2 27B's $0.65/1M/1M. The strongest choice for serious software engineering work, provided you can absorb the output-side pricing. Compare it first if Gemma 2 27B's pricing is the thing stopping you.
What is a faster alternative to Gemma 2 27B?
Gemini 2.5 Pro Preview 06-05 — deliberate against Gemma 2 27B's fast, with 1.0M tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
Get notified when Gemma 2 27B pricing changes
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.