UseRightAI
HomeModelsAsk AIComparePricingWhat's New
UseRightAICut through AI hype. Pick what works.

Independent AI model tracker. Live pricing, real benchmarks, zero vendor bias.

X (Twitter)LinkedInUpdatesContact

Compare

Opus 4.8 vs Opus 4.7Fable 5 vs Opus 4.8New AI Models 2026ChatGPT vs ClaudeGPT-4o vs Claude SonnetClaude vs GeminiDeepSeek vs ChatGPTMistral vs ClaudeGemini Flash vs GPT-4o MiniLlama vs ChatGPTAll comparisons →Build your own →

Best For

CodingWritingDevelopersProduct ManagersDesignersSalesBest Cheap AIBest Free AI

Pricing & Data

API Token PricingCost per TaskPrice HistoryBenchmark ScoresPrivacy & SafetySubscription PlansCost CalculatorWhich AI is Cheapest?Cheapest AI APIs

Company

About UseRightAIContactWhat ChangedAll ModelsEditorial PolicyDisclosuresPrivacy PolicyTerms of Service

© 2026 UseRightAI. Independent · Free forever · Not affiliated with any AI provider.

Affiliate links are clearly labeled. See disclosures.

Home/Cheapest Moonshot Model Worth Using
Best budget pickMoonshot · Pricing

Cheapest Moonshot Model Worth Using

Kimi K2.7 Code is Moonshot's cheapest model at $0.95/1M input tokens — 68% less than the flagship Kimi K3. It is also the best capability-per-dollar pick in the lineup.

Last verified Aug 6, 2026/Model data modified Aug 6, 2026
Rankings refresh dailyScored on 6 criteriaNo paid rankings
MoonshotBudget
Input cost
$0.95/1M
Context
256k tokens
Speed
Fast

Clear recommendation block

The shortest way to see the safest default, the lower-cost option, and the specialist pick before you read deeper.

Best overall model

Kimi K2.7 Code

View
Why this recommendation

Kimi K2.7 Code is the safest overall answer here when you want the strongest default instead of the lowest list price.

MoonshotBudget
Best for
Cost-efficient agentic coding
Price
$0.95/1M
Context
256k tokens
Best budget model

Mistral: Mistral Nemo

View
Why this recommendation

Mistral: Mistral Nemo is the lower-cost option to start with when you still need useful output at scale.

MistralBudget
Best for
Teams needing a cheap, fast, multilingual workhorse for classification, summarization, or light coding tasks at scale.
Price
$0.02/1M
Context
131k tokens
Best for speed

Kimi K3

View
Why this recommendation

Kimi K3 is the better pick when response speed matters more than maximum reasoning depth.

MoonshotPremium
Best for
Frontier-level reasoning and agentic coding
Price
$3.00/1M
Context
1M tokens

Why this page recommends it

Kimi K2.7 Code is the lowest-cost Moonshot model: $0.95/1M input, $4/1M output.

Kimi K2.7 Code is the best capability-per-dollar pick (budget score 85/100).

Kimi K3 costs 3x more on input — reserve it for work where quality is the bottleneck.

Decision notes

Choose Kimi K2.7 Code for high-volume, low-stakes tasks like classification, extraction, and drafts.

Choose Kimi K2.7 Code as the everyday default if you want one budget model.

Route only the hardest tasks to Kimi K3 — a two-tier setup usually cuts spend 60–80%.

Interactive decision lab

Test the recommendation against your priority

Switch the scoring lens to see whether the top answer changes when you care more about cost, speed, or long-document work.

#1Kimi K388 pts
#2Kimi K2.7 Code73 pts
Quality first

Kimi K3

Moonshot / Premium / Aug 6, 2026

88

Closest Chinese challenger to the frontier — #4 overall on intelligence.

Ranks models by the broadest mix of coding, writing, research, and long-context usefulness.

Cost
$3.00/1M
$15.00/1M out
Speed
Deliberate
2/5 score
Context
1M tokens
input window
View model
Data-backed recommendation
Avoid this pick if

You need fast responses or predictable output costs — always-on thinking burns tokens.

Recommended comparisons

The fastest way to see where the recommendation shifts when your priority changes.

MoonshotBudgetBest budget pick

Kimi K2.7 Code

Value coding specialist — 1T MoE agentic coder at budget prices.

Best use case
Cost-efficient agentic coding
Input
$0.95/1M
Pricing
Budget
Speed
Fast
Context
256k tokens
Open weightsCodingBudget
MoonshotPremiumOption 2

Kimi K3

Closest Chinese challenger to the frontier — #4 overall on intelligence.

Best use case
Frontier-level reasoning and agentic coding
Input
$3.00/1M
Pricing
Premium
Speed
Deliberate
Context
1M tokens
Open weightsReasoningFlagship

Side-by-side specs

Every figure below is the provider's list price or a published capability score — the same numbers the recommendation on this page is built from.

ModelInputOutputEst. monthContextSpeedCodingWritingResearch
Kimi K2.7 CodeMoonshot$0.95/1M$4.00/1M$18256k tokensFast886870
Kimi K3Moonshot$3.00/1M$15.00/1M$601M tokensDeliberate969093

Capability scores are out of 100 and reflect our own weighting of published benchmarks and production signals — see how we evaluate models. “Est. month” assumes 10M input and 2M output tokens at list price, with no batch or caching discounts applied, so treat it as a ceiling.

The case for each model

What each one is genuinely good at, where it falls down, and the situations we would steer you away from it — not just the headline score.

Kimi K2.7 Code

Best budget pickMoonshot

An open-weight 1T-parameter MoE (32B active) coding specialist tuned for long-horizon agentic software engineering with markedly better token efficiency than its predecessor.

Input
$0.95/1M
Output
$4.00/1M
Context
256k tokens
Speed
Fast

What people actually use it for

  • Agentic coding with the Kimi Code terminal CLI at $0.95/1M input
  • High-volume code review and refactoring where thinking-token burn matters (~30% fewer than K2.6)
  • Self-hosted coding infra under a modified MIT license

Where it wins

  • +21.8% over Kimi K2.6 on Kimi Code Bench v2 while using roughly 30% fewer thinking tokens
  • Only 32B active params per token — fast and cheap to serve at $0.95/$4.00 per 1M (cache hits $0.19)
  • Modified MIT license with weights on Hugging Face; pairs with the Kimi Code terminal CLI

Where it falls down

  • 256K context is a quarter of what 2026 rivals offer for large-repo agent work
  • Headline gains are on Moonshot's own in-house benchmark; general reasoning lags the Western frontier

Skip it if

Your agent needs big-repo context (256K cap) or frontier general reasoning.

Our verdict

The value pick among coding specialists. K3 superseded it at the frontier a month later, but for pure coding-agent volume at a quarter of K3's input price, K2.7 Code remains the smarter buy.

Model ID kimi-k2.7-code; weights on Hugging Face June 12, 2026. Kimi Code membership from $19/mo.

Kimi K3

Moonshot

Moonshot's 2.8-trillion-parameter multimodal reasoning flagship with always-on thinking — the largest open-weight model ever released and the closest Chinese challenger to the Western frontier.

Input
$3.00/1M
Output
$15.00/1M
Context
1M tokens
Speed
Deliberate

What people actually use it for

  • Hardest reasoning tasks — #4 of all models on AA Intelligence Index v4.1 (57.1), ahead of Claude Opus 4.8
  • Agentic coding at 81.2 FrontierSWE and 88.3 Terminal-Bench 2.0 (Moonshot-reported)
  • 1M-context research synthesis with always-on extended thinking

Where it wins

  • AA Intelligence Index v4.1: 57.1 — #4 overall, behind only Claude Fable 5 and GPT-5.6 Sol, ahead of Claude Opus 4.8
  • FrontierSWE 81.2 and Terminal-Bench 2.0 88.3 — frontier-grade agentic coding numbers
  • Open weights (July 26, 2026) — at 2.8T parameters, the largest open-weight release in history

Where it falls down

  • Most expensive Chinese-lab model ever ($3/$15) with always-on thinking driving high output-token burn and slow responses
  • 2.8T size makes self-hosting impractical despite open weights; consumer signups were paused July 19 over GPU capacity

Skip it if

You need fast responses or predictable output costs — always-on thinking burns tokens.

Our verdict

The first Chinese model to genuinely crowd the Western frontier — #4 on aggregate intelligence ahead of Opus 4.8. The always-on thinking makes it slow and output-heavy, so cost per task runs above the sticker price. A serious Opus-class alternative if latency isn't critical.

Released July 16, 2026; open weights July 26. Cache-hit input $0.30/1M. Subscriptions: Adagio (free) to Vivace $199/mo; full 1M context only on Allegro ($99) and up. New signups paused July 19 near GPU capacity, reopening in batches.

Explore related decisions

Moonshot
Kimi K2.7 CodeValue coding specialist — 1T MoE agentic coder at budget prices.Read guide
Guide
MoonshotSee the full breakdown and our current recommendation.Read guide
Guide
Best Cheap AIThe cheapest AI models ranked by real value: GPT-4o Mini at $0.15/1M, Gemini Flash at $0.075/1M, DeepSeek V3 at $0.07/1M. Find which budget AI is actually…Read guide
Guide
Best Cheap AI API in 2026The cheapest AI APIs ranked by actual value — DeepSeek V3 at $0.07/1M, Gemini Flash at $0.075/1M, GPT-4o Mini at $0.15/1M. Includes real cost examples, no-code…Read guide
Tool
AI API cost calculatorModel your monthly spend from real token prices — input and output sides both counted.Read guide
Pricing
AI API pricing comparisonInput and output cost per million tokens for every model, updated when providers change prices.Read guide
Moonshot · Coding
Best Moonshot Model for CodingEvery Moonshot model ranked for coding — capability scores, price per 1M tokens, and context windows, with a clear top pick and a budget option.Read guide
Moonshot · Writing
Best Moonshot Model for WritingEvery Moonshot model ranked for writing — capability scores, price per 1M tokens, and context windows, with a clear top pick and a budget option.Read guide

Quick links

Browse all modelsCompare pricingView Kimi K2.7 CodeView Kimi K3

How we evaluate AI models

UseRightAI recommendations are based on practical decision factors people actually feel in day-to-day use.

Newsletter

Get updates when cheapest moonshot model worth using changes

Useful if you care about ranking shifts, pricing changes, or a better recommendation appearing in this decision path.

No spam. Useful updates only. Affiliate disclosures always clearly labeled.

FAQ

What is the cheapest Moonshot model?

Kimi K2.7 Code at $0.95/1M input and $4/1M output tokens. Value coding specialist — 1T MoE agentic coder at budget prices.

Is the cheapest Moonshot model good enough for real work?

Kimi K2.7 Code is the best capability-per-dollar pick in Moonshot's lineup (budget score 85/100). It handles cost-efficient agentic coding well — step up to Kimi K3 only where quality visibly falls short.

How much cheaper is Kimi K2.7 Code than Moonshot's flagship?

Kimi K2.7 Code costs $0.95/1M input vs $3/1M for Kimi K3 — a 68% saving on input tokens.

Which cheap Moonshot model has the largest context window?

Kimi K3 — 1M tokens at $3/1M input. Context is where budget models are least compromised: you usually lose reasoning depth before you lose window size, so a cheap model is often a perfectly good choice for summarising or extracting from long documents.

What do you give up with Kimi K2.7 Code?

256K context is a quarter of what 2026 rivals offer for large-repo agent work. Headline gains are on Moonshot's own in-house benchmark; general reasoning lags the Western frontier. Avoid it if your agent needs big-repo context (256K cap) or frontier general reasoning.

What does Kimi K2.7 Code cost per month in practice?

On a moderate workload of 10M input and 2M output tokens, Kimi K2.7 Code runs about $17.50 against $60.00 for Kimi K3 — a difference of $42.50 a month at the same volume. Output tokens dominate the bill on both, so the length of the responses you generate matters far more than the length of your prompts.

Should I use one cheap Moonshot model or mix tiers?

Mixing is almost always cheaper for the same quality. Route high-volume, low-stakes work — classification, extraction, first drafts, routine agent steps — to Kimi K2.7 Code, and reserve Kimi K3 for the calls where a wrong answer costs real time. Teams that split this way typically cut spend substantially without a quality drop anyone notices, because most tokens in a real workload are not hard problems.