Claude Fable 5
Anthropic's new Mythos-class flagship and the most capable coding model anyone can use — 80.3% SWE-Bench Pro, an 11-point jump over Opus 4.8. 1M context, 128K output, native parallel subagents. Released June 9, 2026.
Closest Chinese challenger to the frontier — #4 overall on intelligence.
Frontier-level reasoning and agentic coding
You need fast responses or predictable output costs — always-on thinking burns tokens.
Compare every model's knowledge cutoff, max output, and context window.
Released July 16, 2026; open weights July 26. Cache-hit input $0.30/1M. Subscriptions: Adagio (free) to Vivace $199/mo; full 1M context only on Allegro ($99) and up. New signups paused July 19 near GPU capacity, reopening in batches.
AA Intelligence Index v4.1: 57.1 — #4 overall, behind only Claude Fable 5 and GPT-5.6 Sol, ahead of Claude Opus 4.8
FrontierSWE 81.2 and Terminal-Bench 2.0 88.3 — frontier-grade agentic coding numbers
Open weights (July 26, 2026) — at 2.8T parameters, the largest open-weight release in history
Most expensive Chinese-lab model ever ($3/$15) with always-on thinking driving high output-token burn and slow responses
2.8T size makes self-hosting impractical despite open weights; consumer signups were paused July 19 over GPU capacity
What people actually use Kimi K3 for.
Hardest reasoning tasks — #4 of all models on AA Intelligence Index v4.1 (57.1), ahead of Claude Opus 4.8
Agentic coding at 81.2 FrontierSWE and 88.3 Terminal-Bench 2.0 (Moonshot-reported)
1M-context research synthesis with always-on extended thinking
The nearest models people weigh against it, and what actually separates them.
vs Claude Fable 5 — Against Claude Fable 5 (Anthropic), Kimi K3 runs about 70% cheaper per token. Take Kimi K3 unless you specifically need what Claude Fable 5 does better.
vs Claude Fable 5.1 — Against Claude Fable 5.1 (Anthropic), Kimi K3 runs about 70% cheaper per token. Take Kimi K3 unless you specifically need what Claude Fable 5.1 does better.
vs Claude Opus 4.7 — Against Claude Opus 4.7 (Anthropic), Kimi K3 runs about 40% cheaper per token. Take Kimi K3 unless you specifically need what Claude Opus 4.7 does better.
Price History
→0% since Aug 7
38 data points · tracked daily since Aug 7, 2026
Frontier-level reasoning and agentic coding. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
Anthropic's new Mythos-class flagship and the most capable coding model anyone can use — 80.3% SWE-Bench Pro, an 11-point jump over Opus 4.8. 1M context, 128K output, native parallel subagents. Released June 9, 2026.
Anthropic's September 1, 2026 frontier release and the new capability ceiling for coding, agents, and scientific work. Base pricing is unchanged at $10/$50, but cache reads dropped 75% to $0.25/1M — roughly 25% cheaper on typical workloads and up to 45% cheaper on agentic ones. 1M context, 128K output, adaptive thinking always on.
Anthropic's previous Opus flagship, now superseded by Opus 4.8. Still the second-best coding model publicly available at the same $5/$25 price.
Kimi K3 costs $3 per million input tokens and $15 per million output tokens on the API, with cached input at $0.3 per million. A month of 10M input and 2M output tokens runs about $60.00 at list price, before any batch or caching discounts.
Kimi K3 has a 1M tokens context window, with up to 131k tokens of output per response. That is the total of prompt plus response the model can hold in one request.
Kimi K3 is best for frontier-level reasoning and agentic coding. It is a strong fit when that workflow matters more than the tradeoffs around premium pricing and deliberate speed.
You need fast responses or predictable output costs — always-on thinking burns tokens.
GPT-5.1-Codex-Max (OpenAI) at $1.25/1M/1M input against Kimi K3's $3.00/1M/1M — roughly 38% less per token all in. The strongest choice for serious software engineering work, provided you can absorb the output-side pricing. Compare it first if Kimi K3's pricing is the thing stopping you.
Claude Fable 5 — deliberate against Kimi K3's deliberate, with 1M tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.