Claude Fable 5
Anthropic's new Mythos-class flagship and the most capable coding model anyone can use — 80.3% SWE-Bench Pro, an 11-point jump over Opus 4.8. 1M context, 128K output, native parallel subagents. Released June 9, 2026.
Best self-hostable multimodal model — European, dense, MIT-licensed.
Self-hostable European multimodal coding
API price-performance is all that matters — DeepSeek V4-Pro is stronger and cheaper hosted.
Compare every model's knowledge cutoff, max output, and context window.
Released April 29, 2026 (model ID mistral-medium-3-5-26-04). Pricing verified on docs.mistral.ai. Le Chat Pro $14.99/mo does not include API usage.
77.6% SWE-bench Verified — the strongest dense open-weights coding score at release
Single-checkpoint multimodality with function calling and structured outputs
Open weights under a modified MIT license, practical to self-host and fine-tune at 128B dense
Trails GPT-5.6, Opus-class, and Gemini frontier models on complex multi-step reasoning; sparse published benchmark disclosure
256K context is a quarter of the 1M frontier norm, and $7.50/1M output is dear for the tier
What people actually use Mistral Medium 3.5 for.
Coding at 77.6% SWE-bench Verified — within ~2 points of Claude Sonnet 4.6 at roughly half the price
Document Q&A and vision tasks from a single checkpoint with structured outputs
EU-compliant self-hosted deployments — 128B dense is far easier to run than trillion-parameter MoE rivals
The nearest models people weigh against it, and what actually separates them.
vs Claude Fable 5 — Against Claude Fable 5 (Anthropic), Mistral Medium 3.5 runs about 85% cheaper per token, gives up 3.9x on context and answers faster. Take Mistral Medium 3.5 unless you specifically need what Claude Fable 5 does better.
vs Claude Fable 5.1 — Against Claude Fable 5.1 (Anthropic), Mistral Medium 3.5 runs about 85% cheaper per token, gives up 3.9x on context and answers faster. Take Mistral Medium 3.5 unless you specifically need what Claude Fable 5.1 does better.
vs Claude Opus 4.7 — Against Claude Opus 4.7 (Anthropic), Mistral Medium 3.5 runs about 70% cheaper per token, gives up 3.9x on context and answers faster. Take Mistral Medium 3.5 unless you specifically need what Claude Opus 4.7 does better.
Price History
→0% since Aug 7
37 data points · tracked daily since Aug 7, 2026
Self-hostable European multimodal coding. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
Anthropic's new Mythos-class flagship and the most capable coding model anyone can use — 80.3% SWE-Bench Pro, an 11-point jump over Opus 4.8. 1M context, 128K output, native parallel subagents. Released June 9, 2026.
Anthropic's September 1, 2026 frontier release and the new capability ceiling for coding, agents, and scientific work. Base pricing is unchanged at $10/$50, but cache reads dropped 75% to $0.25/1M — roughly 25% cheaper on typical workloads and up to 45% cheaper on agentic ones. 1M context, 128K output, adaptive thinking always on.
Anthropic's previous Opus flagship, now superseded by Opus 4.8. Still the second-best coding model publicly available at the same $5/$25 price.
Mistral Medium 3.5 costs $1.5 per million input tokens and $7.5 per million output tokens on the API. A month of 10M input and 2M output tokens runs about $30.00 at list price, before any batch or caching discounts.
Mistral Medium 3.5 has a 256k tokens context window, with up to 256k tokens of output per response. That is the total of prompt plus response the model can hold in one request.
Mistral Medium 3.5 is best for self-hostable european multimodal coding. It is a strong fit when that workflow matters more than the tradeoffs around balanced pricing and balanced speed.
API price-performance is all that matters — DeepSeek V4-Pro is stronger and cheaper hosted.
GPT-5.1-Codex-Max (OpenAI) at $1.25/1M/1M input against Mistral Medium 3.5's $1.50/1M/1M. The strongest choice for serious software engineering work, provided you can absorb the output-side pricing. Compare it first if Mistral Medium 3.5's pricing is the thing stopping you.
Claude Fable 5 — deliberate against Mistral Medium 3.5's balanced, with 1M tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.