Gemini 3.8 Flash
Google's September 2, 2026 Flash model, positioned as the agentic workhorse of the Gemini 3 family — stronger coding and terminal work than 3.7 Flash at the same $0.75/$3.75 price.
Muse Spark with better tool calling and first-try accuracy, same price.
Coding agents and multimodal workflows that need few turns and clean output
You need published benchmark evidence before choosing a model.
Compare every model's knowledge cutoff, max output, and context window.
Released September 2, 2026. Gateway id meta/muse-spark-1.3. $1.25/$4.25 per 1M. 1,048,576 context and max output; text, image and PDF input. No benchmark table published for this point release. Verified October 10, 2026.
Same $1.25/$4.25 price as the original Muse Spark
1M context with output limits that match it
Meta describes higher first-attempt accuracy and more reliable tool calling than Muse Spark
Meta published no standard benchmark table we could verify for this update
Free in the Meta AI app, but API access runs through third-party hosts
What people actually use Muse Spark 1.3 for.
Coding agents where fewer unnecessary turns cut the real cost
Reading long documents, images and video in one context
Developer tooling that relies on dependable tool calls
The nearest models people weigh against it, and what actually separates them.
vs Gemini 3.8 Flash — Against Gemini 3.8 Flash (Google), Muse Spark 1.3 costs about 18% more per token, takes 1x the context and answers slower. Gemini 3.8 Flash is the one to check first if the price difference matters more than the ceiling.
vs GPT-6 Astra — Against GPT-6 Astra (OpenAI), Muse Spark 1.3 runs about 91% cheaper per token, gives up 1x on context and answers faster. Take Muse Spark 1.3 unless you specifically need what GPT-6 Astra does better.
vs Claude Fable 5.1 — Against Claude Fable 5.1 (Anthropic), Muse Spark 1.3 runs about 91% cheaper per token, takes 1x the context and answers faster. Take Muse Spark 1.3 unless you specifically need what Claude Fable 5.1 does better.
Coding agents and multimodal workflows that need few turns and clean output. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
Google's September 2, 2026 Flash model, positioned as the agentic workhorse of the Gemini 3 family — stronger coding and terminal work than 3.7 Flash at the same $0.75/$3.75 price.
OpenAI's September 3, 2026 frontier release — the first GPT-6 model and OpenAI's answer to Claude Fable 5.1 two days earlier. State of the art on computer use (OSWorld 2.0 72.6% in ~47% less time than GPT-5.6 Sol), agentic coding (Terminal-Bench 4.0 57.9%), and frontier math (FrontierMath Tier 4 97.6%). $10/$50 per 1M tokens, 1.05M context, 128K output, knowledge cutoff April 30, 2026.
Anthropic's September 1, 2026 frontier release and the new capability ceiling for coding, agents, and scientific work. Base pricing is unchanged at $10/$50, but cache reads dropped 75% to $0.25/1M — roughly 25% cheaper on typical workloads and up to 45% cheaper on agentic ones. 1M context, 128K output, adaptive thinking always on.
Pricing moves, ranking shifts, and capability updates.
Meta released Muse Spark 1.3 on September 2, 2026 at the original Muse Spark's $1.25/$4.25 per million tokens. Meta describes higher first-attempt accuracy, more reliable tool calling and native understanding of video, images and documents across a 1M-token context. Meta published no standard benchmark table for this point release, so we have not changed its ranking against other providers' models. Verified October 10, 2026.
View modelMuse Spark 1.3 costs $1.25 per million input tokens and $4.25 per million output tokens on the API. A month of 10M input and 2M output tokens runs about $21.00 at list price, before any batch or caching discounts.
Muse Spark 1.3 has a 1.0M tokens context window, with up to 1.0M tokens of output per response. That is the total of prompt plus response the model can hold in one request.
Muse Spark 1.3 is best for coding agents and multimodal workflows that need few turns and clean output. It is a strong fit when that workflow matters more than the tradeoffs around balanced pricing and balanced speed.
You need published benchmark evidence before choosing a model.
Gemini 3.8 Flash (Google) at $0.75/1M/1M input against Muse Spark 1.3's $1.25/1M/1M — roughly 18% less per token all in. Google's fast agentic workhorse — strong coding at Flash pricing. Compare it first if Muse Spark 1.3's pricing is the thing stopping you.
GPT-6 Astra — deliberate against Muse Spark 1.3's balanced, with 1.1M tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.