Grok 4.5
xAI's first coding- and agent-focused model — the first full-scale deployment of the 1.5T-parameter V9 MoE base, trained with real developer-session data from Cursor.
Strong coding value with 2M context — an underrated pick at this price.
Coding and research at competitive pricing with maximum context
You need the highest writing quality or the most reliable production-grade output — Claude wins both.
Best when you want near-flagship coding quality with a massive context window at a mid-tier price.
75% SWE-bench score — strong coding performance close to top Claude models
2M token context window at $2/$6 per million tokens
Fast and responsive for exploration and open-ended research loops
Claude Opus 4.6 and Sonnet 4.6 lead on pure coding benchmarks
Less established ecosystem and tooling than OpenAI or Anthropic
What people actually use Grok 4 for.
Early-stage research mapping — exploring a new topic before narrowing down
Analyzing large codebases or datasets within a 2M-token context window
Competitive intelligence and market research with broad, fast synthesis
The nearest models people weigh against it, and what actually separates them.
vs Grok 4.5 — Against Grok 4.5 (xAI), Grok 4 lands within a few percent on price and takes 4x the context. Which one wins depends on whether context depth or latency is your constraint.
vs Grok 4.6 — Against Grok 4.6 (xAI), Grok 4 lands within a few percent on price and takes 4x the context. Which one wins depends on whether context depth or latency is your constraint.
vs GPT-6 Astra — Against GPT-6 Astra (OpenAI), Grok 4 runs about 87% cheaper per token, takes 1.9x the context and answers faster. Take Grok 4 unless you specifically need what GPT-6 Astra does better.
Price History
→0% since May 8
46 data points · tracked daily since May 8, 2026
Coding and research at competitive pricing with maximum context. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
xAI's first coding- and agent-focused model — the first full-scale deployment of the 1.5T-parameter V9 MoE base, trained with real developer-session data from Cursor.
xAI's long-horizon agent model — it finishes agentic tasks in roughly half the turns of its rivals, which makes it cheaper in practice than its per-token price suggests.
OpenAI's September 3, 2026 frontier release — the first GPT-6 model and OpenAI's answer to Claude Fable 5.1 two days earlier. State of the art on computer use (OSWorld 2.0 72.6% in ~47% less time than GPT-5.6 Sol), agentic coding (Terminal-Bench 4.0 57.9%), and frontier math (FrontierMath Tier 4 97.6%). $10/$50 per 1M tokens, 1.05M context, 128K output, knowledge cutoff April 30, 2026.
Grok 4 costs $2 per million input tokens and $6 per million output tokens on the API. A month of 10M input and 2M output tokens runs about $32.00 at list price, before any batch or caching discounts.
Grok 4 is best for coding and research at competitive pricing with maximum context. It is a strong fit when that workflow matters more than the tradeoffs around balanced pricing and fast speed.
You need the highest writing quality or the most reliable production-grade output — Claude wins both.
Grok 4.5 (xAI) at $2.00/1M/1M input against Grok 4's $2.00/1M/1M. Best cost-per-solved-task coding agent — efficiency over ceiling. Compare it first if Grok 4's pricing is the thing stopping you.
Grok 4.6 — fast against Grok 4's fast, with 500k tokens of context. Worth the swap when response time is what your users notice rather than the last few points of reasoning depth.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.