UseRightAI
HomeModelsAsk AIComparePricingWhat's New
UseRightAICut through AI hype. Pick what works.

Independent AI model tracker. Live pricing, real benchmarks, zero vendor bias.

X (Twitter)LinkedInUpdatesContact

Compare

Opus 4.8 vs Opus 4.7Fable 5 vs Opus 4.8New AI Models 2026ChatGPT vs ClaudeGPT-4o vs Claude SonnetClaude vs GeminiDeepSeek vs ChatGPTMistral vs ClaudeGemini Flash vs GPT-4o MiniLlama vs ChatGPTAll comparisons →Build your own →

Best For

CodingWritingDevelopersProduct ManagersDesignersSalesBest Cheap AIBest Free AI

Pricing & Data

API Token PricingCost per TaskPrice HistoryBenchmark ScoresPrivacy & SafetySubscription PlansPlan Usage LimitsCost CalculatorWhich AI is Cheapest?Cheapest AI APIs

Company

About UseRightAIContactWhat ChangedAll ModelsEditorial PolicyDisclosuresPrivacy PolicyTerms of Service

© 2026 UseRightAI. Independent · Free forever · Not affiliated with any AI provider.

Affiliate links are clearly labeled. See disclosures.

HomeAI Cost per TaskRun a chatbot for 10,000 chats

Support

What it costs to run a chatbot for 10,000 chats

One eight-turn conversation, priced with the history re-sent on every turn — the way chat actually bills.

Pricing verified: September 2026

The short answer

Between $0.0003 and $0.021 per run, depending on the model. That is a 80× spread for the same job.

Cheapest that fits
Gemma 2 9B
$0.0003
$2.61 at 10,000 conversations a month
Best value
GPT-5.6 Luna
$0.0023
$22.80 at 10,000 conversations a month
Highest quality
Claude Sonnet 5
$0.021
$210 at 10,000 conversations a month

How this is calculated

A system prompt plus eight turns of conversation, with the full history re-sent each turn, sums to ≈ 6,000 input tokens across the conversation. Eight replies ≈ 900 output tokens.

6,000 input tokens900 output tokens7k tokens context needed

Cost = (input tokens ÷ 1M × input price) + (output tokens ÷ 1M × output price). Prices come from our daily-verified model data. Batch and cached-input discounts are not applied — they only ever make these numbers smaller.

Every model, cheapest first

ModelProviderPer runAt 10,000 conversations a monthContext
Gemma 2 9BGoogle$0.0003$2.618k tokens
Llama 3.2 1B InstructMeta$0.0003$3.4260k tokens
Llama 3.1 8B InstructMeta$0.0004$3.7216k tokens
gpt-oss-safeguard-20bOpenAI$0.0006$6.00131k tokens
Mistral Small 3.2 24BMistral$0.0006$6.30128k tokens
GPT-5 NanoOpenAI$0.0007$6.60400k tokens
Ministral 3 3B 2512Mistral$0.0007$6.90131k tokens
Gemini 2.0 Flash LiteGoogle$0.0007$7.201.0M tokens
Mistral 7B Instruct v0.1needs chunkingMistral$0.0008$8.313k tokens
Devstral Small 1.1Mistral$0.0009$8.70131k tokens
Mistral Small CreativeMistral$0.0009$8.7033k tokens
Voxtral Small 24B 2507Mistral$0.0009$8.7032k tokens
Mistral Small 3.1Mistral$0.0009$8.70128k tokens
Gemini 2.0 FlashGoogle$0.0010$9.601.0M tokens
Gemini 2.5 Flash LiteGoogle$0.0010$9.601.0M tokens
Gemini 2.5 Flash Lite Preview 09-2025Google$0.0010$9.601.0M tokens
GPT-4.1 NanoOpenAI$0.0010$9.601.0M tokens
Llama 3 8B InstructMeta$0.0010$9.668k tokens
Ministral 3 8B 2512Mistral$0.0010$10.35262k tokens
Mistral NemoMistral$0.0010$10.35131k tokens
DeepSeek V4-FlashDeepSeek$0.0011$10.921M tokens
Gemma 4 26B A4BGoogle$0.0011$11.40262k tokens
Llama Guard 4 12BMeta$0.0012$12.42164k tokens
GLM-5.3 FlashZ.ai$0.0014$13.501M tokens
Ministral 3 14B 2512Mistral$0.0014$13.80262k tokens
Qwen 3.8 FlashAlibaba$0.0014$13.83991k tokens
GPT-4o MiniOpenAI$0.0014$14.40128k tokens
Mistral Small 4Mistral$0.0014$14.40262k tokens
SabaMistral$0.0017$17.4033k tokens
Grok 3 MinixAI$0.0022$22.50131k tokens
Grok 3 Mini BetaxAI$0.0022$22.50131k tokens
GPT-5.6 LunaOpenAI$0.0023$22.801.1M tokens
Llama 3.2 11B Vision InstructMeta$0.0024$23.80131k tokens
Grok Code Fast 1xAI$0.0026$25.50256k tokens
Mistral Small 3Mistral$0.0026$26.0533k tokens
Codestral 2508Mistral$0.0026$26.10256k tokens
DeepSeek V3DeepSeek$0.0026$26.10128k tokens
Claude 3 HaikuAnthropic$0.0026$26.25200k tokens
Llama 3.1 70B InstructMeta$0.0028$27.60131k tokens
Gemini 3 Flash PreviewGoogle$0.0029$28.501.0M tokens
Llama Guard 3 8BMeta$0.0029$29.07131k tokens
Gemma 4 31BGoogle$0.0032$32.13262k tokens
GPT-5 MiniOpenAI$0.0033$33.00400k tokens
GPT-5.1-Codex-MiniOpenAI$0.0033$33.00400k tokens
DeepSeek V4-ProDeepSeek$0.0034$33.931M tokens
Muse Glimmer 30BMeta$0.0034$34.50131k tokens
Llama 3 70B InstructMeta$0.0037$37.268k tokens
Mixtral 8x7B InstructMistral$0.0037$37.2633k tokens
GPT-4.1 MiniOpenAI$0.0038$38.401.0M tokens
Gemini 2.5 FlashGoogle$0.0040$40.501.0M tokens
Gemini 3.5 Flash-LiteGoogle$0.0040$40.501.0M tokens
Nano Banana (Gemini 2.5 Flash Image)Google$0.0040$40.5033k tokens
Llama 4 ScoutMeta$0.0041$40.80512k tokens
Devstral 2 2512Mistral$0.0042$42.00262k tokens
Devstral MediumMistral$0.0042$42.00131k tokens
Mistral Medium 3Mistral$0.0042$42.00131k tokens
Mistral Medium 3.1Mistral$0.0042$42.00131k tokens
GPT-3.5 TurboOpenAI$0.0043$43.5016k tokens
Mistral Large 3 2512Mistral$0.0043$43.50262k tokens
Gemma 2 27BGoogle$0.0045$44.858k tokens
Llama 4 MaverickMeta$0.0050$50.40256k tokens
DeepSeek R1DeepSeek$0.0053$52.71128k tokens
Gemini 3.1 FlashGoogle$0.0057$57.001M tokens
GPT Audio MiniOpenAI$0.0058$57.60128k tokens
GPT-3.5 Turbo (older v0613)needs chunkingOpenAI$0.0078$78.004k tokens
Codestral 25.01Mistral$0.0078$78.30256k tokens
Gemini 3.6 FlashGoogle$0.0079$78.751.0M tokens
Gemini 3.7 FlashGoogle$0.0079$78.751.0M tokens
Claude 3.5 HaikuAnthropic$0.0084$84.00200k tokens
Claude 4 HaikuAnthropic$0.0084$84.00200k tokens
Kimi K2.7 CodeMoonshot$0.0093$93.00256k tokens
Claude Haiku 4.5Anthropic$0.010$105200k tokens
o3 MiniOpenAI$0.011$106200k tokens
o3 Mini HighOpenAI$0.011$106200k tokens
o4 MiniOpenAI$0.011$106200k tokens
o4 Mini HighOpenAI$0.011$106200k tokens
GPT-3.5 Turbo Instructneeds chunkingOpenAI$0.011$1084k tokens
Muse SparkMeta$0.011$1131.0M tokens
GPT-5.2 MiniOpenAI$0.012$115128k tokens
GLM-5.2Z.ai$0.012$1241M tokens
GLM-5.3Z.ai$0.012$1241M tokens
Mistral Medium 3.5Mistral$0.016$158256k tokens
GPT-5OpenAI$0.017$165400k tokens
GPT-5 ChatOpenAI$0.017$165128k tokens
GPT-5 CodexOpenAI$0.017$165400k tokens
GPT-5.1OpenAI$0.017$165400k tokens
GPT-5.1 ChatOpenAI$0.017$165128k tokens
GPT-5.1-CodexOpenAI$0.017$165400k tokens
GPT-5.1-Codex-MaxOpenAI$0.017$165400k tokens
Gemini 2.5 ProGoogle$0.017$1651.0M tokens
Gemini 2.5 Pro Preview 05-06Google$0.017$1651.0M tokens
Gemini 2.5 Pro Preview 06-05Google$0.017$1651.0M tokens
GPT-5 Image MiniOpenAI$0.017$168400k tokens
Gemini 3.5 FlashGoogle$0.017$1711.0M tokens
Grok 4xAI$0.017$1742M tokens
Grok 4.5xAI$0.017$174500k tokens
Grok 4.6xAI$0.017$174500k tokens
Mixtral 8x22B InstructMistral$0.017$17466k tokens
Pixtral Large 2411Mistral$0.017$174131k tokens
Qwen 3.8 MaxAlibaba$0.017$1741M tokens
GPT-4.1OpenAI$0.019$1921.0M tokens
o3OpenAI$0.019$192200k tokens
o4 Mini Deep ResearchOpenAI$0.019$192200k tokens
Claude Sonnet 5Anthropic$0.021$2101M tokens
GPT-5.6 SolOpenAI$0.021$2101.1M tokens
GPT-3.5 Turbo 16kOpenAI$0.022$21616k tokens
Qwen 3.7 MaxAlibaba$0.022$2171M tokens
GPT-5.6 TerraOpenAI$0.023$2281.1M tokens
Gemini 3.1 ProGoogle$0.023$2282M tokens
Nano Banana Pro (Gemini 3 Pro Image Preview)Google$0.023$22866k tokens
GPT-5.2OpenAI$0.023$231200k tokens
GPT-5.3 ChatOpenAI$0.023$231128k tokens
GPT-5.3-CodexOpenAI$0.023$231400k tokens
GPT AudioOpenAI$0.024$240128k tokens
GPT-4oOpenAI$0.024$240128k tokens
Mistral Large 2Mistral$0.026$261128k tokens
GPT-5.4OpenAI$0.028$285272k tokens
Claude 3.7 Sonnet (thinking)Anthropic$0.032$315200k tokens
Claude Sonnet 4Anthropic$0.032$315200k tokens
Claude Sonnet 4.5Anthropic$0.032$3151M tokens
Claude Sonnet 4.6Anthropic$0.032$3151M tokens
Grok 3xAI$0.032$315131k tokens
Grok 3 BetaxAI$0.032$315131k tokens
Kimi K3Moonshot$0.032$3151M tokens
Claude Opus 4.5Anthropic$0.052$525200k tokens
Claude Opus 4.6Anthropic$0.052$5251M tokens
Claude Opus 4.7Anthropic$0.052$5251M tokens
Claude Opus 4.8Anthropic$0.052$5251M tokens
Claude Opus 5Anthropic$0.052$5251M tokens
GPT-5.5OpenAI$0.057$5701M tokens
Claude 3.5 SonnetAnthropic$0.063$630200k tokens
GPT-5 ImageOpenAI$0.069$690400k tokens
GPT-4 TurboOpenAI$0.087$870128k tokens
GPT-4 Turbo (older v1106)OpenAI$0.087$870128k tokens
GPT-4 Turbo PreviewOpenAI$0.087$870128k tokens
o3 Deep ResearchOpenAI$0.096$960200k tokens
Claude Fable 5Anthropic$0.105$1,0501M tokens
Claude Fable 5.1Anthropic$0.105$1,0501M tokens
Claude Mythos 5Anthropic$0.105$1,0501M tokens
Claude Mythos 5.1Anthropic$0.105$1,0501M tokens
GPT-6 AstraOpenAI$0.105$1,0501.1M tokens
o1OpenAI$0.144$1,440200k tokens
Claude Opus 4Anthropic$0.158$1,575200k tokens
Claude Opus 4.1Anthropic$0.158$1,575200k tokens
o3 ProOpenAI$0.192$1,920200k tokens
GPT-5 ProOpenAI$0.198$1,980400k tokens
GPT-4OpenAI$0.234$2,3408k tokens
GPT-4 (older v0314)OpenAI$0.234$2,3408k tokens
o1-proOpenAI$1.44$14,400200k tokens

“Needs chunking” means the model's context window can't hold this job in a single pass, so the real cost is higher than the figure shown and quality usually suffers.

Should you just pick the cheapest?

Chat bills quadratically, not linearly: every turn re-sends the whole history, so a sixteen-turn conversation costs roughly four times an eight-turn one rather than twice. Prompt caching targets exactly this and can cut the input side by up to 90% on a stable system prompt.

Price your own token counts →best ai chatbot →best cheap api →best ai for customer support →ai api cost calculator →

Frequently asked

How much does it cost to run an AI chatbot?

$0.0003 on Gemma 2 9B, the cheapest capable option, rising to $0.021 on Claude Sonnet 5 at the top end. GPT-5.6 Luna is the value pick at $0.0023 per run. The job is priced at 6,000 input and 900 output tokens — see the working below.

How did you work out the token count for this task?

A system prompt plus eight turns of conversation, with the full history re-sent each turn, sums to ≈ 6,000 input tokens across the conversation. Eight replies ≈ 900 output tokens.

What does this cost at 10,000 conversations a month?

$2.61 a month on Gemma 2 9B, $22.80 on GPT-5.6 Luna, and $210 on Claude Sonnet 5. Batch APIs typically halve these figures for work that can wait, and prompt caching cuts the input side further when the same context is reused.

Is the cheapest model the right choice for this task?

Chat bills quadratically, not linearly: every turn re-sends the whole history, so a sixteen-turn conversation costs roughly four times an eight-turn one rather than twice. Prompt caching targets exactly this and can cut the input side by up to 90% on a stable system prompt.

Are these prices current?

Yes. Every figure on this page is computed from our model pricing data, which is checked daily against each provider's official pricing page. When a provider changes a price, these numbers change with it.

What other AI tasks cost

Documents
Summarize a 100-page PDF
Documents
Summarize a 400-page book
Documents
Translate a 50-page document
Writing
Write a 1,500-word blog post
Writing
Write a tailored cover letter
Writing
Write 1,000 product descriptions

Newsletter

Get model updates before your workflow falls behind

Pricing changes, new model releases, and updated recommendations — delivered when it matters.

No spam. Useful updates only. Affiliate disclosures always clearly labeled.