UseRightAI
HomeModelsAsk AIComparePricingWhat's New
UseRightAICut through AI hype. Pick what works.

Independent AI model tracker. Live pricing, real benchmarks, zero vendor bias.

X (Twitter)LinkedInUpdatesContact

Compare

Opus 4.8 vs Opus 4.7Fable 5 vs Opus 4.8New AI Models 2026ChatGPT vs ClaudeGPT-4o vs Claude SonnetClaude vs GeminiDeepSeek vs ChatGPTMistral vs ClaudeGemini Flash vs GPT-4o MiniLlama vs ChatGPTAll comparisons →Build your own →

Best For

CodingWritingDevelopersProduct ManagersDesignersSalesBest Cheap AIBest Free AI

Pricing & Data

API Token PricingCost per TaskPrice HistoryBenchmark ScoresPrivacy & SafetySubscription PlansCost CalculatorWhich AI is Cheapest?Cheapest AI APIs

Company

About UseRightAIContactWhat ChangedAll ModelsDisclosuresPrivacy PolicyTerms of Service

© 2026 UseRightAI. Independent · Free forever · Not affiliated with any AI provider.

Affiliate links are clearly labeled. See disclosures.

HomeAI Cost per TaskRun a chatbot for 10,000 chats

Support

What it costs to run a chatbot for 10,000 chats

One eight-turn conversation, priced with the history re-sent on every turn — the way chat actually bills.

Pricing verified: August 2026

The short answer

Between $0.0001 and $0.010 per run, depending on the model. That is a 71× spread for the same job.

Cheapest that fits
Mistral: Mistral Nemo
$0.0001
$1.47 at 10,000 conversations a month
Best value
GPT-5.6 Luna
$0.0011
$11.40 at 10,000 conversations a month
Highest quality
Claude Sonnet 5
$0.010
$105 at 10,000 conversations a month

How this is calculated

A system prompt plus eight turns of conversation, with the full history re-sent each turn, sums to ≈ 6,000 input tokens across the conversation. Eight replies ≈ 900 output tokens.

6,000 input tokens900 output tokens7k tokens context needed

Cost = (input tokens ÷ 1M × input price) + (output tokens ÷ 1M × output price). Prices come from our daily-verified model data. Batch and cached-input discounts are not applied — they only ever make these numbers smaller.

Every model, cheapest first

ModelProviderPer runAt 10,000 conversations a monthContext
Mistral: Mistral NemoMistral$0.0001$1.47131k tokens
Google: Gemma 2 9BGoogle$0.0003$2.618k tokens
OpenAI: GPT-5 NanoOpenAI$0.0003$3.30400k tokens
Meta: Llama 3.2 1B InstructMeta$0.0003$3.4260k tokens
Meta: Llama 3.1 8B InstructMeta$0.0004$3.7216k tokens
OpenAI: GPT-4.1 NanoOpenAI$0.0005$4.801.0M tokens
Mistral: Mistral Small 3.2 24BMistral$0.0006$6.30128k tokens
Mistral: Ministral 3 3B 2512Mistral$0.0007$6.90131k tokens
GPT-4oOpenAI$0.0007$7.20128k tokens
Google: Gemini 2.0 Flash LiteGoogle$0.0007$7.201.0M tokens
OpenAI: gpt-oss-safeguard-20bOpenAI$0.0007$7.20131k tokens
Mistral: Mistral 7B Instruct v0.1needs chunkingMistral$0.0008$8.313k tokens
Llama 4 ScoutMeta$0.0009$8.70512k tokens
Mistral: Devstral Small 1.1Mistral$0.0009$8.70131k tokens
Mistral: Mistral Small CreativeMistral$0.0009$8.7033k tokens
Mistral: Voxtral Small 24B 2507Mistral$0.0009$8.7032k tokens
Google: Gemini 2.0 FlashGoogle$0.0010$9.601.0M tokens
Google: Gemini 2.5 Flash LiteGoogle$0.0010$9.601.0M tokens
Google: Gemini 2.5 Flash Lite Preview 09-2025Google$0.0010$9.601.0M tokens
Meta: Llama 3 8B InstructMeta$0.0010$9.668k tokens
Mistral: Ministral 3 8B 2512Mistral$0.0010$10.35262k tokens
DeepSeek V4-FlashDeepSeek$0.0011$10.921M tokens
GPT-5.6 LunaOpenAI$0.0011$11.401.1M tokens
Gemma 4 26B A4BGoogle$0.0011$11.40262k tokens
Gemma 4 31BGoogle$0.0012$12.00262k tokens
Meta: Llama Guard 4 12BMeta$0.0012$12.42164k tokens
Mistral: Ministral 3 14B 2512Mistral$0.0014$13.80262k tokens
GPT-4o MiniOpenAI$0.0014$14.40128k tokens
Mistral: Mistral Small 4Mistral$0.0014$14.40262k tokens
Mistral: SabaMistral$0.0017$17.4033k tokens
Llama 4 MaverickMeta$0.0019$19.20256k tokens
OpenAI: GPT-4.1 MiniOpenAI$0.0019$19.201.0M tokens
Google: Gemini 2.5 FlashGoogle$0.0020$20.251.0M tokens
OpenAI: GPT-3.5 TurboOpenAI$0.0022$21.7516k tokens
xAI: Grok 3 MinixAI$0.0022$22.50131k tokens
xAI: Grok 3 Mini BetaxAI$0.0022$22.50131k tokens
Meta: Llama 3.2 11B Vision InstructMeta$0.0024$23.80131k tokens
xAI: Grok Code Fast 1xAI$0.0026$25.50256k tokens
Mistral Small 3.1Mistral$0.0026$26.05128k tokens
Mistral: Mistral Small 3Mistral$0.0026$26.0533k tokens
Mistral: Codestral 2508Mistral$0.0026$26.10256k tokens
DeepSeek V3DeepSeek$0.0026$26.10128k tokens
Anthropic: Claude 3 HaikuAnthropic$0.0026$26.25200k tokens
Meta: Llama 3.1 70B InstructMeta$0.0028$27.60131k tokens
Google: Gemini 3 Flash PreviewGoogle$0.0029$28.501.0M tokens
Llama Guard 3 8BMeta$0.0029$29.07131k tokens
OpenAI: GPT-5 MiniOpenAI$0.0033$33.00400k tokens
OpenAI: GPT-5.1-Codex-MiniOpenAI$0.0033$33.00400k tokens
DeepSeek V4-ProDeepSeek$0.0034$33.931M tokens
Meta: Llama 3 70B InstructMeta$0.0037$37.268k tokens
Mistral: Mixtral 8x7B InstructMistral$0.0037$37.2633k tokens
Gemini 3.6 FlashGoogle$0.0039$39.381.0M tokens
Gemini 3.5 Flash-LiteGoogle$0.0040$40.501.0M tokens
Google: Nano Banana (Gemini 2.5 Flash Image)Google$0.0040$40.5033k tokens
Mistral: Devstral 2 2512Mistral$0.0042$42.00262k tokens
Mistral: Devstral MediumMistral$0.0042$42.00131k tokens
Mistral: Mistral Medium 3Mistral$0.0042$42.00131k tokens
Mistral: Mistral Medium 3.1Mistral$0.0042$42.00131k tokens
Mistral: Mistral Large 3 2512Mistral$0.0043$43.50262k tokens
Google: Gemma 2 27BGoogle$0.0045$44.858k tokens
Anthropic: Claude Haiku 4.5Anthropic$0.0052$52.50200k tokens
DeepSeek R1DeepSeek$0.0053$52.71128k tokens
OpenAI: o3 MiniOpenAI$0.0053$52.80200k tokens
OpenAI: o3 Mini HighOpenAI$0.0053$52.80200k tokens
OpenAI: o4 MiniOpenAI$0.0053$52.80200k tokens
OpenAI: o4 Mini HighOpenAI$0.0053$52.80200k tokens
Gemini 3.1 FlashGoogle$0.0057$57.001M tokens
OpenAI: GPT Audio MiniOpenAI$0.0058$57.60128k tokens
OpenAI: GPT-3.5 Turbo (older v0613)needs chunkingOpenAI$0.0078$78.004k tokens
Codestral 25.01Mistral$0.0078$78.30256k tokens
Google: Gemini 2.5 ProGoogle$0.0083$82.501.0M tokens
OpenAI: GPT-5 CodexOpenAI$0.0083$82.50400k tokens
Anthropic: Claude 3.5 HaikuAnthropic$0.0084$84.00200k tokens
Claude 4 HaikuAnthropic$0.0084$84.00200k tokens
Gemini 3.5 FlashGoogle$0.0086$85.501.0M tokens
Kimi K2.7 CodeMoonshot$0.0093$93.00256k tokens
OpenAI: GPT-4.1OpenAI$0.0096$96.001.0M tokens
OpenAI: o3OpenAI$0.0096$96.00200k tokens
Claude Sonnet 5Anthropic$0.010$1051M tokens
GPT-5.6 SolOpenAI$0.010$1051.1M tokens
OpenAI: GPT-3.5 Turbo Instructneeds chunkingOpenAI$0.011$1084k tokens
Muse SparkMeta$0.011$1131.0M tokens
GPT-5.6 TerraOpenAI$0.011$1141.1M tokens
OpenAI: GPT-5OpenAI$0.011$114400k tokens
GPT-5.2 MiniOpenAI$0.012$115128k tokens
GLM-5.2Z.ai$0.012$1241M tokens
Anthropic: Claude Sonnet 4.5Anthropic$0.016$1581M tokens
Claude Sonnet 4.6Anthropic$0.016$1581M tokens
Mistral Medium 3.5Mistral$0.016$158256k tokens
Google: Gemini 2.5 Pro Preview 05-06Google$0.017$1651.0M tokens
Google: Gemini 2.5 Pro Preview 06-05Google$0.017$1651.0M tokens
OpenAI: GPT-5 ChatOpenAI$0.017$165128k tokens
OpenAI: GPT-5.1OpenAI$0.017$165400k tokens
OpenAI: GPT-5.1 ChatOpenAI$0.017$165128k tokens
OpenAI: GPT-5.1-CodexOpenAI$0.017$165400k tokens
OpenAI: GPT-5.1-Codex-MaxOpenAI$0.017$165400k tokens
OpenAI: GPT-5 Image MiniOpenAI$0.017$168400k tokens
Grok 4xAI$0.017$1742M tokens
Grok 4.5xAI$0.017$174500k tokens
Mistral Large 2Mistral$0.017$174128k tokens
Mistral: Mixtral 8x22B InstructMistral$0.017$17466k tokens
Mistral: Pixtral Large 2411Mistral$0.017$174131k tokens
Qwen 3.8 MaxAlibaba$0.017$1741M tokens
OpenAI: o4 Mini Deep ResearchOpenAI$0.019$192200k tokens
OpenAI: GPT-3.5 Turbo 16kOpenAI$0.022$21616k tokens
Qwen 3.7 MaxAlibaba$0.022$2171M tokens
Gemini 3.1 ProGoogle$0.023$2282M tokens
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)Google$0.023$22866k tokens
OpenAI: GPT-5.3 ChatOpenAI$0.023$231128k tokens
OpenAI: GPT-5.3-CodexOpenAI$0.023$231400k tokens
OpenAI: GPT AudioOpenAI$0.024$240128k tokens
Claude Opus 5Anthropic$0.026$2631M tokens
GPT-5.4OpenAI$0.028$285272k tokens
Anthropic: Claude 3.7 Sonnet (thinking)Anthropic$0.032$315200k tokens
Anthropic: Claude Sonnet 4Anthropic$0.032$315200k tokens
Kimi K3Moonshot$0.032$3151M tokens
xAI: Grok 3xAI$0.032$315131k tokens
xAI: Grok 3 BetaxAI$0.032$315131k tokens
OpenAI: GPT-4 TurboOpenAI$0.043$435128k tokens
Anthropic: Claude Opus 4.5Anthropic$0.052$525200k tokens
Claude Fable 5Anthropic$0.052$5251M tokens
Claude Opus 4.7Anthropic$0.052$5251M tokens
Claude Opus 4.8Anthropic$0.052$5251M tokens
GPT-5.5OpenAI$0.057$5701M tokens
Anthropic: Claude 3.5 SonnetAnthropic$0.063$630200k tokens
OpenAI: GPT-5 ImageOpenAI$0.069$690400k tokens
OpenAI: o1OpenAI$0.072$720200k tokens
OpenAI: GPT-4 Turbo (older v1106)OpenAI$0.087$870128k tokens
OpenAI: GPT-4 Turbo PreviewOpenAI$0.087$870128k tokens
OpenAI: o3 Deep ResearchOpenAI$0.096$960200k tokens
OpenAI: o3 ProOpenAI$0.096$960200k tokens
OpenAI: GPT-5 ProOpenAI$0.099$990400k tokens
Claude Mythos 5Anthropic$0.105$1,0501M tokens
GPT-5.2OpenAI$0.106$1,062200k tokens
Anthropic: Claude Opus 4Anthropic$0.158$1,575200k tokens
Anthropic: Claude Opus 4.1Anthropic$0.158$1,575200k tokens
Claude Opus 4.6Anthropic$0.158$1,5751M tokens
OpenAI: GPT-4OpenAI$0.234$2,3408k tokens
OpenAI: GPT-4 (older v0314)OpenAI$0.234$2,3408k tokens
OpenAI: o1-proOpenAI$0.720$7,200200k tokens

“Needs chunking” means the model's context window can't hold this job in a single pass, so the real cost is higher than the figure shown and quality usually suffers.

Should you just pick the cheapest?

Chat bills quadratically, not linearly: every turn re-sends the whole history, so a sixteen-turn conversation costs roughly four times an eight-turn one rather than twice. Prompt caching targets exactly this and can cut the input side by up to 90% on a stable system prompt.

Price your own token counts →best ai chatbot →best cheap api →best ai for customer support →ai api cost calculator →

Frequently asked

How much does it cost to run an AI chatbot?

$0.0001 on Mistral: Mistral Nemo, the cheapest capable option, rising to $0.010 on Claude Sonnet 5 at the top end. GPT-5.6 Luna is the value pick at $0.0011 per run. The job is priced at 6,000 input and 900 output tokens — see the working below.

How did you work out the token count for this task?

A system prompt plus eight turns of conversation, with the full history re-sent each turn, sums to ≈ 6,000 input tokens across the conversation. Eight replies ≈ 900 output tokens.

What does this cost at 10,000 conversations a month?

$1.47 a month on Mistral: Mistral Nemo, $11.40 on GPT-5.6 Luna, and $105 on Claude Sonnet 5. Batch APIs typically halve these figures for work that can wait, and prompt caching cuts the input side further when the same context is reused.

Is the cheapest model the right choice for this task?

Chat bills quadratically, not linearly: every turn re-sends the whole history, so a sixteen-turn conversation costs roughly four times an eight-turn one rather than twice. Prompt caching targets exactly this and can cut the input side by up to 90% on a stable system prompt.

Are these prices current?

Yes. Every figure on this page is computed from our model pricing data, which is checked daily against each provider's official pricing page. When a provider changes a price, these numbers change with it.

What other AI tasks cost

Documents
Summarize a 100-page PDF
Documents
Summarize a 400-page book
Documents
Translate a 50-page document
Writing
Write a 1,500-word blog post
Writing
Write a tailored cover letter
Writing
Write 1,000 product descriptions

Newsletter

Get model updates before your workflow falls behind

Pricing changes, new model releases, and updated recommendations — delivered when it matters.

No spam. Useful updates only. Affiliate disclosures always clearly labeled.