Skip to content

AI Model Pricing Calculator

Calculate daily, monthly and yearly LLM API costs per model from your request volume and average input and output tokens.

AI Model Pricing Calculator

Models

Monthly cost comparison

GPT-6 Luna
$13.50
Qwen3.8-Flash
$16.05
Mistral Small 4
$18.00
GPT-OSS 120B (Groq)
$18.00
DeepSeek V4.1 Flash
$36.00
Mistral Large 3
$52.50
Gemini 3.5 Flash-Lite
$55.50
Gemini 3.8 Flash
$101.25
GPT-5.4 Mini
$112.50
Grok 4.3
$112.50
Claude Haiku 4.5
$135.00
DeepSeek V4 Pro
$138.60
Muse Spark 1.3
$138.75
Mistral Medium 3.5
$202.50
Grok 4.7
$210.00
Qwen3.8-Max
$210.00
Claude Sonnet 5.5
$270.00
GPT-6.1 Sol
$270.00
Gemini 3.1 Pro Preview
$300.00
Claude Opus 5.5
$540.00
GPT-5.5
$750.00
Claude Fable 5.1
$1,350.00
GPT-6 Astra
$1,350.00

Cheapest: GPT-6 Luna at $13.50/mo · Most expensive: GPT-6 Astra at $1,350.00/mo

Cost summary

GPT-6 LunaCheapest$0.450$13.50$164.25
Qwen3.8-Flash$0.535$16.05$195.27
Mistral Small 4$0.600$18.00$219.00
GPT-OSS 120B (Groq)$0.600$18.00$219.00
DeepSeek V4.1 Flash$1.20$36.00$438.00
Mistral Large 3$1.75$52.50$638.75
Gemini 3.5 Flash-Lite$1.85$55.50$675.25
Gemini 3.8 Flash$3.38$101.25$1,231.88
GPT-5.4 Mini$3.75$112.50$1,368.75
Grok 4.3$3.75$112.50$1,368.75
Claude Haiku 4.5$4.50$135.00$1,642.50
DeepSeek V4 Pro$4.62$138.60$1,686.30
Muse Spark 1.3$4.63$138.75$1,688.13
Mistral Medium 3.5$6.75$202.50$2,463.75
Grok 4.7$7.00$210.00$2,555.00
Qwen3.8-Max$7.00$210.00$2,555.00
Claude Sonnet 5.5$9.00$270.00$3,285.00
GPT-6.1 Sol$9.00$270.00$3,285.00
Gemini 3.1 Pro Preview$10.00$300.00$3,650.00
Claude Opus 5.5$18.00$540.00$6,570.00
GPT-5.5$25.00$750.00$9,125.00
Claude Fable 5.1$45.00$1,350.00$16,425.00
GPT-6 AstraPriciest$45.00$1,350.00$16,425.00

Break-even calculator

How much do you charge per API call to your users?

Pricing data last updated: October 3, 2026. All prices per 1M tokens.

Updated . Provided as is. Check the output before you rely on it in production.

How to use AI Model Pricing Calculator

  1. 1

    Enter your token usage

    Input your estimated prompt tokens (input) and completion tokens (output) per request. If you're unsure, use the token counter tool first.

  2. 2

    Select models to compare

    Choose several models (GPT, Claude, Gemini, Grok, DeepSeek and more). The calculator shows the cost per request for each.

  3. 3

    View cost breakdowns

    See daily, monthly and yearly cost for every selected model, and sort the table by any of those columns to rank them.

  4. 4

    Find the revenue that covers the bill

    Enter your revenue per API call in the break-even calculator to see how many calls each model needs to pay for itself, and the monthly profit left at your current volume.

  5. 5

    Export and share cost comparison

    Download the comparison to share with stakeholders and justify the model and provider you choose.

Questions and answers

What is AI Pricing Calculator?
LLM APIs bill per million input and output tokens at a different rate for each model. Enter requests per day and average input and output tokens, and this calculator gives daily, monthly and yearly cost for every selected model, with a break-even view per call and CSV export.
How current are the pricing figures?
Rates are read from each vendor's own pricing page, and the date of the last check is shown above the price table. The table lists input and output prices per million tokens and marks which models support prompt caching; the calculator uses the standard input and output rates.
For AI agents: how to call this tool

Machine-readable contract, endpoints and examples. Humans can ignore this section.

Best Path For Builders

Browser workflow

Runs instantly in the browser with private local processing and copy/export-ready output.

Browser Workflow

This tool is optimized for instant in-browser execution with local data handling. Run it here and copy/export the output directly.

/ai-pricing/

For automation planning, fetch the canonical contract at /api/tool/ai-pricing.json.

AI Model Pricing — October 3, 2026

Current API pricing per 1 million tokens for 66 models across 9 providers. Prices in USD.

Model Provider Input / 1M Output / 1M Caching Tier
Claude Fable 5.1 Anthropic $10.00 $50.00 ✓ premium
Claude Opus 5.5 Anthropic $4.00 $20.00 ✓ premium
Claude Opus 5 Anthropic $5.00 $25.00 ✓ premium
Claude Sonnet 5.5 Anthropic $2.00 $10.00 ✓ mid
Claude Sonnet 5 Anthropic $2.00 $10.00 ✓ mid
Claude Haiku 4.5 Anthropic $1.00 $5.00 ✓ budget
Claude Fable 5 Anthropic $10.00 $50.00 ✓ premium
Claude Opus 4.8 Anthropic $5.00 $25.00 ✓ premium
Claude Opus 4.7 Anthropic $5.00 $25.00 ✓ premium
Claude Opus 4.6 Anthropic $5.00 $25.00 ✓ premium
Claude Opus 4.5 Anthropic $5.00 $25.00 ✓ premium
Claude Sonnet 4.6 Anthropic $3.00 $15.00 ✓ mid
Claude Sonnet 4.5 Anthropic $3.00 $15.00 ✓ mid
GPT-6 Astra OpenAI $10.00 $50.00 ✓ premium
GPT-6.1 Sol OpenAI $2.00 $10.00 ✓ mid
GPT-6 Sol OpenAI $2.00 $10.00 ✓ mid
GPT-6 Luna OpenAI $0.10 $0.50 ✓ budget
GPT-5.6 Sol OpenAI $4.00 $20.00 ✓ premium
GPT-5.6 Terra OpenAI $2.00 $12.00 ✓ mid
GPT-5.6 Luna OpenAI $0.20 $1.20 ✓ budget
GPT-5.5 OpenAI $5.00 $30.00 ✓ premium
GPT-5.5 Pro OpenAI $30.00 $180.00 — premium
GPT-5.4 OpenAI $2.50 $15.00 ✓ premium
GPT-5.4 Pro OpenAI $30.00 $180.00 — premium
GPT-5.4 Mini OpenAI $0.75 $4.50 ✓ mid
GPT-5.4 Nano OpenAI $0.20 $1.25 ✓ budget
GPT-5.3-Codex OpenAI $1.75 $14.00 ✓ premium
GPT-5.2 OpenAI $1.75 $14.00 ✓ mid
GPT-5.2 Pro OpenAI $21.00 $168.00 — premium
GPT-5.1 OpenAI $1.25 $10.00 ✓ mid
GPT-5 OpenAI $1.25 $10.00 ✓ mid
GPT-5-mini OpenAI $0.25 $2.00 ✓ budget
GPT-5-nano OpenAI $0.050 $0.40 ✓ budget
o3 OpenAI $2.00 $8.00 ✓ mid
o3-pro OpenAI $20.00 $80.00 — premium
o4-mini OpenAI $1.10 $4.40 ✓ mid
o3-mini OpenAI $1.10 $4.40 ✓ mid
GPT-4.1 OpenAI $2.00 $8.00 ✓ mid
GPT-4.1-mini OpenAI $0.40 $1.60 ✓ budget
GPT-4.1-nano OpenAI $0.10 $0.40 ✓ budget
Gemini 3.8 Flash Google $0.75 $3.75 ✓ mid
Gemini 3.7 Flash Google $0.75 $3.75 ✓ mid
Gemini 3.6 Flash Google $0.75 $3.75 ✓ mid
Gemini 3.5 Flash Google $1.50 $9.00 ✓ mid
Gemini 3.5 Flash-Lite Google $0.30 $2.50 ✓ budget
Gemini 3.1 Pro Preview Google $2.00 $12.00 ✓ premium
Gemini 2.5 Pro Google $1.25 $10.00 ✓ premium
Gemini 2.5 Flash Google $0.30 $2.50 ✓ mid
Gemini 2.5 Flash-Lite Google $0.10 $0.40 ✓ budget
DeepSeek V4.1 Flash DeepSeek $0.30 $1.20 ✓ budget
DeepSeek V4 Pro DeepSeek $1.32 $3.96 ✓ mid
Grok 4.7 xAI $2.00 $6.00 ✓ premium
Grok 4.6 xAI $2.00 $6.00 ✓ premium
Grok 4.5 xAI $2.00 $6.00 ✓ premium
Grok 4.3 xAI $1.25 $2.50 ✓ mid
Grok 4.20 xAI $1.25 $2.50 ✓ mid
Grok Build 0.1 xAI $1.00 $2.00 ✓ mid
Mistral Medium 3.5 Mistral $1.50 $7.50 — premium
Mistral Small 4 Mistral $0.15 $0.60 — budget
Mistral Large 3 Mistral $0.50 $1.50 — mid
Codestral 25.08 Mistral $0.30 $0.90 — budget
Muse Spark 1.3 Meta $1.25 $4.25 ✓ mid
Qwen3.8-Max Qwen $2.00 $6.00 ✓ premium
Qwen3.8-Flash Qwen $0.15 $0.47 ✓ budget
GPT-OSS 120B (Groq) Groq $0.15 $0.60 — budget
GPT-OSS 20B (Groq) Groq $0.075 $0.30 — budget

OpenAI — Includes flagship reasoning/coding models plus mini and nano budget tiers for high-volume workloads.

Anthropic — Claude family offers premium, balanced, and budget tiers, with strong prompt-caching support.

Google — Gemini family spans high-capability Pro options and low-cost Flash variants.

DeepSeek — Often among the most cost-efficient options for bulk inference and large-scale processing.

xAI — Provides flagship and fast-cost tiers, including high-context-window variants.

Meta — Muse Spark carries a 1M-token context window at a mid-tier price, with a cheaper Contributor variant that lets Meta train on your data.

Qwen — Alibaba Cloud's Max and Flash models both take 1M-token contexts; Flash is among the cheapest current models.

Mistral — Mistral Large 3 is open-weight under Apache 2.0, so it can also be self-hosted.