AI Model Cost Calculator

Enter your monthly token volume and see the estimated cost across every major AI model, ranked cheapest to most expensive.

AI model pricing is split into an input price and an output price per million tokens, and the gap between the cheapest and most expensive models is often 100x or more. This tool multiplies your monthly token volume against every model in our pricing table so you can see the real cost gap for your own workload, not just a generic per-token number.

Estimate Your Monthly Cost

Not sure about your volume? A busy support agent replying to ~2,000 tickets a month with a 1,000-token context each is roughly 2M input / 1M output tokens — scale from there.

Cheapest at this volumeGLM 4.7 Flash$0/mo
Most expensive at this volumeGPT-5.4 Pro$3,300/mo
RankModelVendorEst. monthly cost
1GLM 4.7 FlashZhipu AI$0
2Nova MicroAmazon$3.15
3Nova LiteAmazon$5.40
4Ministral 3 (3B)Mistral$6.00
5GPT-5 NanoOpenAI$6.50
6GLM 4.7 FlashXZhipu AI$7.50
7Devstral Small 2Mistral$8.00
8Gemini 2.5 Flash-LiteGoogle$9.00
9Ministral 3 (8B)Mistral$9.00
10Mistral NeMoMistral$9.00
11DeepSeek V4 FlashDeepSeek$9.80
12Ministral 3 (14B)Mistral$12.00
13Mistral Small 4Mistral$13.50
14Llama 4 ScoutMeta$15.10
15Qwen3.6 FlashAlibaba$20.80
16Command LightCohere$21.00
17GLM 4.5 AirZhipu AI$21.00
18Llama 4 MaverickMeta$21.70
19GPT-5.6 LunaOpenAI$22.00
20GPT-5.4 NanoOpenAI$22.50
21CodestralMistral$24.00
22MiniMax M3MiniMax$27.00
23MiniMax M2.7MiniMax$27.00
24MiniMax M2.5MiniMax$27.00
25MiniMax M2.1MiniMax$27.00
26MiniMax M2MiniMax$27.00
27Gemini 3.1 Flash-LiteGoogle$27.50
28DeepSeek V4 ProDeepSeek$30.45
29GPT-5 MiniOpenAI$32.50
30Qwen3.6 PlusAlibaba$35.75
31Gemini 3.5 Flash-LiteGoogle$40.00
32Gemini 2.5 FlashGoogle$40.00
33Mistral Large 3Mistral$40.00
34Magistral SmallMistral$40.00
35Devstral 2Mistral$40.00
36Command RCohere$40.00
37Mixtral 8x7BMistral$42.00
38GLM 4.7Zhipu AI$52.00
39GLM 4.6Zhipu AI$52.00
40SonarPerplexity$60.00
41Grok BuildxAI$70.00
42CommandCohere$70.00
43Nova ProAmazon$72.00
44GLM 5Zhipu AI$82.00
45GPT-5.4 MiniOpenAI$82.50
46Gemini 3.1 Flash Live PreviewGoogle$82.50
47Grok 4.3xAI$87.50
48Grok 4.20xAI$87.50
49Kimi K2.7 CodeMoonshot AI$87.50
50Kimi K2.6Moonshot AI$87.50
51Claude Haiku 4.5Anthropic$100
52GLM 5 TurboZhipu AI$100
53GLM 5.2Zhipu AI$114
54GLM 5.1Zhipu AI$114
55Gemini 3.6 FlashGoogle$150
56Mistral Medium 3.5Mistral$150
57Magistral MediumMistral$150
58Grok 4.5xAI$160
59Mixtral 8x22BMistral$160
60GPT-5OpenAI$163
61Gemini 2.5 ProGoogle$163
62Gemini 3.5 FlashGoogle$165
63Sonar Reasoning ProPerplexity$180
64Sonar Deep ResearchPerplexity$180
65Claude Sonnet 5Anthropic$200
66GPT-5.6 TerraOpenAI$220
67Gemini 3.1 ProGoogle$220
68Command R+Cohere$225
69GPT-5.3-CodexOpenAI$228
70GPT-5.2 ChatOpenAI$228
71Nova PremierAmazon$250
72GPT-5.4OpenAI$275
73Claude Sonnet 4.6Anthropic$300
74Claude Sonnet 4.5Anthropic$300
75Sonar ProPerplexity$300
76Command R+ (04-2024)Cohere$300
77Kimi K3Moonshot AI$300
78Claude Opus 5Anthropic$500
79Claude Opus 4.8Anthropic$500
80Claude Opus 4.7Anthropic$500
81Claude Opus 4.6Anthropic$500
82Claude Opus 4.5Anthropic$500
83GPT-5.6 SolOpenAI$550
84GPT-5.5OpenAI$550
85Claude Fable 5.1Anthropic$1,000
86Claude Mythos 5.1Anthropic$1,000
87Claude Fable 5Anthropic$1,000
88Claude Mythos 5Anthropic$1,000
89GPT-5.5 ProOpenAI$3,300
90GPT-5.4 ProOpenAI$3,300

Estimate only — treats every token as list-priced with no prompt caching. Cached input tokens (where a vendor offers them) cost far less on repeated context; check the full pricing table for cached rates and source links before you commit spend.

How the Estimate Works

The calculator takes two numbers — your monthly input tokens and monthly output tokens, both in millions — and multiplies each against every model's published per-million-token price. Input and output are priced separately because output almost always costs several times more than input.

  • Monthly cost = (input tokens in millions × input price) + (output tokens in millions × output price)
  • Every model is ranked from cheapest to most expensive at your specific volume
  • Prices come from the same ledger as our full AI model pricing page, sourced from each vendor's official pricing page

From Runs Per Day to Monthly Cost

To price a scheduled job, measure the tokens one run uses, multiply by runs per day, then multiply by 30. That gives you the monthly token volume the calculator above asks for. Most people plan automation in runs per day, so this is the missing step between the two.

Here is a worked example. A small business runs a daily digest job. It reads 20 pages, summarizes what changed, and emails the team a short recap. Swap in your own numbers at each step.

  • One run reads 20 pages at about 2,000 tokens each = 40,000 input tokens
  • One run writes a recap of about 2,000 output tokens
  • Running it 4 times a day = 160,000 input and 8,000 output tokens per day
  • Over 30 days = 4,800,000 input and 240,000 output tokens per month
  • In the calculator's units, that is 4.8 million input and 0.24 million output

Now price it. Claude Haiku 4.5 lists at $1 per million input tokens and $5 per million output tokens on Anthropic's pricing page. So input costs 4.8 × $1 = $4.80, output costs 0.24 × $5 = $1.20, and the job runs about $6.00 a month. Raise it to 24 runs a day and the same math lands near $36.00.

A scheduled job costs nothing between runs. It wakes on its schedule, does the work, and exits. It is not sitting there refreshing, and it does not bill you while you have the page closed. You pay for tokens the run actually used, so cost scales with how often you schedule it.

For how to set the schedule itself, see our guide on how to run a script on a schedule. Automation bills per token through an API key, not through a seat plan — our guide on an API key versus a subscription covers which one you need.

Why Cost Varies So Much Between Models

The cheapest models on this list, such as Amazon Nova Micro or DeepSeek V4 Flash, cost a small fraction of a cent per million tokens. The most capable flagship models, such as Claude Fable 5 or GPT-5.6 Sol, cost tens of dollars per million tokens because they are built for the hardest reasoning, coding, and research work.

A cheaper model is not automatically the right choice. Route routine, high-volume work to a budget tier, and reserve a flagship model for the smaller share of tasks where quality actually changes the outcome.

What This Calculator Does Not Cover

This is a token-price estimate, not a full bill. It does not include prompt-caching discounts (several vendors price repeated context far cheaper on a cache hit), per-request fees some vendors add on top of token price, or seat-based subscription plans like ChatGPT Plus or Claude Pro, which are priced flat per user rather than per token.

Use the result to compare models against each other for your workload, then confirm the current price and any caching terms on the vendor's official pricing page before you commit spend.


Frequently Asked Questions

  • Start from a real workload. Count roughly how many requests you send per month and the average input and output length per request, then multiply. A support agent answering 2,000 tickets a month with a 1,000-token context is close to 2 million input and 1 million output tokens.
  • Multiply the tokens one run uses by your per-token rate. A daily digest job that reads 20 pages (about 40,000 input tokens) and writes a 2,000-token recap costs about 5 cents per run on Claude Haiku 4.5, at $1 per million input and $5 per million output tokens. Run it 4 times a day for 30 days and that is about $6.00 a month. The job costs nothing between runs — it wakes, works, and exits.
  • Yes, and that is the reason to run the numbers instead of trusting a single per-token price. Because input and output are priced separately, a model that looks cheap on input can slip down the ranking once you weight it by a heavy-output workload (long generated answers, not long prompts). Enter your own volume above to see the ranking for your actual mix, not a generic list.
  • No. It prices every token at the standard list rate. Several vendors, including Anthropic and Google, offer a much cheaper cached-input rate for repeated context. If your workload reuses a large system prompt or document, your real cost will be lower than this estimate.
  • No. This tool estimates API (pay-per-token) costs only. Seat-based subscription plans are a flat monthly price regardless of usage — see the full pricing table for those.
  • Real bills depend on your exact token counts, caching hit rate, and any per-request fees a vendor charges (Perplexity Sonar, for example, adds a search fee per request). Treat this as a starting comparison, then confirm the current price on the vendor's official pricing page.

Not Sure Which Model Fits Your Workload?

Token price is only one input. Our free AI workflow audit maps the right model, plan, and routing strategy to your exact tasks, budget, and compliance needs.

Book a Consultation