Workload planning

AI Token Cost Calculator

How do you calculate AI token cost?

AI token cost depends on the selected model, normal input tokens, cached input where supported, output tokens, and request volume. Add the priced input, cache, and output groups to get cost per request, then multiply by daily and monthly traffic. This token cost calculator uses dated official list prices and returns an estimate, not an invoice.

Verified Last verified: Source: Official provider pricing
Applies to 0 input tokens.
Cost estimateEstimated cost
Input tokens0
Cached input tokens0
Cost per request$0.0120
Cost per 1,000 requests$12.0000
Daily cost$0.0120
Monthly cost$0.3600
Annual cost$4.3200
Applied rate tierStandard context
Context window usage0.10%

This estimate uses published text-token rates. It excludes tool calls, media, storage, taxes, regional pricing, platform fees, and provider-side token differences.

Same workload

Compare this request across models

“Lowest estimated cost” means the lowest result for these inputs and listed rates, not a universal claim about the cheapest or best API.

Provider / modelInput tokensInput / 1MCached / 1MOutput / 1MRequestMonthlyContextAccuracyVerified
DeepSeek V4 FlashDeepSeek0$0.14$0.0028$0.28$0.0003$0.00841,000,000Estimate2026-08-09Needs review
DeepSeek V4 ProDeepSeek0$0.435$0.003625$0.87$0.0009$0.02611,000,000Estimate2026-08-09Needs review
GPT-5.6 LunaOpenAI0$0.2Over 272K: $0.4$0.02Over 272K: $0.04$1.2Over 272K: $1.8$0.0012$0.03601,050,000Compatible2026-08-09Verified
GPT-5.4 nanoOpenAI0$0.2$0.02$1.25$0.0013$0.0375400,000Compatible2026-08-09Verified
Gemini 3.5 Flash-LiteGoogle0$0.3$0.03$2.5$0.0025$0.07501,048,576Estimate2026-08-09Verified
Gemini 2.5 FlashGoogle0$0.3$0.03$2.5$0.0025$0.07501,048,576Estimate2026-08-09Verified
GPT-5.4 miniOpenAI0$0.75$0.075$4.5$0.0045$0.1350400,000Compatible2026-08-09Verified
Claude Haiku 4.5Anthropic0$1$0.1$5$0.0050$0.1500200,000Estimate2026-08-09Verified
Gemini 3.6 FlashGoogle0$1.5$0.15$7.5$0.0075$0.22501,048,576Estimate2026-08-09Verified
Gemini 3.5 FlashGoogle0$1.5$0.15$9$0.0090$0.27001,048,576Estimate2026-08-09Verified
Claude Sonnet 5Anthropic0$2$0.2$10$0.0100$0.30001,000,000Estimate2026-08-09VerifiedValid through 2026-08-31
GPT-5.6 TerraOpenAI0$2Over 272K: $4$0.2Over 272K: $0.4$12Over 272K: $18$0.0120$0.36001,050,000Compatible2026-08-09Verified
Gemini 3.1 Pro PreviewGoogle0$2Over 200K: $4$0.2Over 200K: $0.4$12Over 200K: $18$0.0120$0.36001,048,576Estimate2026-08-09Verified
Claude Sonnet 4.6Anthropic0$3$0.3$15$0.0150$0.45001,000,000Estimate2026-08-09Verified
Claude Opus 5Anthropic0$5$0.5$25$0.0250$0.75001,000,000Estimate2026-08-09Verified
GPT-5.6 SolOpenAI0$5Over 272K: $10$0.5Over 272K: $1$30Over 272K: $45$0.0300$0.90001,050,000Compatible2026-08-09Verified
Claude Fable 5Anthropic0$10$1$50$0.0500$1.50001,000,000Estimate2026-08-09Verified

The 17 model rows use USD list prices per one million text tokens, checked against the linked official provider pages. OpenAI and Gemini long-context tiers are applied automatically from the entered input count. Batch, audio, image, video, tools, storage, regional, and platform fees may differ.

How much does one AI API request cost?

The token cost calculator divides input into normal and cached groups, applies the listed per-million rates, and adds expected output cost. That sum is the estimated cost per request. The token cost calculator then multiplies the request result by daily volume and active days, with an annual projection equal to twelve monthly scenarios.

Use Paste text in the token cost calculator while editing a prompt. Use Enter token count when provider telemetry already supplies usage totals. Review the token cost calculator output again after real usage data becomes available.

How do cached tokens change API cost?

In the token cost calculator, eligible cached input uses the listed cache-read or cache-hit rate instead of the normal input rate. The token cost calculator caps cached tokens at total input, but it cannot prove cache eligibility or a hit. Model cache writes, storage, retention, minimum prefix size, and invalidation separately when the provider charges for them. Rerun the token cost calculator with measured cache-hit data before approving a budget.

How can I compare model cost for one workload?

Keep input, cached input, output, request volume, and active days identical in the token cost calculator, then sort the model table by Lowest estimated cost. Start with a typical request and rerun a long-output or retry case. The lowest result is conditional on those assumptions and does not identify the best model quality or total implementation cost. The token cost calculator should be one input to the model decision, not the decision itself.

Compare the token cost calculator result with LLM API pricing, then inspect the GPT cost calculator, Claude cost calculator, or GPT vs Claude comparison.

Five token cost calculator checks before trusting a monthly estimate

Source and method

Which sources and measured checks support this token cost calculator?

The token cost calculator covers 17 current model records and measured zero, 1, 1,000, 1 million, and 10 million-token cases, long-context tiers, cached input, missing prices, zero output, and high request volume. Its limitation is scope: calculations use direct-API text list prices and exclude tools, media, storage, regional terms, taxes, retries, and provider token differences. OpenAI official pricing. Last verified: .

Which sources and measured checks support this token cost calculator? Evidence and limitations for the current page.
EvidenceObserved result or boundary
Measured cost cases0 to 10 million tokens, 0% to 100% cached input, and large request volume.
Provider coverageThe token cost calculator compares OpenAI, Anthropic Claude, Google Gemini, and DeepSeek standard text rates.
Answers

Token cost calculator questions

What does the token cost calculator include?

The token cost calculator includes normal input, cached input when a separate rate is listed, expected output, requests per day, days per month, and a twelve-month annual projection.

Does entering cached input guarantee cache hits?

No. The token cost calculator treats it as a planning assumption. Cache eligibility, write charges, retention, and prefix matching vary by provider and request.

Why can the final API bill differ?

The token cost calculator cannot see provider tokenization, tool schemas, message framing, thinking tokens, media, retries, storage, regional rates, platform fees, or taxes that can change the final charge.

Which model is cheapest?

The token cost calculator table can identify the lowest estimated cost for the entered workload and current list prices. It cannot establish the best model or a guaranteed cheapest total implementation.

Close the token cost calculator estimate with real usage

This token cost calculator is suitable for planning and model comparison, but it is not for invoicing. Before approving a budget, check the provider pricing source and reconcile the token cost calculator estimate against production usage records.