LLM API pricing usually depends on the selected model, normal input tokens, cached input where supported, and generated output tokens. Providers publish separate USD rates per one million tokens. The table below compares standard direct-API text rates; actual workload cost also depends on output length, retries, tools, media, regional service, and platform terms.
What are the current LLM input and output token prices?
This LLM API pricing table lists standard direct-API USD prices per one million text tokens for 17 current GPT, Claude, Gemini, and DeepSeek models. Long-context alternatives appear beneath affected rates, and the workload calculator applies them automatically. Each row links to its official provider source and shows the verified date. A missing value is unavailable, never assumed to be zero.
The 17 model rows use USD list prices per one million text tokens, checked against the linked official provider pages. OpenAI and Gemini long-context tiers are applied automatically from the entered input count. Batch, audio, image, video, tools, storage, regional, and platform fees may differ.
Workload model
What does one selected model cost for this workload?
Start with the 5 million input and 1 million output token reference workload, then replace those values with your own input, cached input, output, and request volume. Switch the selected model to test another row from the complete comparison above; the result is conditional on those assumptions and the currently verified rates.
Applies to 0 input tokens.
Cost estimateEstimated cost
Input tokens5,000,000
Cached input tokens0
Cost per request$38.0000
Cost per 1,000 requests$38,000.0000
Daily cost$38.0000
Monthly cost$1,140.0000
Annual cost$13,680.0000
Applied rate tierOver 272K input tokens
Context window usage100.00%
This estimate uses published text-token rates. It excludes tool calls, media, storage, taxes, regional pricing, platform fees, and provider-side token differences.
When was the LLM API pricing verified?
Every LLM API pricing row was checked on against the linked OpenAI, Anthropic, Google, or DeepSeek pricing page. Recheck the LLM API pricing source before a purchasing decision. Claude Sonnet 5 shows the official introductory rate and its August 31, 2026 valid-through date. DeepSeek LLM API pricing remains flagged for review because its official page warns of a planned price increase.
Why are output tokens important in LLM API pricing?
Output often uses a higher per-million rate than input. A model with a low input price can produce a higher total when answers are long, retries are common, or a workflow generates large structured results.
Compare typical and high-output cases before choosing a model from the LLM API pricing table.
How does cached input change LLM API pricing?
LLM API pricing for cached input applies only when a provider recognizes an eligible cache hit. Writes, storage, retention, minimum prefix size, and invalidation can change the economics. The calculator models reads, not automatic eligibility.
Does the lowest LLM API pricing mean the best API?
No. In LLM API pricing, “Lowest estimated cost” describes one selected workload. Quality, latency, context, tool support, data handling, availability, and engineering fit can outweigh the token line item.
Where does LLM API pricing compare GPT and Claude?
Which sources and measured checks support this LLM API pricing comparison?
The LLM API pricing comparison checked 17 active model rows against current OpenAI, Anthropic Claude, Google Gemini, and DeepSeek first-party pages, then measured identical token workloads with one cost formula and applicable long-context rate tiers. Its limitation is commercial scope: batch, media, tools, storage, regional service, taxes, credits, and partner-platform pricing are excluded. OpenAI official pricing. Last verified: .
Which sources and measured checks support this LLM API pricing comparison? Evidence and limitations for the current page.
Evidence
Observed result or boundary
Verified model rows
17 models across OpenAI, Anthropic Claude, Google Gemini, and DeepSeek.
Measured workload
Normal input, cached input, output, request volume, monthly and annual cost.
Answers
LLM API pricing questions
How current is this LLM API pricing comparison?
This LLM API pricing table was checked against each linked official provider page on 2026-08-09. DeepSeek is marked Needs review because its official page still announces a planned price increase.
Does LLM API pricing include free tiers?
No. This LLM API pricing table compares standard paid direct-API text rates. Free quota, promotional credits, batch discounts, enterprise contracts, and platform fees are separate.
Which model has the lowest LLM API pricing?
The lowest LLM API pricing estimate depends on the selected models, input, cache, output, and request volume. The sorted result applies only to that workload and current published rates.
Does LLM API pricing include unified API partners?
No verified unified affiliate provider is active in the LLM API pricing table. Any future partner must have a real URL, rights review, disclosure, and independently checked terms.
Choose LLM API pricing from workload evidence, not one rate cell
This LLM API pricing comparison is suitable for screening a defined text workload, but it is not for creating a universal model ranking. Use realistic input, output, cache, and retry data, then recheck the official source before a purchase or deployment decision.