LLM API pricing

LLM API Pricing Comparison

Compare estimated monthly cost across major LLM APIs using the same request volume, input tokens, output tokens, cached input, and budget buffer.

Compare LLM API prices

Choose a workload, enter monthly requests and token size, then compare provider cost before committing.

  1. 1Pick workloadUse the closest API usage scenario.
  2. 2Enter tokensAdd request volume, input, output, and cache.
  3. 3Compare modelsCheck monthly cost and effective rate.
  4. 4Share or exportCopy the link or download the CSV.
Lowest estimated monthly cost$0.00
Lowest-cost model-
Budget with buffer$0.00
Monthly tokens0
Models compared0

Next: copy the share link or export the CSV, then request a worksheet or send a pricing correction if the assumptions need review.

ProviderModelInput / 1MCached / 1MOutput / 1MDiscountMonthly costCost / request
Comparison estimate: model price is only one decision input. Check latency, output quality, context limits, tool support, and data controls before moving production traffic.
Same workload

Compare apples to apples

Use one traffic and token profile across all listed models instead of comparing headline prices.

Cached input

Model repeated context

Include discounted cached input when system prompts, retrieval context, or examples repeat.

Budgeting

Add a launch buffer

Plan for traffic variance, retries, evaluation runs, and production monitoring overhead.

How to compare LLM API pricing

Start with monthly request volume, average input tokens, average output tokens, and the share of input that is likely to be cached. This page then ranks models by estimated monthly spend for that exact workload.

After narrowing the model shortlist, use more specific calculators for the next decision: AI Cost Calculator for full product unit economics, OpenAI vs Claude Cost Calculator for a focused provider comparison, and Token Counter when you need to estimate prompt size from real text.

QuestionUse this page whenFollow-up
Which LLM API is cheapest?You have rough request and token assumptions.AI API Pricing Calculator
OpenAI or Claude?You want a narrower comparison before testing quality.OpenAI vs Claude Cost Calculator
How much will the product cost?You need per-user and per-action AI cost, not just model spend.AI Cost Calculator
How many tokens are in my prompt?You need a first estimate from pasted text.Token Counter

FAQ

Which LLM API is cheapest?

The cheapest LLM API depends on your input tokens, output tokens, cached input share, and model quality needs. Run the same workload across models before choosing.

Why compare LLM API prices by workload?

Input and output prices are different, and some providers discount cached input. A model that looks cheap per input token may be more expensive for long outputs or uncached traffic.

Sources

Prices in this site were last checked on 2026-07-04. Pricing can change without notice.