AI API pricing

AI API Pricing Calculator

Compare estimated monthly API cost across major LLM models from the same request volume, input tokens, output tokens, and cached-input assumptions.

Compare AI API pricing

Choose a workload, enter monthly requests and token size, then compare provider cost before committing.

  1. 1Pick workloadUse the closest API usage scenario.
  2. 2Enter tokensAdd request volume, input, output, and cache.
  3. 3Compare modelsCheck monthly cost and effective rate.
  4. 4Share or exportCopy the link or download the CSV.
Lowest estimated monthly cost$0.00
Lowest-cost model-
Budget with buffer$0.00
Monthly tokens0
Models compared0

Next: copy the share link or export the CSV, then request a worksheet or send a pricing correction if the assumptions need review.

ProviderModelInput / 1MCached / 1MOutput / 1MDiscountMonthly costCost / request
Comparison estimate: model prices and availability can change. Use this page for first-pass budgeting, then confirm with each provider's official pricing page.
Compare

Same workload, many models

Keep request volume and token assumptions constant while comparing provider-level cost.

Cached input

Account for repeated context

Apply a cached-input share when long instructions or retrieval context repeat across requests.

Provider pages

Drill into details

Use dedicated OpenAI, Claude, or Gemini pages when provider-specific billing rules matter.

When to use this AI API pricing calculator

Use this comparison when you are choosing a model family or estimating a new feature. It works best after you have a rough average for input tokens, output tokens, and monthly request volume.

For provider-specific rules like Claude prompt cache writes, Gemini grounding, Mistral batch discounts, xAI tool pricing, or OpenAI tokenization, use the dedicated calculator pages and source links after this first comparison.

NeedBest pageWhy
OpenAI-specific budgetOpenAI API Cost CalculatorIncludes OpenAI cached input and budget buffer assumptions.
Claude prompt cachingClaude Token Cost CalculatorSeparates cache reads, cache writes, and output cost.
Gemini grounding or Batch APIGemini Cost CalculatorIncludes context caching, Batch API, storage, and grounding assumptions.
Raw text token estimateToken CounterStarts from pasted prompt text before you know average token counts.

FAQ

Is the cheapest model always the best choice?

No. Cost is only one input. Evaluate latency, context window, quality, tool support, data controls, and reliability before choosing a production model.

Why do cached input prices matter?

Repeated system prompts, policies, examples, or retrieval context can materially change cost when a provider supports discounted cached input.

Sources

Prices in this site were last checked on 2026-07-04. Pricing can change without notice.