Skip to main content

Token cost calculator

Token Cost Calculator

Calculate how input tokens, cached input tokens, and output tokens contribute to the cost of each AI API request.

Input and output token cost

Choose a model and edit the input and output token counts to inspect per-request and monthly cost.
OpenAIgpt-5.6-sol

Input price

$5.00 / 1M tokens

Output price

$30.00 / 1M tokens

OpenAIgpt-5.6-solStable
Official source

Verified Jul 14, 2026. Estimates vary by usage and provider pricing conditions.

Usage assumptions

Estimate traffic and token usage for an average request.
Active seats, customers, or internal users.
Average AI calls per user each day.
Prompt, history, and retrieved context per request.
Generated answer length; SaaS founders should test long replies.
Reusable cached prompt or context tokens per request.
Use 30 for always-on products or fewer for batch jobs.

Estimated results

Run the calculator to see projected cost and usage volume.

Enter your usage details, then select Calculate estimate to see your projected cost.

Estimated cost = input usage cost + output usage cost + supported optional charges.

Price requests from token components

Token cost is not one blended number. Many providers price input and output tokens separately, and some offer lower cached-input pricing for repeated prompt or context segments. This page focuses on the cost mechanics behind each request.

Calculator shortcut

Use the calculator with provider and model pricing to estimate input, cached input, output, and monthly cost.

Calculate token costs

Benefits

Input cost clarity

Estimate the cost of instructions, user text, chat history, and retrieved context.

Output cost planning

Model how generated answer length affects cost when output tokens carry a higher unit price.

Cached-input awareness

Account for cached input pricing where providers support discounted repeated context.

Related planning resources

Continue with the most relevant provider, guide, comparison, or calculator for this page's distinct planning intent.

Use cases

Per-request cost checks

Estimate one average request before multiplying it across traffic.

Prompt optimization

Measure how shorter instructions or smaller context windows can reduce input-token cost.

Response limits

Compare concise and verbose output settings before setting product defaults.

Pricing estimation warning

Provider pricing can change and may include special tiers, batch discounts, or terms not captured by a simple calculator.

Launch checklist

Make the estimate more useful

A few practical checks help developers and founders avoid surprises after real users arrive.

Common cost mistakes

Forgetting retries, long context, power users, and generated output length.

How to reduce AI API costs

Shorten prompts, cap output length, cache repeated answers, and route simple tasks to cheaper models.

Cheaper vs stronger models

Use stronger models when accuracy or reasoning changes the outcome; use cheaper models for routine work.

Before launching an AI feature

Ask who triggers requests, how often, how long responses are, and what happens during usage spikes.

FAQ

How are AI API costs calculated?

Most AI API estimates multiply input tokens, output tokens, and request volume by the selected model's token prices.

Why are output tokens more expensive?

Output tokens often cost more because the model is generating new text, which usually requires more inference work than reading input context.

What are cached input tokens?

Cached input tokens are repeated prompt or context tokens that some providers can reuse at a discounted price.