Source-backed AI Wiki tool

AI API cost and context planner

Turn a token workload into per-request, monthly, and annual estimates. Compare official prices and see which model context windows can actually hold the request.

No API keyInputs stay in your browserFacts reviewed 2026-07-23
This is a planning estimate, not a quote or invoice. Review the selected model's effective dates and caveats, then confirm the current provider price before committing a budget.

Private, browser-only estimate

No API key is needed. Your workload inputs stay on this device and are not saved or sent to a provider.

Workload

Tokens and request volume

Cached input is a subset of input, never an additional token count.

Pricing mode

A published cached-input rate is applied only to the cached subset. Otherwise those tokens use the regular input rate.

Selected estimate

GPT-5.6 Terra

gpt-5.6-terra

Per request

$0.0688

22K total tokens

Per month

$1,512.50

22K requests

Per year

$18,150.00

12 equal months

Cost composition

Input
$0.0388
Output
$0.0300
Cache savings
−$0.0113

Published rates

Input / 1M
$2.50
Output / 1M
$15.00
Cached input / 1M
$0.2500

Context and output fit

22K combined input + requested output

2.1% used

Context: 1.1M

Max output: 128K

1M context tokens remain before the published window limit.

ActiveRates effective Jul 23, 2026

Standard short-context processing. Requests above 272K input tokens receive long-context multipliers.

Ranked comparison

Compare sourced API prices

Sorted by estimated monthly cost. Select a row to inspect its rates and constraints above.

API model costs for the entered workload, ranked by monthly estimate
ModelPer requestPer monthContext fitSource
$0.002350$51.702.1%Official ↗
$0.004200$92.408.6%Official ↗
$0.009650$212.302.1%Official ↗
$0.009650$212.302.1%Official ↗
$0.0130$286.00Does not fitOfficial ↗
$0.0255$561.0011%Official ↗
$0.0267$586.8511%Official ↗
$0.0275$605.002.1%Official ↗
$0.0281$617.1011%Official ↗
$0.0383$841.502.1%Official ↗
$0.0394$866.252.1%Official ↗
$0.0413$907.502.1%Official ↗
$0.0510$1,122.002.2%Official ↗
$0.0550$1,210.002.1%Official ↗
$0.0688$1,512.502.1%Official ↗
$0.1275$2,805.002.2%Official ↗
$0.1375$3,025.002.1%Official ↗
$0.2550$5,610.002.2%Official ↗
$0.2600$5,720.0017%Official ↗
$0.3825$8,415.0011%Official ↗
$0.3825$8,415.0011%Official ↗
$0.7200$15,840.00Does not fitOfficial ↗
Model facts reviewed Jul 23, 2026. Estimates exclude taxes, minimum commitments, rate tiers, storage, web search, tools, grounding, fine-tuning, and provider-specific regional or partner-platform charges.

Transparent methodology

What the estimate includes

01

Split input correctly

Cached tokens are carved out of total input. In Standard mode, only that subset receives a recorded cache rate; Batch uses its own published input rate.

02

Scale request volume

Input cost and output cost form a per-request estimate. Requests per day multiplied by working days produces the monthly volume; twelve equal months produces the annual figure.

03

Check the constraints

Input plus requested output is compared with the context window. Output is also checked independently when the provider publishes a separate maximum.

Common questions

AI API pricing FAQ

How is monthly AI API cost calculated?

The calculator multiplies the selected model's published per-million-token input and output rates by your tokens per request, then multiplies the request estimate by requests per working day and working days per month. The yearly figure uses twelve equal months.

Are cached input tokens added to regular input tokens?

No. Cached input is treated as a subset of the total input. The cached subset receives a published cached-input rate when one is recorded; otherwise it receives the ordinary input rate.

Does Batch always cost half as much?

No. Batch mode appears only when the model registry contains official Batch input and output rates. The calculator does not assume a percentage discount, combine Batch with a cached-input discount, or infer missing prices.

Does the estimate include every provider fee?

No. It covers the recorded text-token prices only. Taxes, regional pricing, minimum commitments, long-context tiers, storage, searches, grounding, tools, fine-tuning, media, partner platforms, and other feature charges may change the actual bill.