Home / Pricing Methodology

Pricing methodology & sources

A price without a source is a rumor. This page documents where every number on this site comes from, when it was last checked, and exactly what the calculators compute.

✓ Last verified: September 26, 2026 · Manually verified

1. How prices get on this site

2. GPT-6 prices we use (short context)

Source: OpenAI official pricing. Per 1M tokens, USD. These are the short-context Standard rates — the ones our calculators use.

ModelInput / 1MCached input / 1MOutput / 1M
GPT-6 Luna$0.10$0.01$0.50
GPT-6 Sol$2.00$0.20$10.00
GPT-6 Astra$10.00$1.00$50.00

3. Long-context rates exist — and we don't hide them

OpenAI's pricing page shows a second column: long-context rates that apply to very large prompts. They are higher:

ModelInput / 1MCached input / 1MOutput / 1M
GPT-6 Luna$0.20$0.02$0.75
GPT-6 Sol$4.00$0.40$15.00
GPT-6 Astra$20.00$2.00$75.00

Which one does the calculator use? Short-context rates — they cover the overwhelming majority of API workloads. If your prompts routinely exceed ~200K tokens, treat our estimate as a lower bound and check the long-context column on the official page.

4. Service tiers: what the 0.5× means

The official pricing page lets you switch between service tiers. Our calculator models three of them:

Deliberately excluded: Fast mode (priority throughput) has no public price multiplier, so we don't guess one. FedRAMP and regional-processing endpoints carry a 10% surcharge — also excluded; our numbers assume standard global endpoints.

5. Exactly what the calculator computes

For one request, with c = cached-input share (0–100%):

per_request = (fresh_input_tokens / 1M) × input_price + (cached_input_tokens / 1M) × cached_input_price + (output_tokens / 1M) × output_price per_month = per_request × requests_per_month × tier_multiplier per_year = per_month × 12

Key point: cached input replaces the standard input rate for the cached portion — the two are never stacked. The tier multiplier (1.0× / 0.5×) applies once, to the total.

Don't double-count discounts. The most common spreadsheet mistake is applying the cached-input rate and the Batch 50% to the same tokens as if they compounded beyond what's shown above. They don't — cached tokens are billed at the cached rate, then the tier multiplier applies to the whole request. Our calculator follows this order.

6. Other providers: Claude & Gemini

Our comparison pages also show verified Anthropic and Google prices — with one honest caveat: caching works differently per provider, so cross-provider comparisons are close estimates, not exact invoices.

ModelInput / 1MCached input / 1MOutput / 1MSource
Claude Opus 5.5$4.00$0.20$20.00Anthropic
Claude Sonnet 5$2.00$0.20$10.00Anthropic
Claude Haiku 4.5$1.00$0.10$5.00Anthropic
Gemini 2.5 Pro$1.25$0.125$10.00Google
Gemini 2.5 Flash$0.30$0.03$2.50Google
Gemini 2.5 Flash-Lite$0.10$0.01$0.40Google

7. What we won't do

Spot a price that's changed? The official pages are the source of truth — OpenAI, Anthropic, Google — and this page is updated whenever our weekly check confirms a change.

8. Changelog