Home / Pricing Methodology
A price without a source is a rumor. This page documents where every number on this site comes from, when it was last checked, and exactly what the calculators compute.
✓ Last verified: September 26, 2026 · Manually verifiedSource: OpenAI official pricing. Per 1M tokens, USD. These are the short-context Standard rates — the ones our calculators use.
| Model | Input / 1M | Cached input / 1M | Output / 1M |
|---|---|---|---|
| GPT-6 Luna | $0.10 | $0.01 | $0.50 |
| GPT-6 Sol | $2.00 | $0.20 | $10.00 |
| GPT-6 Astra | $10.00 | $1.00 | $50.00 |
OpenAI's pricing page shows a second column: long-context rates that apply to very large prompts. They are higher:
| Model | Input / 1M | Cached input / 1M | Output / 1M |
|---|---|---|---|
| GPT-6 Luna | $0.20 | $0.02 | $0.75 |
| GPT-6 Sol | $4.00 | $0.40 | $15.00 |
| GPT-6 Astra | $20.00 | $2.00 | $75.00 |
Which one does the calculator use? Short-context rates — they cover the overwhelming majority of API workloads. If your prompts routinely exceed ~200K tokens, treat our estimate as a lower bound and check the long-context column on the official page.
The official pricing page lets you switch between service tiers. Our calculator models three of them:
Deliberately excluded: Fast mode (priority throughput) has no public price multiplier, so we don't guess one. FedRAMP and regional-processing endpoints carry a 10% surcharge — also excluded; our numbers assume standard global endpoints.
For one request, with c = cached-input share (0–100%):
Key point: cached input replaces the standard input rate for the cached portion — the two are never stacked. The tier multiplier (1.0× / 0.5×) applies once, to the total.
Don't double-count discounts. The most common spreadsheet mistake is applying the cached-input rate and the Batch 50% to the same tokens as if they compounded beyond what's shown above. They don't — cached tokens are billed at the cached rate, then the tier multiplier applies to the whole request. Our calculator follows this order.
Our comparison pages also show verified Anthropic and Google prices — with one honest caveat: caching works differently per provider, so cross-provider comparisons are close estimates, not exact invoices.
| Model | Input / 1M | Cached input / 1M | Output / 1M | Source |
|---|---|---|---|---|
| Claude Opus 5.5 | $4.00 | $0.20 | $20.00 | Anthropic |
| Claude Sonnet 5 | $2.00 | $0.20 | $10.00 | Anthropic |
| Claude Haiku 4.5 | $1.00 | $0.10 | $5.00 | Anthropic |
| Gemini 2.5 Pro | $1.25 | $0.125 | $10.00 | |
| Gemini 2.5 Flash | $0.30 | $0.03 | $2.50 | |
| Gemini 2.5 Flash-Lite | $0.10 | $0.01 | $0.40 |
Spot a price that's changed? The official pages are the source of truth — OpenAI, Anthropic, Google — and this page is updated whenever our weekly check confirms a change.