GPT-6.1 Sol API Cost Calculator
Calculate GPT-6.1 Sol API costs with long-context pricing, cached reads, cache writes, Fast, Flex, Batch and EU regional processing.
Official OpenAI rates · Last verified September 30, 2026
GPT-6.1 Sol pricing
| Model | Input / 1M | Cached read / 1M | Cache write / 1M | Output / 1M |
|---|---|---|---|---|
| GPT-6.1 Sol | $2.00 | $0.10 | $2.50 | $10.00 |
Long context
| Model | Input / 1M | Cached read / 1M | Cache write / 1M | Output / 1M |
|---|---|---|---|---|
| GPT-6.1 Sol | $4.00 | $0.20 | $5.00 | $15.00 |
Worked workload examples
For 100K input and 10K output with one Standard short-context request, no cache: 100,000 × $2 / 1M + 10,000 × $10 / 1M = $0.30.
With 50% cached reads: 50,000 × $2 / 1M + 50,000 × $0.10 / 1M + 10,000 × $10 / 1M = $0.205.
With 50% reads and 10% writes: 40,000 × $2 / 1M + 50,000 × $0.10 / 1M + 10,000 × $2.50 / 1M + 10,000 × $10 / 1M = $0.210. The three input buckets sum to 100K.
300K input, 10K output, no cache: 300,000 × $4 / 1M + 10,000 × $15 / 1M = $1.35. Fast: $2.70; Flex or Batch: $0.675; EU Standard: $1.485.
Current and previous Sol
GPT-6.1 Sol replaces GPT-6 Sol in the site's current lineup. The API model id is gpt-6.1-sol, with a 1,050,000-token context window and 128,000 maximum output tokens.
Compare GPT-6.1 Sol vs GPT-6 Sol → · Previous GPT-6 Sol calculator · What changed at release?
Frequently asked questions
How much does GPT-6.1 Sol cost?
Standard short-context rates per 1M tokens are $2.00 fresh input, $0.10 cached read, $2.50 cache write and $10.00 output.
When does long-context pricing apply?
Input above 272,000 tokens uses long rates for the full request: $4.00 fresh input, $0.20 cached read, $5.00 cache write and $15.00 output. Exactly 272,000 input tokens remains short context.
Does Fast mode work with EU regional processing?
No. Fast costs 2× Standard with global processing. EU regional processing adds 10% and supports Standard, Flex and Batch; the calculator blocks EU plus Fast.
Is GPT-6.1 Sol always 50% cheaper than GPT-6 Sol?
No. Cached reads cost 50% less. Fresh input, cache writes and output have the same published rates. Total savings depend on the cached-read share and workload.
Official sources
GPT-6.1 Sol model documentation · OpenAI API pricing · Pricing methodology