Home / GPT-6.1 Sol API Cost Calculator

GPT-6.1 Sol API Cost Calculator

Calculate GPT-6.1 Sol API costs with long-context pricing, cached reads, cache writes, Fast, Flex, Batch and EU regional processing.

Official OpenAI rates · Last verified September 30, 2026

GPT-6.1 Sol pricing

ModelInput / 1MCached read / 1MCache write / 1MOutput / 1M
GPT-6.1 Sol$2.00$0.10$2.50$10.00

Long context

ModelInput / 1MCached read / 1MCache write / 1MOutput / 1M
GPT-6.1 Sol$4.00$0.20$5.00$15.00

Worked workload examples

For 100K input and 10K output with one Standard short-context request, no cache: 100,000 × $2 / 1M + 10,000 × $10 / 1M = $0.30.

With 50% cached reads: 50,000 × $2 / 1M + 50,000 × $0.10 / 1M + 10,000 × $10 / 1M = $0.205.

With 50% reads and 10% writes: 40,000 × $2 / 1M + 50,000 × $0.10 / 1M + 10,000 × $2.50 / 1M + 10,000 × $10 / 1M = $0.210. The three input buckets sum to 100K.

300K input, 10K output, no cache: 300,000 × $4 / 1M + 10,000 × $15 / 1M = $1.35. Fast: $2.70; Flex or Batch: $0.675; EU Standard: $1.485.

Current and previous Sol

GPT-6.1 Sol replaces GPT-6 Sol in the site's current lineup. The API model id is gpt-6.1-sol, with a 1,050,000-token context window and 128,000 maximum output tokens.

Compare GPT-6.1 Sol vs GPT-6 Sol → · Previous GPT-6 Sol calculator · What changed at release?

Frequently asked questions

How much does GPT-6.1 Sol cost?

Standard short-context rates per 1M tokens are $2.00 fresh input, $0.10 cached read, $2.50 cache write and $10.00 output.

When does long-context pricing apply?

Input above 272,000 tokens uses long rates for the full request: $4.00 fresh input, $0.20 cached read, $5.00 cache write and $15.00 output. Exactly 272,000 input tokens remains short context.

Does Fast mode work with EU regional processing?

No. Fast costs 2× Standard with global processing. EU regional processing adds 10% and supports Standard, Flex and Batch; the calculator blocks EU plus Fast.

Is GPT-6.1 Sol always 50% cheaper than GPT-6 Sol?

No. Cached reads cost 50% less. Fresh input, cache writes and output have the same published rates. Total savings depend on the cached-read share and workload.

Official sources

GPT-6.1 Sol model documentation · OpenAI API pricing · Pricing methodology