OpenAI APIStableChecked 2026-09-08Official source ↗

GPT-5.6 Sol pricing, context window, and cost calculator

An OpenAI model with tracked short-context, long-context, cached-input, and cache-write pricing.

Verified fields

Quick facts

API access providerOpenAI
API model IDgpt-5.6-sol
Default input / 1M$4.00
Default output / 1M$20.00
Context window1,050,000
Default profilestandard / short context
Acceptstext, image
Producestext
Knowledge cutoff2026-02-16
StatusStable
Maximum output128,000 tokens

Supported capabilities

  • Streaming
  • Function calling
  • Structured outputs
  • Prompt caching
  • Batch API
  • Reasoning

Only capabilities explicitly tracked from provider documentation are shown.

Source-tracked rates

Verified pricing profiles

Amounts are USD per 1M tokens unless the column states an hourly storage unit.

ProfileInput / 1MCached input / 1MCache write / 1MOutput / 1MNotes
standard / short contextDefault$4.00 USD$0.40 USD$5.00 USD$20.00 USDUsed for default estimates.
standard / long context$8.00 USD$0.80 USD$10.00 USD$30.00 USDTracked separately; select explicitly when modeling this profile.
Official price history

GPT-5.6 Sol price changes

Verified changes for the same model and pricing profile. Cross-model generational price differences are not treated as history.

OpenAI · Standard · Short Context

GPT-5.6 Sol

First recorded 2026-09-03
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$5.00$4.00-$1.00-20.0%
Cached input / 1M$0.50$0.40-$0.10-20.0%
Output / 1M$30.00$20.00-$10.00-33.3%
Cache write / 1M$6.25$5.00-$1.25-20.0%
Checked 2026-09-03Official source ↗

Price change recorded from the current official provider pricing source.

OpenAI · Standard · Long Context

GPT-5.6 Sol

First recorded 2026-09-03
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$10.00$8.00-$2.00-20.0%
Cached input / 1M$1.00$0.80-$0.20-20.0%
Output / 1M$45.00$30.00-$15.00-33.3%
Cache write / 1M$12.50$10.00-$2.50-20.0%
Checked 2026-09-03Official source ↗

Price change recorded from the current official provider pricing source.

Workload estimate

GPT-5.6 Sol cost calculator

Change the workload assumptions. The calculator uses the same shared pricing resolver as StackLens Compare.

Estimated monthly total$140.00
Monthly input cost
$40.00
Monthly output cost
$100.00
Per 1,000 requests
$14.00
Active profile
standard / short context

Estimate excludes untracked provider-specific charges.

StackLens cost guidance

When GPT-5.6 Sol switches to long-context pricing

GPT-5.6 Sol uses a higher tracked rate when a request exceeds 272,000 input tokens. These single-request examples show the pricing change on either side of that boundary.

standard / short context

Below the long-context threshold

Input
250,000 tokens · $1.00
Output
10,000 tokens · $0.20
Estimated token cost
$1.20

At 250,000 input tokens, the shared resolver keeps the request on the tracked short-context profile.

standard / long context

Above the long-context threshold

Input
300,000 tokens · $2.40
Output
10,000 tokens · $0.30
Estimated token cost
$2.70

At 300,000 input tokens, the tracked long-context input and output rates apply to the request.

Boundary: The threshold example shows token pricing, not model quality, latency, retries, or the value of using a larger context window.

Explicit assumptions

Common workload examples

Input only

1M input tokens

$4.00 under the standard.short_context profile, excluding output and other charges.

Output only

1M output tokens

$20.00 under the standard.short_context profile, excluding input and other charges.

Default workload

10,000 monthly requests

At 1,000 input and 500 output tokens per request: $140.00 under standard / short context.

Cache scenario

50% cached input

The same default workload with 50% cached input is estimated at $122.00. Cache savings apply only to the tracked cached-input rate.

Research layer

Experience and rollout notes

Official facts are separated from user reports. Community reports are useful signals, not controlled benchmarks.

Official fact1 official source

Official long-context pricing boundary

OpenAI documents a 1.05M-token context window and 128K maximum output for GPT-5.6 Sol. Requests above 272K input tokens use a higher published input and output pricing tier.

What this does not prove: The pricing threshold is an official billing rule, not a statement about quality at long context.

Recurring user report2 independent community sources reviewed

Implementation scope and usage

A recurring counter-signal is that Sol may overbuild straightforward implementations or consume more tokens than expected, which can offset gains on simpler tasks.

What this does not prove: Measure completed-task cost and review time on your own workload rather than relying on token price alone.

No controlled StackLens benchmark yet

StackLens has not run a controlled GPT-5.6 Sol benchmark. The notes above separate official documentation from independent user reports.

Research reviewed 2026-07-13. Reports may change as GPT-5.6 Sol reaches more workflows.

Verification

Sources and methodology

Pricing and limits are source-tracked and may change. Verify current values with the provider before making production purchasing decisions.

Calculator profile rule: standard.short_context is the verified default. Requests with more than 272,000 input tokens use the higher long-context rate for the full request.