RunInfraby RightNow
  • Model APIsNew
  • Pricing
  • Research
  • Contact
DashboardSign inGet started
Loading related term
Loading related term
Loading related term
Loading related term
Home/Catalog/GPT-5.6 Sol
Model reference

GPT-5.6 Sol

This reference covers the API-only flagship capability tier in a model family whose tiers advance on independent cadences. It keeps sibling pricing changes and hosted availability separate from capability claims.

GPT-5.6 Sol is the vendor reference for gpt-5.6-sol, as of 2026-08-12. Context window: 1,050,000, as of 2026-08-12. RunInfra has not measured this model; every figure below belongs to its named source, cited and dated.

Cited specifications

GroupFactSource-cited display valueSource and date
IdentityAPI model idgpt-5.6-sol
SourceRetrieved 2026-08-12
IdentitySeries launchWe're launching the GPT-5.6 family of models for general availability following our limited preview: our new flagship, Sol, alongside Terra, a balanced model for everyday work, and Luna, our most cost-efficient model.
SourceRetrieved 2026-08-12
IdentitySeries launch publication dateJuly 9, 2026
SourceRetrieved 2026-08-12
IdentitySeries naming systemIn this new naming system introduced with GPT-5.6, the number identifies a model's generation, while Sol, Terra, and Luna identify durable capability tiers that can advance on their own cadence.
SourceRetrieved 2026-08-12
IdentityPreview publication dateJune 26, 2026
SourceRetrieved 2026-08-12
IdentityVendor positioningGPT-5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.
SourceRetrieved 2026-08-12
Architecture---
ContextContext window1,050,000
SourceRetrieved 2026-08-12
ContextMaximum input922,000 tokens
SourceRetrieved 2026-08-12
ContextMaximum output128,000 tokens
SourceRetrieved 2026-08-12
ContextKnowledge cutoffFebruary 16, 2026
SourceRetrieved 2026-08-12
ModalitiesInput and outputtext and image input; text output
SourceRetrieved 2026-08-12
ModalitiesReasoning effortnone, low, medium (default), high, xhigh, and max
SourceRetrieved 2026-08-12
ModalitiesMaximum reasoning effortmax gives GPT-5.6 even more time than xhigh to reason and explore alternatives, run checks, and revise its approach.
SourceRetrieved 2026-08-12
ModalitiesUltra modeultra goes further by coordinating four agents in parallel by default, trading higher token use for stronger results and faster time-to-result on demanding tasks.
SourceRetrieved 2026-08-12
License---
PricingOpenAI standard input$5.00 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI standard cached input$0.50 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI standard cache write$6.25 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI standard output$30.00 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI long-context input$10.00 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI long-context cached input$1.00 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI long-context cache write$12.50 per 1M tokens
SourceRetrieved 2026-08-12
PricingOpenAI long-context output$45.00 per 1M tokens
SourceRetrieved 2026-08-12
PricingLong-context thresholdPrompts with >272K input tokens are priced at 2x input and 1.5x output
SourceRetrieved 2026-08-12
PricingCache economicscache writes are billed at 1.25x the model's uncached input rate, while cache reads continue to receive the 90% cached-input discount.
SourceRetrieved 2026-08-12
PricingSibling price cutUpdate on July 30, 2026: OpenAI reduced the price of GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20%.
SourceRetrieved 2026-08-12
PricingOpenRouter listing$5.00/M input tokens and $30.00/M output tokens; listed July 9, 2026
SourceRetrieved 2026-08-12
AvailabilityLaunch availabilityGPT-5.6 is available starting today across ChatGPT, Codex, and the OpenAI API.
SourceRetrieved 2026-08-12
AvailabilityTier rate limitsTier 5: 15,000 RPM and 40,000,000 TPM
SourceRetrieved 2026-08-12
AvailabilitySol ProPro and Enterprise users can also select GPT-5.6 Sol Pro for the highest-quality results on complex tasks.
SourceRetrieved 2026-08-12

What the vendor says is new

"We're launching the GPT-5.6 family of models for general availability following our limited preview: our new flagship, Sol, alongside Terra, a balanced model for everyday work, and Luna, our most cost-efficient model."

SourceRetrieved 2026-08-12

"ultra goes further by coordinating four agents in parallel by default, trading higher token use for stronger results and faster time-to-result on demanding tasks."

SourceRetrieved 2026-08-12

Vendor-claimed benchmarks

These results are vendor-claimed, not independently measured by RunInfra.

BenchmarkVendor-claimed valueSource and date
Agents' Last Exam52.7%
SourceRetrieved 2026-08-12
Artificial Analysis Intelligence Index v4.158.9
SourceRetrieved 2026-08-12
Artificial Analysis Coding Agent Index v1.180
SourceRetrieved 2026-08-12
SWE-Bench Pro64.6%
SourceRetrieved 2026-08-12
DeepSWE v1.172.7%
SourceRetrieved 2026-08-12
Terminal-Bench 2.188.8%
SourceRetrieved 2026-08-12
BrowseComp90.4%
SourceRetrieved 2026-08-12
OSWorld 2.062.6%
SourceRetrieved 2026-08-12
GPQA Diamond94.6%
SourceRetrieved 2026-08-12
Capture-the-Flag96.7%
SourceRetrieved 2026-08-12
FrontierMath Tiers 1-389%
SourceRetrieved 2026-08-12
MMMU Pro, no tools83%
SourceRetrieved 2026-08-12

Serving support

Listed rows have a cited upstream support signal. They are not RunInfra measurements.

No open-engine serving signal applies to this model. It is served only through the vendor's own API.

Serving concepts

  • Context length->
  • Prefix caching->
  • KV cache->
  • Admission control->

What RunInfra measures when we measure it

RunInfra has not measured this model yet. When measurement is published, the record will state throughput, latency, memory, quality, serving conditions, and reproducible evidence.

  • Measurement methodology->
  • Measured package catalog->
  • Published benchmarks->

Questions about this model reference

What is GPT-5.6 Sol?

As of 2026-08-12, api model id: gpt-5.6-sol.

What does the GPT-5.6 Sol API cost?

As of 2026-08-12, openai standard input: $5.00 per 1M tokens. As of 2026-08-12, openai standard cached input: $0.50 per 1M tokens. As of 2026-08-12, openai standard cache write: $6.25 per 1M tokens. As of 2026-08-12, openai standard output: $30.00 per 1M tokens. As of 2026-08-12, openai long-context input: $10.00 per 1M tokens. As of 2026-08-12, openai long-context cached input: $1.00 per 1M tokens. As of 2026-08-12, openai long-context cache write: $12.50 per 1M tokens. As of 2026-08-12, openai long-context output: $45.00 per 1M tokens. As of 2026-08-12, long-context threshold: Prompts with >272K input tokens are priced at 2x input and 1.5x output. As of 2026-08-12, cache economics: cache writes are billed at 1.25x the model's uncached input rate, while cache reads continue to receive the 90% cached-input discount. As of 2026-08-12, sibling price cut: Update on July 30, 2026: OpenAI reduced the price of GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20%. As of 2026-08-12, openrouter listing: $5.00/M input tokens and $30.00/M output tokens; listed July 9, 2026.

Can I run GPT-5.6 Sol myself?

As of 2026-08-12, launch availability: GPT-5.6 is available starting today across ChatGPT, Codex, and the OpenAI API.

RunInfraby RightNow

© 2026 RunInfra. All rights reserved.

System status
Pipeline BuilderModel APIsCost CalculatorPricingStartupsBenchmarksDocsResearchNewsContact
Backed by
YCombinator
AICPA Type II
SOC 2
NVIDIA Inception ProgramNVIDIA Inception Program
Ask AI about RunInfra
Part of RightNow
SecurityDPAAUPCookiesTermsPrivacy