gpt-5.6-sol · Released 09 July 2026
| Provider | OpenAI |
|---|---|
| Family | GPT-5.6 |
| Modality | multimodal |
| Context | 1,050,000 tokens |
| Max output | 128,000 tokens |
| Reasoning | Yes |
| Tool calling | Yes |
| Structured output | Yes |
| Open weights | No / not recorded |
| Licence | — |
No verified Cybatar-run benchmark result has been published for this model yet.
Versioned observations tied to the exact deployment and source date.
Promotional pricing published after the August 21 update; promotional period stated to run at least through November 21, 2026.
Evidence ↗OpenAI updates GPT-5.6 Sol pricing with a temporary promotional reduction.
Evidence ↗Official-source release record for the exact GPT-5.6 Sol model version.
Evidence ↗GPT-5.6 becomes generally available across Sol, Terra and Luna tiers.
Evidence ↗Reported by independent evaluators or the model provider. These are contextual research observations, not Cybatar-run scores.
| Benchmark | Score | Configuration / harness | Provenance | Source |
|---|---|---|---|---|
| Artificial Analysis Intelligence Index v4.1.1 | 61 | Model evaluation · reasoning effort: max | Independent | Artificial Analysis |
| GDPval-AA v2 | 1679 Elo | Model evaluation · reasoning effort: xhigh | Independent | Artificial Analysis |
| Artificial Analysis Coding Agent Index v1.3 | 67 | Codex · reasoning effort: max | Independent | Artificial Analysis |
| DeepSWE 1.1 | 69 % | Codex · reasoning effort: max | Independent | Artificial Analysis |
| Terminal-Bench 2.1 | 88 % | Codex · reasoning effort: max | Independent | Artificial Analysis |
| SWE-Atlas-QnA 2026 | 43 % | Codex · reasoning effort: max | Independent | Artificial Analysis |
| GPQA Diamond 2026 | 94.6 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| FrontierMath Tier 1-3 v2 | 89 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| FrontierMath Tier 4 v2 | 83 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| MMMU-Pro (no tools) 2026 | 83 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| MMMU-Pro (with tools) 2026 | 84.6 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| AutomationBench 2026 | 18.1 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| GDP.pdf 2026 | 30.7 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| HealthBench Professional 2026 | 60.5 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| Terminal-Bench 2.1 | 88.8 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |
| DeepSWE 1.1 | 72.7 % | Model evaluation · provider reported configuration: OpenAI launch evaluation | Provider Reported | OpenAI |