GPT-5.6
OpenAI’s GPT-5.6 family for reasoning, coding, and agents. GPT-6 Astra (2026-09-03) is the new peak on a staged rollout.
OpenAI’s GPT-5.6 family (Sol flagship, Terra balanced, Luna cost-efficient) for complex reasoning, coding, multimodal understanding, and agentic tool use. The gpt-5.6 API alias routes to Sol.
Why GPT-5.6 matters
GPT-5.6 remains the default OpenAI brain many teams ship: mature tool calling, multimodal understanding, and Sol / Terra / Luna cost routing. GPT-6 Astra is the new capability ceiling (API gpt-6-astra) but is still rolling out, costs more at Standard list rates, is generally available in GitHub Copilot (2026-09-04), and is not available in Cursor.
Vision · Audio · Tool calling · Thinking · MCP · Coding
Last reviewed: 4 September 2026
When to choose GPT-5.6
Decision guidance for architects—not a feature list.
Best for
- Complex multi-step reasoning
- Software engineering agents
- Multimodal analysis
- MCP / tool-heavy workflows
Avoid if
- You require open weights or on-prem-only inference
- Cost and vendor lock-in are hard blockers
Strengths
Qualitative snapshot for architects—not a public ranking.
- Tool calling / agents★★★★★
- Ecosystem integrations★★★★★
- Coding★★★★★
- Open weights / self-host★☆☆☆☆
- Cost efficiency★★★☆☆
Ecosystem
Built by
- OpenAI
GPT-5 is OpenAI’s flagship foundation-model family.
Competes with
- Claude Opus
Primary rival for enterprise assistants, coding agents, and careful instruction following.
- Gemini 3.1 Pro
Competes on multimodal frontier capability and cloud platform distribution.
Works with
Recommended for
- AI Agents
Mature tool calling and structured output for agent workloads.
- Large language models
Reference proprietary frontier model for LLM fundamentals.
How GPT-5.6 evolved
Key moments in chronological order.
- Model
OpenAI launches GPT-6 Astra (gpt-6-astra) as the new capability peak. This GPT-5.6 family (Sol/Terra/Luna) remains the volume and cost-routing path. Staged ChatGPT/API/AWS rollout; Enterprise off by default; standard API $10/$50 per 1M tokens. Not supplied to Cursor.
- Platform
GPT-5.6 Sol promotional API pricing
Sol API drops to $4 input / $20 output per 1M tokens (−20% / −33%); promotional pricing through at least Nov 21, 2026. Applies to API and eligible Codex/Work credits; ChatGPT subscription prices unchanged.
- Platform
Ultrafast mode preview for Sol
Limited API preview of Ultrafast for GPT-5.6 Sol (Cerebras-backed): up to ~750 output tokens/sec and up to ~14× Standard. Pricing and GA not published.
- Product
ChatGPT Sol update + Luna for Free/Go
Plus/Pro get an updated GPT-5.6 Sol with a reasoning slider; Free/Go move to Luna defaults with unlimited text chats rolling out. Distinct from July Codex/Work API builds.
- Platform
Terra/Luna API price reductions
Luna −80% and Terra −20% API pricing; Sol unchanged. Fast mode replaces Priority Processing in the API.
- Model
GPT-5.6 family (Sol / Terra / Luna)
Capability tiers make routing explicit: Sol for peak, Terra balanced, Luna cost-efficient.
- API
gpt-5.6 API routing
API aliases route to Sol by default with Terra/Luna for cost control.
- Product
MCP and agent tooling deepen
Tool use, structured output, and MCP-style workflows become first-class for agents.
- Model
Multimodal + long context
Text, image, and audio understanding with million-token-class context for research agents.
- Research
Reasoning-model era (o1 lineage)
OpenAI popularizes deliberate reasoning traces that later feed GPT-5 thinking modes.
Overview
Capabilities
- Vision: Yes
- Audio: Yes
- Tool calling: Yes
- Thinking: Yes
- MCP: Yes
- Coding: Yes
- Structured output: Yes
Technical specifications
- Provider
- OpenAI
- License
- Proprietary
- Context window
- 1.1M
- Parameters
- Undisclosed
- Architecture
- Transformer (proprietary)
- Release
- 2026-07
- Modalities
- Text, Image, Audio
- Vision
- Yes
- Audio
- Yes
- Tool calling
- Yes
- Thinking
- Yes
- MCP
- Yes
- Open weights
- No
- API
- Yes
- Pricing (input)
- Sol ~$4 / Terra ~$2 / Luna ~$0.20 per 1M tokens
- Pricing (output)
- Sol ~$20 / Terra ~$12 / Luna ~$1.20 per 1M tokens
Supported modalities
Text · Image · Audio
Context window
1.1M (1,050,000 tokens)
Pricing
Aug 21 2026: Sol promotional API pricing −20% input / −33% output ($4/$20) through at least Nov 21, 2026. Jul 30 2026: Luna −80%, Terra −20%. Sol Fast mode ~2× Sol price. Ultrafast (Aug 13 preview) has no published price yet—verify OpenAI pricing.
Availability
API: gpt-5.6-sol / terra / luna. ChatGPT (Aug 6 2026): Plus/Pro use updated Sol with a reasoning slider; Free/Go default to Luna (unlimited text chats rolling out). Ultrafast Sol (Aug 13): limited API preview via Cerebras (~14× Standard / up to ~750 tok/s); pricing and GA not published. Codex/Work remain on prior July API builds unless OpenAI says otherwise. OpenAI notified SpaceX it intends to stop supplying OpenAI models to Cursor (proposed shutoff 2026-11-12); ChatGPT, API, and Codex are unchanged. GPT-6 Astra (gpt-6-astra) launched 2026-09-03 as the new OpenAI peak on a staged ChatGPT/API/AWS rollout — this slug remains the GPT-5.6 family.
Use cases
- Complex multi-step reasoning and research
- Software engineering and code agents
- Multimodal analysis (docs, charts, images)
- Tool-using assistants and MCP workflows
Strengths
- Clear Sol / Terra / Luna routing for capability vs cost
- Mature tool-calling and structured output ecosystem
- Broad third-party integrations
Limitations
- Closed weights; limited inspectability
- Sol cost and rate limits at high volume
- Vendor lock-in for proprietary features
Related guides
Related benchmarks
Related research
Related GitHub
Related tools
Related rankings
Related comparisons
Companies
Explore more models
- GPT-6 AstraOpenAIOpenAI’s GPT-6 Astra peak model (API id gpt-6-astra) for computer use, coding agents, professional artifacts, science, and cybersecurity-adjacent defender workflows. Staged launch 2026-09-03; GPT-5.6 Sol/Terra/Luna remain the volume and cost-routing family.
- Claude FableAnthropicAnthropic’s Claude Fable 5.1 — the most capable widely released Claude for long-horizon agents, deep reasoning, and demanding coding workflows. Mythos 5.1 is the same underlying model with relaxed cyber/life-sciences safeguards for trusted-access programs.
- Claude OpusAnthropicAnthropic’s Claude Opus 5 tier for complex agentic coding, enterprise work, long-context analysis, and careful instruction following. Claude Fable 5.1 sits above Opus for peak widely released capability.
- Claude SonnetAnthropicAnthropic’s Claude Sonnet 5 tier — best combination of speed and intelligence for most production agents and coding, at lower cost than Opus.
- Gemini 3.1 ProGoogleGoogle’s current Pro-class Gemini for hard reasoning and native multimodal work. Prefer API id gemini-3.1-pro-preview; Gemini 3.5 Pro remains partner-testing. Legacy gemini-2.5-pro is scheduled for shutdown Oct 16, 2026.
- Muse SparkMetaMeta Superintelligence Labs’ Muse Spark 1.3 — closed multimodal reasoning model for agentic tasks, long-horizon coding, computer use, and 1M-context workflows via Muse Code and the Meta Model API. Succeeds Spark 1.2; max reasoning is still withheld for safety testing.