Ranked #2 Everyday Ecosystem — The Leading AI Assistants
OpenAI

GPT-5.6

GPT-5.6 is not one louder chatbot. It is a three-model work crew inside a newly expanded ChatGPT: Sol for the jobs that deserve the expensive brain, Terra for most daily work, Luna for the flood. ChatGPT Work and the merged desktop app are the glue that turn that roster into a serious digital colleague.

Updated July 10, 2026 AgenticChatGPT WorkSol / Terra / Luna
9.9out of 10
Official Website
Best for

GPT-5.6 is not one louder chatbot. It is a three-model work crew inside a newly expanded ChatGPT: Sol for the jobs that deserve the expensive brain, Terra for most daily work, Luna for the flood. ChatGPT Work and the merged desktop app are the glue that turn that roster into a serious digital colleague.

Why It Wins

General availability across ChatGPT, Codex, and the API. Sol leads OpenAI's agentic coding and computer-use story; Terra brings GPT-5.5-competitive everyday work at half Sol's token price; Luna is the fast volume tier. ChatGPT Work can carry a project across connected apps, files, browser, and desktop. Ultra adds parallel agents for the hard jobs.

Watch out

The family is broadly available, but the useful knobs are not identical on every plan: Sol, max effort, and ultra have different access rules in ChatGPT Work and Codex. Long Work tasks consume more plan usage, connected apps need deliberate permissions, and Sol's stronger cyber safeguards can create friction on legitimate edge cases. Benchmarks show a strong lead in agentic coding, not a clean sweep of every coding test.

01

What It Actually Is

If GPT-5.5 was a very good colleague, GPT-5.6 is a small department with a traffic controller. Sol handles the hard, expensive thinking. Terra takes the normal work. Luna clears the queue. The important launch is not only those three names; it is the place OpenAI put them.

ChatGPT Work is now the agent for jobs too large to fit in a polite little chat bubble. Give it a project, source files, and connected apps; it can gather context, break work into steps, keep going in the background, then produce editable docs, slides, sheets, Sites, or web apps. Scheduled Tasks can keep repetitive work moving. You remain the editor, approver, and person who has to explain the spreadsheet at Monday’s meeting.

On desktop, the former Codex app is merging into the new ChatGPT desktop app. Codex is still the coding agent—now alongside Chat and Work—with inline diff editing, pull-request review, faster computer use, and multi-repository projects. The app also has a built-in browser, while Computer Use can click, type, and move files across your local apps when you approve the access. That is a much bigger deal than a fresh model badge: OpenAI is joining chat, work, code, and desktop actions into one surface.

Pick the brain, not a religion

Job Route it to Why
Complex investigation, serious codebase work, research with several moving parts Sol Highest capability and best agentic-coding results in the family
Analysis, writing, normal product work, ordinary multi-step agents Terra GPT-5.5-competitive work at half Sol’s API sticker price
Classification, extraction, drafting, and high-volume routine automation Luna Fastest and cheapest GPT-5.6 tier
One giant, parallelizable mission Sol + ultra Multiple agents work in parallel; use only when the extra work is worth it

The reason that routing matters is efficiency. OpenAI reports Sol at 53.6% on Agents’ Last Exam, 80 on the Artificial Analysis Coding Agent Index with max reasoning, and 62.6% on OSWorld 2.0. Those are impressive launch numbers, and the company also reports sharply lower time, tokens, or estimated cost than several competitors on selected tests. They are not magic beans. Benchmark setup, effort setting, tool harness, and the job itself all matter.

The headline usability change is availability: GPT-5.6 launched across ChatGPT, Codex, and the API. But the three-tier family is not a buffet with identical access. Chat users on Plus and above can use Sol at medium or higher effort; Work and Codex expose different plan-based choices; ultra is more restricted. Start by checking the model and effort controls you actually see, rather than buying a plan for a feature that has not reached your account yet.

The honest catch

More agency means more responsibility. A long Work task may consume far more included usage than a chat reply. Connected apps and Computer Use are useful precisely because they can reach real company context, which makes permissions and approval points non-negotiable. And Sol’s stronger safeguards can make legitimate security-adjacent work slower or fussier.

So the winning habit is not “always Sol.” It is Sol for the mountain, Terra for the commute, Luna for the conveyor belt—with ChatGPT Work as the project desk that keeps the three from stepping on each other.

02

Strengths and honest limitations

Key Strengths

  • ChatGPT Work is the product shift: It can break ambitious projects into steps, work for hours, and turn connected-app context into editable documents, slides, sheets, Sites, and web apps. This is not merely a new model picker; it is a new way to hand ChatGPT a workflow.
  • A sensible three-brain roster: Sol is the flagship for difficult, long-horizon work; Terra is the lower-cost everyday model with GPT-5.5-competitive performance; Luna is the fastest, most affordable tier. The API price ladder is equally clear: Sol $5/$30, Terra $2.50/$15, Luna $1/$6 per million input/output tokens.
  • More force only when it pays: max gives GPT-5.6 more room to reason, check, and revise. ultra coordinates multiple agents in parallel—four by default—so a large job can move on several fronts instead of making one agent pace the hallway.
  • Computer use and artifact quality finally matter: Sol posted 62.6% on OSWorld 2.0 in OpenAI’s release table, while the company says the family is markedly better at following document, spreadsheet, and presentation templates. For normal office work, that practical fidelity may matter more than another abstract IQ point.
  • Less orchestration glue for developers: In the Responses API, Programmatic Tool Calling lets GPT-5.6 write and run lightweight in-memory programs to coordinate tools and filter intermediate results. That means fewer model round trips on tool-heavy work, rather than simply throwing a larger context window at the mess.

Honest Limitations

  • Availability is broad; capability is tiered: Sol in Chat is for Plus, Pro, Business, and Enterprise users at medium or higher effort. Free and Go users get Terra in ChatGPT Work and Codex. Ultra is limited to Pro and Enterprise in Work, and Plus and higher in Codex.
  • Long-running work spends real allowance: ChatGPT Work uses the same usage structure as Codex. A quick request may be cheap; a task that researches, browses, revises files, and waits on approvals can use a meaningful share of a plan’s included agentic usage.
  • Sol is deliberately more cautious in sensitive domains: OpenAI says its Sol cyber safeguards block substantially more potentially harmful activity than earlier models. That is sensible for misuse prevention, but defensive security, systems, and biology-adjacent work can meet more review, delay, or refusal than users expect.
  • Do not read one chart as a universal crown: Sol leads several agentic and coding evaluations, but Claude Fable 5 still has the stronger published SWE-Bench Pro result in this leaderboard. Match the model to your real task instead of treating every benchmark as the same sport.
  • Connected work needs adult supervision: ChatGPT Work can touch valuable context across plugins, browser tabs, local files, and desktop apps. Keep permissions narrow, review important actions, and check the finished spreadsheet, deck, or site before sending it into the world.
03

Benchmark Snapshot

Agents' Last Exam — Sol 53.6%

OpenAI reports a new high on this long-horizon professional-workflow evaluation across 55 fields. Treat the comparison and cost estimates as launch-period results, not a substitute for testing your own workflow.

Artificial Analysis Coding Agent Index — Sol 80

Sol with max reasoning leads OpenAI's release comparison on this independent coding-agent index, 2.8 points above Claude Fable 5 in the company's table.

OSWorld 2.0 — Sol 62.6%

OpenAI's computer-use result. It is a useful signal for desktop and web task execution, but not proof that every app workflow will be reliable unattended.

API price ladder — $5/$30 · $2.50/$15 · $1/$6

Official input/output price per 1M tokens for Sol, Terra, and Luna. Cache reads receive a 90% input discount; finished-task cost still depends on routing and agent fan-out.

04

The Verdict

GPT-5.6 now takes #2 in the everyday ecosystem: Opus 5 has the stronger launch-week blend of independent intelligence, knowledge work, computer use, and price, while ChatGPT remains the broadest all-round work system. ChatGPT Work, plugins, browser, Computer Use, Sites, and desktop Codex make Sol useful from research through artifact and code. Choose GPT-5.6 when one integrated platform matters most; choose Opus 5 when judgment per dollar is the deciding factor.

05

Frequently Asked Questions