Monday Jun 29

OpenAI Previews GPT-5.6 Sol

26JUN
ULTRA MODEFLAGSHIPSUBAGENTS

OpenAI teased its strongest model yet. GPT-5.6 comes in three sizes, with an ultra mode that runs subagents. Best coding and cyber scores yet, shipping in weeks, not today.

The preview is limited. Only trusted partners get GPT-5.6 right now, and OpenAI told the government first.

Three models share the family. Sol is the flagship, Terra the daily driver, Luna the cheap one. The new ultra mode goes past a single agent and leans on subagents for hard work.

Sol set a new top score on Terminal-Bench for command-line tasks. All three rate high on bio and cyber, so the safety stack got heavier too.

full brief & sources

Why this matters

  • First look at the model meant to fix the reward-audit problems behind the Goblin Incident.
  • Ultra mode signals OpenAI is baking multi-agent orchestration into the base model, not bolting it on.
  • High bio and cyber ratings mean tighter access controls for every enterprise buyer.

🔍 What happened

  • June 26: OpenAI previewed GPT-5.6 Sol, Terra, and Luna.
  • Sol is the flagship, Terra balanced, Luna fast and cheap.
  • A new max reasoning effort lets Sol think longer.
  • A new ultra mode uses subagents to accelerate complex work.
  • Sol sets state of the art on Terminal-Bench 2.1.
  • General availability is planned in coming weeks; preview limited to trusted partners shared with government.

💬 Smart takes

  • OpenAI: Sol is its strongest and most capable cybersecurity model yet.
  • Skeptic: a limited preview is not a launch, and GPT-5.6 already slipped past its June window.

🧭 Where this goes

  1. Likelygeneral availability lands in July after the IPO quiet period and reward-audit validation.
  2. Likelyrivals copy the in-model subagent pattern within two quarters.
  3. Possiblehigh bio and cyber ratings trigger stricter enterprise gating and a slower rollout.
  4. Wild Cardultra mode makes single-agent pricing obsolete and resets how API cost is billed.

🥄 The Spoon Take

The model is becoming the orchestrator. OpenAI is putting subagents inside GPT-5.6 instead of leaving them to outside frameworks. If that holds, the agent-orchestration layer everyone is building gets absorbed into the model itself. The interesting fight is no longer the model. It is who owns the loop.

🤔 Pushback

A preview shared only with trusted partners tells us little about real cost, latency, or whether ultra mode beats a well-built external agent loop.