New

Seedream 5.0 Pro is here — ByteDance's most powerful image editing model, now on Fullmira.

Try it free
Released July 9, 2026 · Sol · Terra · Luna

Use GPT-5.6 free — all three tiers

Chat with Sol, Terra and Luna right here. Switch tiers mid-conversation, compare them against Claude and Gemini, and pay for one subscription instead of five.

GPT-5.6 Sol
Ask me anything — I'm running as Sol, the flagship tier. Switch to Terra or Luna above and ask the same question to feel the difference in speed and depth.
The GPT-5.6 series is temporarily unavailable for maintenance. Please check back later.
Typically 1–5 credits per message, based on length
Want file uploads, saved history and every other model?Open GPT-5.6 in the full studio
No watermarkNo credit cardSwitch tiers anytime
The family

Sol, Terra, Luna — which one do you need?

The number marks the generation; the names are capability tiers that advance on their own cadence. Most people should start on Terra.

FLAGSHIP

GPT-5.6 Sol

For your hardest work

State of the art on coding, computer use and long-horizon agentic tasks — and it gets there with fewer tokens than the models it beats.

  • 88.8% on Terminal-Bench 2.1
  • 62.6% on OSWorld 2.0 computer use
  • Supports max and ultra reasoning

GPT-5.6 Terra

The everyday default

Half the flagship price while performing competitively with the older GPT-5.5. For most day-to-day work you won't notice the gap.

  • 87.4% on Terminal-Bench 2.1
  • Beats GPT-5.5 at a lower cost
  • Best balance of price and quality

GPT-5.6 Luna

Fast and cheap

The fastest and most affordable of the three, and still ahead of several previous-generation frontier models. Built for volume.

  • 84.7% on Terminal-Bench 2.1
  • Cheapest tier by a wide margin
  • Ideal for drafts and bulk tasks
Model comparison

GPT-5.6 vs GPT-5.5, Claude and Gemini

Published benchmark results, including the ones where GPT-5.6 doesn't come first. Higher is better in every row.

BenchmarkGPT-5.6 SolGPT-5.6 TerraGPT-5.6 LunaGPT-5.5Claude Opus 4.8Gemini 3.1 Pro
Terminal-Bench 2.1· CLI workflows88.8%87.4%84.7%85.6%78.9%70.7%
OSWorld 2.0· computer use62.6%50.2%45.6%47.5%54.8%
BrowseComp· agentic browsing90.4%87.5%83.3%84.4%84.3%85.9%
Agents' Last Exam· long-horizon work52.7%50.4%50.3%46.9%45.2%32.1%
SWE-Bench Pro· real codebases64.6%63.4%62.7%59.4%69.2%54.2%
GPQA Diamond· graduate science94.6%92.9%92.3%93.6%92.0%94.3%
FrontierMath Tier 1–389.0%84.9%78.6%85.3%80.0%59.6%
MMMU Pro· multimodal83.0%80.7%78.4%81.2%80.5%
Toolathlon· tool use58.0%53.1%53.4%55.6%59.9%48.8%

Figures from OpenAI's GPT-5.6 announcement, July 9 2026. Note the honest picture: Claude Opus 4.8 still leads on SWE-Bench Pro and Toolathlon.

What changed

What GPT-5.6 does that 5.5 couldn't

More work per token

The headline change isn't just a higher score — it's reaching that score with far fewer output tokens, which lowers your cost per finished task.

Parallel agents with ultra

The ultra setting runs four agents at once by default, splitting a demanding task across parallel workstreams to reach a stronger answer.

Real design judgement

Given only high-level direction it produces interfaces that hold together, then inspects the rendered result and fixes visual problems before handing it back.

Decks, docs and models

Reads a reference deck's layouts, typography and spacing rules, then applies them consistently to new material, including equations and financial models.

Long context that holds

Substantially better recall across very long inputs — 90.7% on GraphWalks BFS at 256k, against 73.7% for GPT-5.5.

Stronger safeguards

Ships with layered protections and a reasoning monitor. Note: safeguards are tuned conservatively, so some benign requests get caught.

Why use it here

GPT-5.6 on Fullmira vs everywhere else

The model is the same. What differs is what surrounds it.

On Fullmira

  • One subscription covers every model — GPT-5.6, Claude, Gemini, plus image, video and audio models.
  • Switch models mid-conversation. Re-run the same thread on another model without re-pasting your context.
  • Compare side by side. Send one prompt to Sol and Claude at once and judge the answers yourself.
  • Sign up free and get 500 credits instantly — no card needed.
  • Everything in one library — chats, images, video and audio saved together.

Single-vendor apps

  • A separate subscription for each provider you want access to.
  • Locked to one vendor's models — no switching when another does the job better.
  • Comparing means copying your prompt into a different tab, by hand.
  • Card details usually required before you can evaluate anything properly.
  • Your work spread across several apps with separate histories.

Which GPT-5.6 tier should you actually use?

The honest answer for most people is Terra. It costs half of Sol and performs competitively with GPT-5.5, which was a frontier model until recently. Unless your work is genuinely difficult, you will struggle to tell the difference in everyday writing, analysis and light coding.

When Sol earns its price

Reach for the flagship when the task is long-horizon and expensive to get wrong: refactoring across a real codebase, driving a computer through a multi-step workflow, or research that runs for hours. The gap is widest on computer use, where Sol scores 62.6% on OSWorld 2.0 against Terra's 50.2%.

When Luna is the right call

Volume work. Classification, summarising a hundred documents, first drafts you'll rewrite anyway. At $1 per million input tokens it's roughly a fifth of Sol's price, and it still scores 84.7% on Terminal-Bench 2.1.

About max and ultra

Max gives the model more time to reason, check itself and revise. Ultra goes further, coordinating four agents in parallel by default. A practical habit: draft on Terra, and escalate to Sol with max only when the first attempt isn't good enough.

Where GPT-5.6 isn't the winner

On SWE-Bench Pro, Claude Opus 4.8 scores 69.2% against Sol's 64.6%. On Toolathlon, Opus 4.8 edges ahead at 59.9% versus 58.0%. No single model wins everything — which is the whole argument for having them all in one place.

FAQ

GPT-5.6 questions, answered

What is GPT-5.6?
GPT-5.6 is OpenAI's model family released on July 9, 2026. It ships in three tiers: Sol (flagship), Terra (balanced everyday model) and Luna (fastest and cheapest).
Can I use GPT-5.6 for free?
Yes — sign up free and get 500 credits instantly, no credit card required, so you can try all three tiers before deciding on a paid plan.
What's the difference between Sol, Terra and Luna?
Sol is $5 per million input tokens and $30 output. Terra is $2.50 / $15. Luna is $1 / $6, fastest and cheapest. Most people should default to Terra and escalate to Sol only when a task is genuinely hard.
Is GPT-5.6 better than GPT-5.5?
On the published benchmarks, yes — and notably more efficient. Sol scores 88.8% on Terminal-Bench 2.1 against GPT-5.5's 85.6%, and 62.6% on OSWorld 2.0 against 47.5%.
What do the max and ultra settings do?
Max gives the model more time to reason and revise. Ultra coordinates four agents in parallel — trading higher token use for stronger results on demanding tasks.
How much does the GPT-5.6 API cost?
Per one million tokens: Sol is $5 input / $30 output, Terra is $2.50 / $15, and Luna is $1 / $6. Cached input reads keep a 90% discount.
Is GPT-5.6 better than Claude or Gemini?
It depends on the task. Sol leads on Terminal-Bench, OSWorld, BrowseComp and Agents' Last Exam. Claude Opus 4.8 is ahead on SWE-Bench Pro and Toolathlon. The table above shows all the rows including the ones GPT-5.6 loses.
Do I need a separate subscription for GPT-5.6?
No. One Fullmira account covers GPT-5.6 alongside every other model on the platform. Paid plans raise your limits across all of them at once.