跳到正文
北京时间
原文
The Decoder:AI News(RSS)· Maximilian Schreiner·· 2026-07-01精选AI 评分70

OpenAI论文揭示GPT-5.6三个Pro变体,打破单一顶级策略

OpenAI paper reveals three GPT-5.6 Pro models, breaking with single top-tier strategy

AI 导读

OpenAI论文首次列出GPT-5.6的三个Pro变体:Luna Pro、Terra Pro和Sol Pro,取代以往单一Pro模式。在基因组学基准中,Sol Pro通过率31.5%居60个测试模型之首,领先标准Sol(28.7%)和Claude Opus 4.8(16.0%)。Pro相比标准版本提升逐级递减:Luna Pro提升7.1个百分点(16.5%→23.6%),Terra Pro提升5.2(23.3%→28.5%),Sol Pro仅提升2.8(28.7%→31.5%)。Terra Pro(28.5%)几乎与标准Sol(28.7%)持平。论文未披露Pro运行的token用量,也不清楚该分层是否会在ChatGPT中实际推出。

推荐理由

论文意外曝光 GPT-5.6 Pro 将有三个变体,Pro 不再只是一个最强模型,而是让用户按推理需求选版本,这才是匹配 200 美元月费该有的逻辑。

正文 · 原文

Image description

Key Points

  • An OpenAI paper lists three Pro models for GPT-5.6 for the first time: Luna Pro, Terra Pro, and Sol Pro. Until now, Pro was always a single top-tier model.
  • Pro users may soon be able to choose between speed, throughput, and maximum reasoning power.
  • The paper doesn't say whether this lineup will actually ship in ChatGPT, and token usage for the Pro runs stays undisclosed.

An OpenAI benchmark paper suggests that the Pro tier of GPT-5.6 could ship in three variants. That would be the first major change to ChatGPT Pro's structure since the plan launched.

OpenAI officially unveiled the GPT-5.6 generation in late June, splitting it into three models. Sol handles the hardest tasks, Terra targets high-volume business workloads, and Luna covers faster, cheaper everyday queries. Pro variants weren't part of the announcement.

Now a new OpenAI paper on a genomics benchmark reveals Pro models for the first time. The results table includes rows for "GPT-5.6 Luna Pro," "Terra Pro," and "Sol Pro," each labeled as "Pro (Extended)" runs.

Pro is no longer just one top-tier model

In the benchmark, Sol Pro hits a pass rate of 31.5 percent, making it the strongest of all 60 tested models. It beats the standard Sol at 28.7 percent and the best non-GPT score, Claude Opus 4.8 at 16.0 percent. The pass rate measures how often a model completes the full multi-step analysis without errors and arrives at the correct final answer.

Until now, ChatGPT Pro was simply the single best model available, being one tier above everything else. The paper suggests that's changing. It lists three parallel Pro variants that mirror the standard GPT-5.6 lineup: a fast one, a high-volume one, and a max-performance one.

Comparing each standard tier at its highest reasoning setting ("max") against its Pro variant shows how the gains play out. All values are pass rates on the full 129-task suite:

In this case, the Pro boost shrinks as you move up the ladder. Luna Pro gains a full seven points over its standard version, while Sol Pro picks up less than three. Extra compute lifts weaker tiers more: Terra Pro lands at 28.5 percent, nearly matching standard Sol at 28.7 percent, which means a high-volume Pro variant performs almost as well as the best standard flagship.

A break from how Pro has always worked

This split would be the first major change to the Pro offering since ChatGPT Pro launched. Instead of one expensive top tier, Pro could become its own three-model lineup where users pick between speed, throughput, and maximum reasoning power based on the task at hand.

Whether this tiered structure will actually show up in ChatGPT isn't clear from the paper. The names come only from the benchmark table so far.

One detail stays hidden, too. For the standard GPT models, the paper reports average token usage as a rough proxy for compute cost, about 33,200 tokens for Sol at its highest setting. For the Pro runs, that number is missing. The authors say no comparable token accounting was available, but the more likely explanation is that OpenAI simply doesn't want to share those figures.

来源:The Decoder:AI News(RSS) · the-decoder.com