메뉴
BL
The Decoder 28일 전

오픈AI, 단일 최상위 모델 전략 벗어난 GPT-5.6 프로 3종 공개

IMP
8/10
핵심 요약

오픈AI의 새로운 논문에 따르면 기존의 단일 최상위 모델이었던 ChatGPT Pro(프로) 체제를 변경하여 GPT-5.6 모델에 '루나 프로(Luna Pro)', '테라 프로(Terra Pro)', '솔 프로(Sol Pro)' 등 세 가지 버전을 도입할 것으로 보입니다. 이를 통해 사용자는 작업의 특성에 맞춰 처리 속도, 처리량(Throughput), 최대 추론 능력 중 최적의 옵션을 선택할 수 있게 되었습니다. 다만 해당 모델들이 실제 ChatGPT 서비스에 적용될지는 아직 명확히 공개되지 않았습니다.

번역된 본문

오픈AI 논문, GPT-5.6을 위한 세 가지 프로 모델 공개... 단일 최상위 모델 전략 폐지 작성자: Maximilian Schreiner (2026년 7월 1일)

핵심 요약

  • 오픈AI의 논문에서 처음으로 GPT-5.6용 세 가지 프로 모델인 루나 프로(Luna Pro), 테라 프로(Terra Pro), 솔 프로(Sol Pro)가 언급되었습니다.
  • 지금까지 프로(Pro)는 항상 단일 최상위 모델을 의미했습니다.
  • 프로 유저들은 조만간 속도, 처리량(Throughput), 최대 추론 능력 중에서 선택할 수 있게 될 수도 있습니다.
  • 다만 이 논문에는 해당 라인업이 실제 ChatGPT에 출시될지에 대한 여부가 명시되어 있지 않으며, 프로 버전의 토큰 사용량 또한 공개되지 않았습니다.

오픈AI의 벤치마크 논문에 따르면, GPT-5.6의 프로(Pro) 등급이 세 가지 변형으로 출시될 가능성이 시사되고 있습니다. 이는 ChatGPT Pro 플랜이 출시된 이래 처음으로 이루어지는 대대적인 구조적 변화가 될 것입니다.

오픈AI는 지난 6월 말 GPT-5.6 세대를 공식적으로 발표하며 이를 세 가지 모델로 분할했습니다. 솔(Sol)은 가장 어려운 작업을 처리하고, 테라(Terra)는 대규모 비즈니스 작업량을 목표로 하며, 루나(Luna)는 더 빠르고 저렴한 일상적인 질의를 담당합니다. 당시 발표에서는 프로 변형에 대한 언급이 없었습니다.

하지만 최근 유전체학(Genomics) 벤치마크에 관한 새로운 오픈AI 논문에서 처음으로 프로 모델들이 공개되었습니다. 결과 표에는 'GPT-5.6 루나 프로(Luna Pro)', '테라 프로(Terra Pro)', '솔 프로(Sol Pro)' 행이 포함되어 있으며, 각각 '프로(확장형)(Pro Extended)' 실행으로 표기되어 있습니다.

프로(Pro)는 더 이상 단일 최상위 모델이 아닙니다

이번 벤치마크에서 솔 프로(Sol Pro)는 31.5%의 통과율(pass rate)을 기록하며 테스트된 총 60개 모델 중 가장 강력한 성능을 보여주었습니다. 이는 28.7%를 기록한 표준 솔(Sol) 모델과 16.0%를 기록한 비-GPT 최고 점수 모델인 클로드 오퍼스 4.8(Claude Opus 4.8)을 모두 뛰어넘는 수치입니다. 여기서 통과율은 모델이 오류 없이 전체 다단계 분석을 완료하고 올바른 최종 답변에 도달하는 빈도를 측정하는 지표입니다.

지금까지 ChatGPT Pro는 단순히 사용 가능한 단일 최고 모델로서, 다른 모든 모델보다 한 단계 높은 위치에 있었습니다. 그러나 이번 논문은 이러한 구조가 변화하고 있음을 시사합니다. 표준 GPT-5.6 라인업을 반영하듯 빠른 처리, 대용량 처리, 최대 성능 등 세 가지 병렬 프로 변형이 모델 라인업으로 제시되었습니다.

각 표준 모델의 최고 추론 설정('최대')과 프로 변형을 비교해 보면 성능 향상의 차이를 알 수 있습니다. 모든 수치는 전체 129개 작업에 대한 통과율을 나타냅니다.

모델 등급 표준 (최대) 프로 (확장형) 격차
GPT-5.6 루나(Luna) 16.5% 23.6% +7.1%p
GPT-5.6 테라(Terra) 23.3% 28.5% +5.2%p
GPT-5.6 솔(Sol) 28.7% 31.5% +2.8%p

이 결과를 보면 하위 등급으로 적용될수록 프로 버전으로 얻는 향상 폭은 줄어드는 경향이 있습니다. 루나 프로(Luna Pro)는 표준 버전 대비 7%p 포인트 가까이 크게 향상된 반면, 솔 프로(Sol Pro)의 향상 폭은 3%p 미만에 그쳤습니다.

추가적인 컴퓨팅 투자는 상대적으로 성능이 낮은 등급을 더 크게 끌어올리는 효과가 있습니다. 테라 프로(Terra Pro)는 28.5%를 기록하며 28.7%인 표준 솔(Sol) 모델과 거의 비슷한 수준을 보여주었습니다. 이는 곧 대용량 처리에 특화된 프로 변형이 표준 플래십 모델의 최상위 버전과 거의 맞먹는 성능을 발휘한다는 것을 의미합니다.

기존 프로 운영 방식과의 결별

이러한 세분화는 ChatGPT Pro가 출시된 이래 처음으로 이루어지는 프로 제공 방식의 대대적인 변화입니다. 단일 비싸고 최상위 등급이던 기존 방식에서 벗어나, 프로는 사용자가 당면한 작업에 따라 속도, 처리량, 최대 추론 능력 중에서 선택할 수 있는 독자적인 3개 모델 라인업으로 거듭날 것입니다.

다만, 이러한 계층적 구조가 실제 ChatGPT에 적용될지는 이 논문만으로는 명확하지 않습니다. 현재로서는 해당 명칭이 벤치마크 표에만 등장하기 때문입니다.

한 가지 세부 사항은 여전히 숨겨져 있습니다. 논문은 표준 GPT 모델에 대해 컴퓨팅 비용의 대략적인 척도로 평균 토큰 사용량을 보고하는데, 예를 들어 최고 설정에서 실행되는 솔(Sol) 모델의 경우 약 33,200개의 토큰을 사용한다고 명시했습니다. 그러나 프로 버전의 경우 해당 수치가 빠져있습니다. 저자들은 비교 가능한 토큰 통계가 없다고 밝혔지만, 실질적인 이유는 오픈AI가 단지 해당 수치를 공개하고 싶지 않아서일 가능성이 큽니다.

원문 보기
원문 보기 (영어)
OpenAI paper reveals three GPT-5.6 Pro models, breaking with single top-tier strategy Maximilian Schreiner View the LinkedIn Profile of Maximilian Schreiner Jul 1, 2026 Nano Banana Pro prompted by THE DECODER Key Points An OpenAI paper lists three Pro models for GPT-5.6 for the first time: Luna Pro, Terra Pro, and Sol Pro. Until now, Pro was always a single top-tier model. Pro users may soon be able to choose between speed, throughput, and maximum reasoning power. The paper doesn't say whether this lineup will actually ship in ChatGPT, and token usage for the Pro runs stays undisclosed. Ask about this article… Search An OpenAI benchmark paper suggests that the Pro tier of GPT-5.6 could ship in three variants. That would be the first major change to ChatGPT Pro's structure since the plan launched. OpenAI officially unveiled the GPT-5.6 generation in late June , splitting it into three models. Sol handles the hardest tasks, Terra targets high-volume business workloads, and Luna covers faster, cheaper everyday queries. Pro variants weren't part of the announcement. Now a new OpenAI paper on a genomics benchmark reveals Pro models for the first time. The results table includes rows for "GPT-5.6 Luna Pro," "Terra Pro," and "Sol Pro," each labeled as "Pro (Extended)" runs. Ad Pro is no longer just one top-tier model In the benchmark, Sol Pro hits a pass rate of 31.5 percent, making it the strongest of all 60 tested models. It beats the standard Sol at 28.7 percent and the best non-GPT score, Claude Opus 4.8 at 16.0 percent. The pass rate measures how often a model completes the full multi-step analysis without errors and arrives at the correct final answer. Ad DEC_D_Incontent-1 Until now, ChatGPT Pro was simply the single best model available, being one tier above everything else. The paper suggests that's changing. It lists three parallel Pro variants that mirror the standard GPT-5.6 lineup: a fast one, a high-volume one, and a max-performance one. Comparing each standard tier at its highest reasoning setting ("max") against its Pro variant shows how the gains play out. All values are pass rates on the full 129-task suite: Ad Model tier Standard (max) Pro (Extended) Gap GPT-5.6 Luna 16.5% 23.6% +7.1 points GPT-5.6 Terra 23.3% 28.5% +5.2 points GPT-5.6 Sol 28.7% 31.5% +2.8 points In this case, the Pro boost shrinks as you move up the ladder. Luna Pro gains a full seven points over its standard version, while Sol Pro picks up less than three. Extra compute lifts weaker tiers more: Terra Pro lands at 28.5 percent, nearly matching standard Sol at 28.7 percent, which means a high-volume Pro variant performs almost as well as the best standard flagship. A break from how Pro has always worked This split would be the first major change to the Pro offering since ChatGPT Pro launched . Instead of one expensive top tier, Pro could become its own three-model lineup where users pick between speed, throughput, and maximum reasoning power based on the task at hand. Ad DEC_D_Incontent-2 Whether this tiered structure will actually show up in ChatGPT isn't clear from the paper. The names come only from the benchmark table so far. Ad One detail stays hidden, too. For the standard GPT models, the paper reports average token usage as a rough proxy for compute cost, about 33,200 tokens for Sol at its highest setting. For the Pro runs, that number is missing. The authors say no comparable token accounting was available, but the more likely explanation is that OpenAI simply doesn't want to share those figures. AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: OpenAI