메뉴
BL
The Decoder 32일 전

OpenAI's GPT-5.6 Sol launches to rival Claude Mythos under government access rules it calls unsustainable

IMP
3/10
핵심 요약

[요약 오류] OpenAI's GPT-5.6 Sol launches to rival Claude Mythos under government access rules it calls unsustainable

원문 보기
원문 보기 (영어)
OpenAI's GPT-5.6 Sol launches to rival Claude Mythos under government access rules it calls unsustainable Matthias Bastian View the LinkedIn Profile of Matthias Bastian Jun 26, 2026 Nano Banana Pro prompted by THE DECODER Key Points OpenAI's new GPT-5.6 generation includes the flagship Sol and two cheaper tiers, Terra and Luna. Sol matches or beats Anthropic's Claude Mythos 5 across benchmarks, with a clear lead in agentic coding and better token efficiency in cybersecurity. The US government is restricting access to select partners for now. OpenAI says the policy hurts developers and businesses. Ask about this article… Search OpenAI's new flagship GPT-5.6 Sol claims a lead over Anthropic's Claude Mythos in agentic coding and goes toe to toe with it in cybersecurity. Access stays limited to a handful of partners for now. OpenAI has unveiled GPT-5.6 Sol, a new generation of models built to compete with Claude's Mythos class . The limited preview is only open to select partners through the API and Codex, at the explicit direction of the US government . The same government previously yanked Anthropic's Mythos-class model Fable 5 off the market . OpenAI isn't subtle about its frustration . "We don’t believe this kind of government access process should become the long-term default. It keeps the best tools from users, developers, enterprises, cyber defenders, and global partners who need them." Ad GPT-5.6 also brings a new layered naming scheme that looks a lot like Claude's. The number (x.6) marks the generation, while Sol, Terra, and Luna are permanent performance tiers that can evolve on their own. Sol is the flagship. Terra matches GPT-5.5 at half the cost. Luna is the budget option. On top of that, there's a "max" mode for deeper reasoning and an "ultra" mode that farms out complex tasks to sub-agents running in parallel. Ad DEC_D_Incontent-1 Sol edges past Claude Mythos in agentic coding OpenAI's benchmark numbers put Sol ahead of Anthropic's Claude Mythos 5 in agentic coding. On Terminal-Bench 2.1, Sol scores 88.8 percent. Sol Ultra hits 91.9, Claude Mythos 5 lands at 88 percent, and Fable 5 trails at 84.3. Sol also shows gains in biology. On GeneBench v1, a benchmark for genomics and quantitative biology, it beats GPT-5.5 (30 percent vs. 22 percent best case) while burning fewer tokens. Ad On ExploitBench , which tests how well AI agents can find and exploit real security flaws in Google's V8 JavaScript engine all the way to full code execution, Sol matches Mythos Preview's performance while using roughly a third of the output tokens, OpenAI says. On ExploitGym , a benchmark built by UC Berkeley researchers with OpenAI and other labs, all three GPT-5.6 models get better as reasoning effort goes up. That points to room for scaling with more compute. Claude numbers for this benchmark aren't available yet. Ad DEC_D_Incontent-2 OpenAI calls Sol its most capable cybersecurity model yet but frames it as a defender, not an attacker. The model is better at spotting and fixing flaws than at running full end-to-end attacks on its own, the company says. Mythos pulled that off in a different benchmark . Ad In tests with Chromium and Firefox, Sol found bugs and exploitation primitives but never produced an autonomous full-chain exploit. OpenAI says GPT-5.6 Sol is still below the "Cyber Critical" threshold in its Preparedness Framework . Pricing, availability, and a Cerebras launch in July Per million tokens, OpenAI charges $5 input and $30 output for Sol, $2.50 and $15 for Terra, and $1 and $6 for Luna. The company has also revamped its prompt caching system with explicit cache breakpoints and a guaranteed minimum lifetime of 30 minutes. Cache writes cost 1.25x the regular input price. Cache reads still get a 90 percent discount. Since Sol uses fewer tokens to match or beat competitors across several benchmarks, the effective cost per task could end up lower than previous generations. That would push back against the trend of AI models getting pricier with each release , a frequent criticism lately, and a competitive weak spot against cheaper Chinese models . In July, Sol is set to go live on Cerebras at up to 750 tokens per second. AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: OpenAI
관련 소식
TC
TechCrunch AI 32일 전
IMP 9

미 정부 요청으로 OpenAI, GPT-5.6 출시 제한

미국 트럼프 행정부의 요청에 따라 OpenAI가 차세대 AI 모델인 GPT-5.6의 출시를 소수의 신뢰할 수 있는 파트너로 제한했습니다. 이는 최첨단 AI 모델에 대한 정부의 강력한 사전 검토 및 통제가 시작됨을 의미하며, AI 업계는 이러한 규제가 혁신을 저해할 수 있다고 우려하고 있습니다.

OpenAI GPT-5.6 AI 규제
HN
Hacker News 32일 전
IMP 8

미국 정부, GPT-5.6 사용 권한 통제 예정

최신 AI 모델로 예상되는 GPT-5.6의 사용자 접근 권한을 미국 정부가 직접 통제하려는 움직임을 보이고 있습니다. 이는 강력한 인공지능 기술의 안전과 국가 안보를 확보하기 위한 규제의 일환으로 풀이됩니다. 따라서 향후 최고 성능의 모델을 활용하려는 개발자나 기업은 정부 승인이라는 새로운 허들을 고려해야 할 전망입니다.

AI 규제 정책 미국 정부
TC
TechCrunch AI 32일 전
IMP 8

OpenAI, 우버 인도 총괄 영입…미국 외 최대 시장 공략

OpenAI가 미국 다음으로 가장 큰 핵심 시장인 인도의 본격적인 사업 확장을 위해 우버 인도 및 남아시아 총괄이었던 프라브짓 싱(Prabhjeet Singh)을 초대 이사역(MD)으로 영입했습니다. 그는 소비자 성장, 기업 도입, 파트너십 및 규제 대응을 총괄하며, 이는 막대한 인구와 개발자를 바탕으로 경쟁이 치열해지는 인도 AI 시장에서 OpenAI의 영향력을 강화하기 위한 핵심 전략으로 평가됩니다.

OpenAI 인도 글로벌
HN
Hacker News 32일 전
IMP 1

GPT-5.6 Sol 차세대 모델 살펴보기

이 게시물은 다음 세대 AI 모델로 예상되는 'GPT-5.6 Sol'의 정보를 해커뉴스(Hacker News)를 통해 미리 공개하는 내용입니다. 본문 전체가 0, 5, 6, 점(.)으로 이루어진 추상적인 기호로 작성되어 있어, 단순한 장난(April Fools)이나 암호화된 이스터에그일 가능성이 높습니다. 실무자들에게는 기술적 참고 자료라기보다는 커뮤니티 내의 밈(Meme) 또는 퍼포먼스로 해석할 수 있습니다.

AI 모델 해커뉴스 가짜 뉴스
MR
MIT Tech Review 32일 전
IMP 5

The Download: brain-melting heatwaves and unprecedented OpenAI restrictions

TC
TechCrunch AI 33일 전
IMP 8

안전 우려로 백악관, OpenAI에 신모델 출시 연기 요청

트럼프 행정부는 안전상의 우려로 인해 OpenAI의 최신 모델을 일반 대중이 아닌 선별된 소수 파트너에게만 우선 공개하도록 요구했습니다. 이는 막강한 AI 모델이 사이버 범죄에 악용되는 것을 막기 위해 최근 정부가 AI 모델에 대한 연방 차원의 통제 및 사전 검토를 강화하고 있는 중요한 정책적 변화를 보여줍니다.

오픈AI 정책 안전