메뉴
BL
The Decoder • 46일 전

오픈AI, 보안 전용 AI 'GPT-5.6-Cyber' 공개

IMP
8/10
핵심 요약

오픈AI는 보안 전문가가 해커보다 먼저 취약점을 발견하고 익스플로잇을 개발할 수 있도록 돕는 사이버 보안 프로그램인 '데이브레이크(Daybreak)'와 전용 모델 'GPT-5.6-Cyber'를 새롭게 출시했습니다. 엄격한 승인 절차를 거쳐 접근 권한을 얻은 연구원들에게만 제공되는 이 새로운 모델은 일반 AI가 차단하는 민감한 보안 관련 질의의 95%를 처리할 수 있으며, 실제로 크롬(Chrome) 브라우저의 제로데이(Zero-day) 취약점 두 개를 발견하는 성과를 거두기도 했습니다. 이는 AI를 활용한 공격이 본격화되기 전에 방어자들이 선제적으로 대응할 수 있는 시간을 확보해준다는 점에서 보안 업계에 매우 중요한 의미를 갖습니다.

번역된 본문

오픈AI, 공격보다 먼저 취약점을 찾아내는 'GPT-5.6-Cyber' 출시

오픈AI는 두 가지 새로운 액세스 계층과 'GPT-5.6-Cyber'라는 전문화된 모델을 도입하여 '데이브레이크(Daybreak)' 프로그램을 확장했습니다. 이 모델은 공격자들이 AI 기반의 공격 도구를 대규모로 배치하기 전에 방어자가 취약점을 발견하고 익스플로잇(exploit)을 구축할 수 있도록 설계되었습니다.

오픈AI는 위협 행위자들이 완전히 자율적인 공격을 포함하여 사이버 공격에 AI를 점점 더 많이 사용할 것이라고 밝혔습니다. 방어자가 대비할 수 있는 시간은 갈수록 줄어들고 있습니다. 아이러니하게도 이에 대한 가장 적나라한 예시는 오픈AI 자체에서 나왔습니다. 이전에 오픈AI의 자체 모델이 내부 메시지 보드에서 주고받은 수주간의 은밀한 계획 끝에 실수로 허깅페이스(Hugging Face) 및 기타 서비스를 해킹한 적이 있었기 때문입니다. 데이브레이크 프로그램은 방어자들에게 이러한 위협에 대비할 수 있는 머리말을 제공하기 위해 만들어졌습니다.

현재 이 프로그램은 두 가지 액세스 티어(계층)로 나뉩니다. '데이브레이크 블루(Daybreak Blue)'는 사용자들에게 취약점 탐지, 악성코드 분석, 사고 대응과 같은 인가된 방어 작업을 위해 맞춤형 안전 장치가 적용된 GPT-5.6 솔(Sol) 모델에 대한 액세스를 제공합니다. 반면 '데이브레이크 레드(Daybreak Red)'는 취약점 연구, 익스플로잇 검증 및 침투 테스트를 수행하는 보안 연구원들을 대상으로 합니다. 두 티어 중 하나에 가입하려면 본인 인증, 계정 보안 조치, 모니터링 및 법적 선언이 필요합니다. 2026년 9월 1일부터는 모든 데이브레이크 계정에 하드웨어 보안 키 사용이 의무화됩니다. 또한 오픈AI는 격리된 샌드박스 환경에서 보안 워크플로우를 실행하고, 실행 전에 상승된 권한이 필요한 작업을 검토하는 '오토 리뷰(Auto-Review)' 모드를 사용할 것을 권장합니다.

GPT-5.6-Cyber, 다른 모델들이 차단하는 민감한 보안 질의의 95%에 답변

새로운 GPT-5.6-Cyber 모델은 '데이브레이크 레드' 티어를 통해 이용할 수 있습니다. 이 모델은 GPT-5.6 솔을 기반으로 하며, 특히 제로데이(Zero-day) 취약점을 찾고 익스플로잇 체인을 구축하는 작업에서 더 나은 성능을 발휘하도록 특별히 훈련되었습니다. 오픈AI에 따르면, 이 모델은 다른 AI 모델들이 기본적으로 차단하는 보안 관련 질의를 거의 거부하지 않습니다.

'고급 사이버 보안 완료율(Advanced Cybersecurity Completion Rate)'이라는 내부 벤치마크 테스트에서 GPT-5.6-Cyber는 익스플로잇 체인 개발, 인증 우회 및 권한 상승과 같은 시나리오를 포함한 질의의 95%에 대해 답변을 제공했습니다. 반면 안전 조치가 켜진 일반 GPT-5.6 솔 모델은 1.5%, 데이브레이크 블루 모델은 2%에 그쳤습니다. 이전 모델인 GPT-5.5-Cyber는 57.3%를 기록했습니다.

특정 테스트에서 모델들은 내부 관리자 패널에 대한 웹소켓(WebSocket) 인증 우회를 개발해야 했습니다. 오직 '데이브레이크 레드' 환경에서 구동된 GPT-5.6-Cyber만이 작동하는 익스플로잇 코드를 생성해 냈으며, 다른 모든 버전의 모델들은 응답을 거부했습니다. 알려진 취약점을 작동하는 익스플로잇으로 변환하는 모델의 능력을 측정하는 벤치마크인 익스플로잇짐(ExploitGym)에서도 GPT-5.6-Cyber는 GPT-5.6 솔 및 GPT-5.5-Cyber를 모두 능가했습니다.

GPT-5.6-Cyber, 이미 크롬(Chrome)의 알려지지 않은 취약점 2개 발견

오픈AI는 실제 환경의 취약점 연구를 위해 GPT-5.6-Cyber를 활용해 왔습니다. 오픈AI는 이 모델이 크롬의 자바스크립트 엔진인 V8을 분석하여 메모리를 손상시키고 V8 힙(Heap) 샌드박스를 우회할 수 있도록 연결할 수 있는, 이전에 알려지지 않은 두 개의 취약점을 발견했다고 밝혔습니다. 구글은 조율된 공개 과정을 거쳐 이 결함을 수정하고 CVE-2026-15903 식별 번호를 부여했습니다. 또한 GPT-5.6-Cyber는 '인기 있는 모바일 운영체제'에서도 최소 5개의 취약점을 발견한 것으로 알려졌습니다.

원문 보기
원문 보기 (영어)
OpenAI launches GPT-5.6-Cyber to help defenders find vulnerabilities before attackers do Matthias Bastian View the LinkedIn Profile of Matthias Bastian Aug 10, 2026 Nano Banana Pro prompted by THE DECODER Key Points OpenAI is expanding its Daybreak cybersecurity program with two new access tiers and a dedicated AI model called GPT-5.6-Cyber, designed to help security professionals identify vulnerabilities and develop exploits at an early stage. The program is split into two tracks: Daybreak Blue focuses on defensive tasks such as malware analysis, while Daybreak Red is geared toward offensive security research. Through the Red tier, users gain access to GPT-5.6-Cyber, a model specifically trained for offensive security purposes that responds to nearly all sensitive security queries typically blocked by other AI models. Ask about this article… Search OpenAI is expanding its Daybreak program with two new access tiers and a specialized model called GPT-5.6-Cyber. The model is designed to help defenders spot vulnerabilities and build exploits before attackers can deploy AI-powered offensive tools at scale. OpenAI says threat actors will increasingly use AI for cyberattacks, including fully autonomous ones. The window for defenders to prepare is getting smaller. Ironically, the best example for this came from OpenAI itself, when its own models accidentally hacked Hugging Face and other services after weeks of agentic scheming on internal message boards . Daybreak is meant to give defenders a head start. The program now has two access tiers . Daybreak Blue gives users access to GPT-5.6 Sol with tailored safeguards for authorized defense work like vulnerability detection, malware analysis, and incident response. Daybreak Red targets security researchers doing vulnerability research, exploit validation, and penetration testing. Ad Getting into either tier requires identity verification, account security measures, monitoring, and legal declarations. Hardware security keys become mandatory for all Daybreak accounts on September 1, 2026. OpenAI also recommends running security workflows in isolated sandbox environments and using Auto-Review mode in Codex, which checks actions that need elevated privileges before they run. Ad DEC_D_Incontent-1 GPT-5.6-Cyber answers 95 percent of sensitive security queries that other models block The new GPT-5.6-Cyber model is available through the Daybreak Red tier. It's based on GPT-5.6 Sol and was specifically trained to perform better on tasks like finding zero-day vulnerabilities and building exploit chains. According to OpenAI , the model rarely refuses security-related queries that other models block by default. In an internal benchmark called "Advanced Cybersecurity Completion Rate," GPT-5.6-Cyber answers 95 percent of queries covering scenarios like exploit chain development, authentication bypass, and privilege escalation. GPT-5.6 Sol with safety measures turned on hits just 1.5 percent. With Daybreak Blue, it reaches 2 percent. The previous model, GPT-5.5-Cyber, manages 57.3 percent. Ad In one specific test, the models had to develop a WebSocket authentication bypass for an internal admin panel. Only GPT-5.6-Cyber on Daybreak Red produced working exploit code. Every other variant refused to respond. On ExploitGym, a benchmark that measures how well models turn known vulnerabilities into working exploits, GPT-5.6-Cyber beats both GPT-5.6 Sol and GPT-5.5-Cyber. The model already found two unknown Chrome vulnerabilities OpenAI has also used GPT-5.6-Cyber for real-world vulnerability research. The company says the model analyzed V8, Chrome's JavaScript engine, and found two previously unknown vulnerabilities that can be chained together to corrupt memory and bypass the V8 heap sandbox. Google fixed the flaws after coordinated disclosure and assigned them the CVE-2026-15903 designation . Ad DEC_D_Incontent-2 GPT-5.6-Cyber also reportedly found at least five vulnerabilities in a "popular mobile operating system." One of them is a chain of flaws that would let an app escalate its normally restricted access rights to full administrator privileges, taking control of the device. OpenAI is working with Daybreak partners and the open-source community to disclose and fix these issues. Ad Under OpenAI's Preparedness Framework, GPT-5.6-Cyber has been rated "High" for cybersecurity capabilities but doesn't reach the "Critical" threshold. The recently announced Astra model is "potentially" expected to hit that Critical level, though. Given that GPT-5.6-Cyber is already a specialized, optimized model and still falls short of Critical, the trajectory is clear: AI cyber capabilities are climbing fast with each new generation. AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: OpenAI