메뉴
BL
The Decoder • 38일 전

AI 사이버보안 위험 커져… OpenAI, 모델 개발 속도 조절

IMP
8/10
핵심 요약

OpenAI가 다가오는 'Astra' 모델이 위험한 사이버공격 능력에 근접할 수 있다는 이유로 모델 개발 속도를 조절하고 있다. 강화학습을 2주간 중단했고 최대 규모 프런티어 RL 학습도 보류 중이며, 보안 요건을 충족하지 못한 작업도 일시 중단되었다. 연구 환경 강화, 30분 이내 경보를 보내는 새 모니터링 시스템 도입, 정렬(alignment) 연구 투자 확대 등의 조치가 이어지고 있다.

번역된 본문

OpenAI는 AI 사이버보안 위험이 너무 위험해지면서 모델 개발 속도를 조절하고 있다고 밝혔다. 부분적으로는 다가오는 'Astra' 모델이 핵심적인 사이버공격 능력을 확보할 수 있는 수준에 가까워지고 있기 때문이다.

이 회사는 2주간 강화학습(RL)을 중단했으며, "가장 큰 규모로 계획된 프런티어 RL 학습"은 여전히 보류 중이고, 새로운 보안 요건을 충족하지 못한 작업 부하들도 일시 중단되었다. Hugging Face 보안 사고와 "내부 연구의 빠른 진전"도 이번 속도 조절의 배경이 되었다.

이후 OpenAI는 더 나은 네트워크 격리와 더 엄격한 샌드박스를 통해 연구 환경을 강화했다고 밝혔다. 새로운 모니터링 시스템은 의심스러운 행동을 감지하면 30분 이내에 경보를 보내며, 작업 부하에 따라 감독용 추론 컴퓨팅의 약 20퍼센트를 사용한다.

회사는 준비 프레임워크(Preparedness Framework)를 확장하고 정렬(alignment) 연구에 더 많이 투자할 계획이다. 다만 해당 프레임워크를 만들었던 팀은 해체했으며, 그 책임을 다른 팀들로 이관했다.

비판론자들은 OpenAI가 시간과 관심을 벌기 위해 공포를 조장한다고 계속 비난할 가능성이 크다. 한편 독립 정부 기관인 AISI는 유사한 유해한 모델 행동을 문서화한 바 있어, OpenAI의 주장에 어느 정도 무게를 실어준다.

출처: OpenAI

원문 보기
원문 보기 (영어)
OpenAI says it's "pacing model development" as AI cybersecurity risks grow too dangerous Matthias Bastian View the LinkedIn Profile of Matthias Bastian Aug 18, 2026 OpenAI says it's "pacing model development," partly because the upcoming "Astra" model may be close to gaining critical cyberattack capabilities . The company paused reinforcement learning for two weeks, its "largest planned frontier RL run" remains on hold, and workloads that haven't met new security requirements are suspended. The Hugging Face security incident and "rapid progress in our internal research" also prompted the slowdown. Since then, OpenAI says research environments have been hardened with better network isolation and stricter sandboxes. A new monitoring system alerts within 30 minutes of detecting suspicious behavior, using roughly 20 percent of supervised inference compute depending on workload. The company plans to expand its Preparedness Framework and invest more in alignment research . It has, however, disbanded the team behind that framework , shifting responsibilities to other teams. Ad Critics will likely keep accusing OpenAI of fear-mongering to buy time and attention. The independent government agency AISI has documented similar harmful model behavior , lending some weight to OpenAI's claims. Ad AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: OpenAI Ask about this article… Search
관련 소식